... Machine Learning, Scikit-learn, XgBoost, SQL GeoIQ (a Lenskart subsidiary) is India's leading hyperlocal location AI platform, helping businesses make precise, data-driven decisions using street-level intelligence across demographics, income, infrastructure, and commercial activity. As a Data Scientist, you will be a core ...
... change, applying first-principles problem solving, and using data to learn and adapt along the way. Enterprise Data & AI Mission Enterprise Data & AI provides the data and systems to enable our enterprise transformation from human-powered operations to data-driven, AI-powered, and human-led agentic operations. This data and ...
Data Scientist – Mortgage Analytics & AI Experience – 5 to 10 years About the Role We are looking for an experienced Data Scientist to join our team and help build the next generation of data-driven solutions for the mortgage and real estate industry. In this role, you will work at the intersection of data science, machine ...
Job Role: Senior Platform Data Engineer (Databricks) Architect and operate large-scale distributed data platforms using Databricks and Spark. Architect Databricks deployments Drive performance standards 8+ years of data or cloud engineering experience - Production ownership experience - Responsible for the end-to-end administration, ...
... seasoned Senior Data Scientist with e-commerce expertise to join our dynamic team. Must-Have Skills: Machine Learning Algorithms, Python or R Programming, Big Data Technologies (Hadoop, Spark, or equivalent), Data Visualization Tools (Tableau, Power BI, or similar), SQL and Data Modeling Expertise. Data Analysis & Insights ...
... compelling and clear business focused insights. Can work with multiple databases and big data solutions. Work with the data engineering team to develop & maintain data pipelines. Identify, design, and implement internal process improvements and tools to automate data processing and ensure data integrity while meeting data security ...
... behavior. - Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery. - Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against ...
... behavior. - Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery. - Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against ...
... Intelligence Platform team, you will: Build out performant APIs to serve consistent data that serve as sources of truth for many systems in Abnormal. Establish data pipelines that ensure our data is updated regularly and reliably Enhance our frameworks to allow us to make changes to our APIs and datasets in an agile manner ...
... and real time data pipelines Build and maintain robust ETL ELT processes for ingesting transforming and loading data from multiple sources Develop and manage data lakes data warehouses and metadata frameworks Collaborate with cross functional teams to understand data requirements and deliver data solutions Ensure data quality ...
... stakeholders. Data Pipeline Design Operations: Design, develop, and support data pipelines to extract, transform, and load (ETL/ELT) data into the new harmonized data platform. Utilize Azure Data Factory (ADF) for orchestration (scheduling, workflow management) and Azure Databricks (Spark) for large-scale data transformation ...
... behavior. - Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery. - Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against ...
... Science & Supply Chain talent. Collect business requirements from the various stakeholders (Data Scientists, Product Managers, IM) High interest in developing E2E data solutions DevOps/ Agile WoW Experience in designing and building advanced reports Programming languages: Python, PySpark, SQL Data enthusiast Big data frameworks: ...
... quality. Big Data - Big Data - Pyspark Database - Database Programming - SQL Data Science and Machine Learning - Data Science and Machine Learning - Apache Spark Data & AI - Data Engineering - Data Quality & Validation Data & AI - Data Engineering - Apache Kafka Big Data - Big Data - Azure databricks Data Science and Machine ...
... reliable, high-quality data, driving excellence in analytics and shaping outcomes that impact industries like H ealthcare and Financial Services . Design and build Data Pipelines using GCP services and BigQuery. Develop and optimize data solutions using Java or Python. Ensure data quality, governance, and compliance across projects. ...
... the development of system enhancements. - Work closely with Data Engineering, Product, Accounting, and Tax team members to investigate root causes, validate data fixes, improve reporting logic, and strengthen tax data controls. - Support indirect tax filing preparation by organizing filing data, validating reports, reviewing ...
... (Pandas, Polars) and modern big data tools such as Databricks (Spark), Flink, or Kafka. - Strong understanding of data pipeline development, ETL frameworks, and data lakes/warehouses. - Extensive experience with cloud data platforms such as AWS (S3, Redshift, Lambda, Glue), GCP (BigQuery, Dataflow, Dataproc), or Snowflake. ...
... processes to seamlessly integrate data from multiple sources into our data warehouse. Implement advanced data modeling techniques and design efficient, optimized databases to support complex business needs. Ensure the highest levels of data quality and integrity through rigorous data validation and cleansing processes. Collaborate ...
... behavior. - Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery. - Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against ...
... Integration: Build and manage data pipelines to ingest data from various sources, including third-party and RESTful APIs. Build ETL/ELT processes using ADF, Fabric Data Flows , and Notebooks. Embed MDM solutions into pipelines, ensuring seamless data sync across systems. Data Storage and Management: Design and implement data ...