... Ensure the reliability, scalability, and efficiency of data pipelines for ingestion, transformation, and storage. Work with cross-functional teams to understand data needs and deliver high-quality solutions. Troubleshoot and resolve data pipeline issues in production environments. Apply data quality and governance best practices, ...
... Own end-to-end delivery of data products—from raw source data ingestion through transformation to governed, consumption-ready datasets. Collaborate with the Data Platform team on pipeline integration, CI/CD workflows, and adherence to shared coding and deployment standards. - Data Quality & Analysis: Implement data quality ...
... MAKE AN IMPACT Job Title: Data Engineer _DEPS Experience: 3–5 Years Location: Bangalore and Pune (Hybrid – Client Office) Job Summary: Seeking a skilled Senior Data DevOps Engineer having experience in Cloudera platforms, data engineering, and DevOps automation. The ideal candidate will manage and optimize Cloudera environments, ...
... efficient code using Scala and Spark - Work with Azure or on-premises Hadoop ecosystem for data processing - Solve real-world engineering problems related to data infrastructure - Collaborate with cross-functional teams to deliver data-driven solutions - Ensure data quality, reliability, and performance across data systems ...
... Python. Develop and optimize large-scale data transformations using PySpark. Ensure data quality, reliability, and performance of data pipelines. GCP BigQuery Data Ingestion & ETL/ELT Data Pipeline Orchestration Python PySpark Strong SQL and data engineering fundamentals hands-on experience in GCP-based data platforms.
... powering self-serve analytics. Write performant Spark/PySpark and SQL; optimize partitioning, storage formats, and query cost. Data Quality & Reliability (10%) Own data quality: validation, freshness/SLA monitoring, and observability so bad data is caught before it reaches consumers. Make the data layer debuggable: lineage, tests, ...
... deploying complex data architectures within Microsoft Fabric and the broader Azure ecosystem.- Advanced proficiency in Python and SQL for building sophisticated data processing pipelines and complex analytical models.- Strong background in data engineering principles, including ETL/ELT design, data modeling, and performance ...
... 5+ years in a Python and PySpark Data Engineering lead role. Educational Background: Bachelor's degree in Computer Science, Engineering, or a related field (Master's degree preferred). Core Technical Skills: Python, PySpark, Databricks, Google ADK, LLMs, FastAPI, Spring Boot, Microservices, Kafka, SQL, Data Mesh, Starburst. ...
... workflows, dose analysis, and program management for radiation safety professionals. We are looking for a seasoned Data Engineer to take a lead role in shaping the data foundation across Landauer's growing portfolio of products. You will set the standards for how data is modeled, moved, and governed — owning the full data lifecycle ...
... of truth. Develops and operationalizes data pipelines to bring data into Costco’s GCP landscape for the delivery of certified data sets Works in tandem with Data Architects, Data Stewards, and Data Quality Engineers to design data pipelines and recommends ongoing optimization of data storage, data ingestion, data quality ...
... Data Engineer will design, build, and maintain scalable data pipelines and solutions, ensuring high performance and reliability across COSMOS DB and related data platforms. Daily responsibilities include modeling and structuring data, implementing ETL processes, optimizing data storage and retrieval, and collaborating ...
... framework (catalog, lineage, ownership)- Ensure data quality, consistency, and compliance- Work closely with business leaders to translate requirements into data solutions- Lead cross-functional teams (data engineers, analysts, data scientists)- Manage vendors, partners, and external technology providers- Drive roadmap ...
... solutions in Client Data Product Management . As a Data Product Manager within our dynamic team , you will be responsible for scaling the CDM(client data management) data product, strengthening operating models, and driving operational excellence through data-driven execution. Drive product development for the CDM data product ...
... hands-on experience in data science.- Strong proficiency in Python and libraries such as NumPy, pandas, scikit-learn, TensorFlow/PyTorch.- Expertise in SQL for data extraction, transformation, and analysis.- Proven experience with PySpark or distributed data processing frameworks.- Deep understanding of machine learning and ...
... We're the world's leading data, insights, and consulting company we shape the brands of tomorrow by better understanding people everywhere. We are looking for a Data Scientist who will work with a team of expert Data Analysts/Scientists/Engineers and a team of expert and dedicated developers. The key focus of this role is ...
... factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based on project guidelines; - Identify and flag factually incorrect, sensitive, inappropriate, or unclear ...
... factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based on project guidelines; - Identify and flag factually incorrect, sensitive, inappropriate, or unclear ...
... factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based on project guidelines; - Identify and flag factually incorrect, sensitive, inappropriate, or unclear ...
... factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based on project guidelines; - Identify and flag factually incorrect, sensitive, inappropriate, or unclear ...
... orchestration, OneLake concepts and Power BI semantic models, including Direct Lake where appropriate. - Experience with data warehouses, data lakes and Azure data services such as Azure Data Lake Storage and Azure Data Factory. - Strong knowledge of dimensional modelling, analytical structures and data-modelling trade-offs. ...