... data engineering, and DevOps automation. The ideal candidate will manage and optimize Cloudera environments, build CI/CD pipelines, and support enterprise-scale data processing workloads. Key Responsibilities Administer and support Cloudera CDP/CDH platforms, including HDFS, Hive, Spark, YARN, Hue, and CDE. Develop, deploy, ...
... efficient code using Scala and Spark - Work with Azure or on-premises Hadoop ecosystem for data processing - Solve real-world engineering problems related to data infrastructure - Collaborate with cross-functional teams to deliver data-driven solutions - Ensure data quality, reliability, and performance across data systems ...
... own hands. DIVERSITY & INCLUSION We believe that diversity of people and perspective gives us a competitive advantage. Job Title: GCP Data Engineer Position: Data Engineer Experience: 5+ Years Work Mode: Hybrid (Capco Office) Design and develop scalable data pipelines using GCP and BigQuery. Build data ingestion pipelines ...
... happens to in-flight work when a worker or a database node dies - Hands-on with a distributed processing engine ( Spark/PySpark or equivalent) on non-trivial data volumes - Experience with an orchestrator ( Dagster , Airflow, Prefect, or equivalent) and a cloud platform (AWS/Azure) - Data-quality mindset : you build validation ...
... deploying complex data architectures within Microsoft Fabric and the broader Azure ecosystem.- Advanced proficiency in Python and SQL for building sophisticated data processing pipelines and complex analytical models.- Strong background in data engineering principles, including ETL/ELT design, data modeling, and performance ...
Roles and Responsibilities Architect and maintain enterprise-grade ELT and ETL data pipelines using Python, PySpark, Kafka, and Databricks to manage large-scale risk data. Build and deploy GenAI agents utilizing Google ADK, Google Flash 2.5+ LLMs, and Model Context Protocol (MCP) integrated with Human-in-the-Loop workflows. ...
... Drive query optimization and data access strategies for APIs serving real-time compliance dashboards, establishing benchmarks and SLOs Lead data infrastructure work for AI/ML features, including dataset curation, feature engineering, and pipeline design supporting cloud-native AI capabilities Define and enforce data quality ...
... basics; exposure to Power BI or Python is a plus.- Curiosity and rigor; willingness to learn under structured https://jobeax.com/link/E56NWE2gL5VKO5x9 the Team Works :Work flows from business demand intake through discovery, solution design, data foundation, dashboard and AI build, testing and adoption, to delivery governance ...
... of truth. Develops and operationalizes data pipelines to bring data into Costco’s GCP landscape for the delivery of certified data sets Works in tandem with Data Architects, Data Stewards, and Data Quality Engineers to design data pipelines and recommends ongoing optimization of data storage, data ingestion, data quality ...
... background. Candidates who are already comfortable working in a remote or hybrid environment and are looking for a new project opportunity are encouraged to apply. The Data Engineer will design, build, and maintain scalable data pipelines and solutions, ensuring high performance and reliability across COSMOS DB and related data platforms. ...
... including model governance and explainability- Establish data governance framework (catalog, lineage, ownership)- Ensure data quality, consistency, and compliance- Work closely with business leaders to translate requirements into data solutions- Lead cross-functional teams (data engineers, analysts, data scientists)- Manage vendors, ...
... SDLC environments, with hands-on agile/scrum delivery experience. Exhibit high data literacy with the ability to process and analyze large datasets. Leverage data analysis/visualization tools (e.g., Alteryx, Python, Tableau) as applicable. Uphold strong data governance orientation (e.g., MDM) when working with critical ...
... hands-on experience in data science.- Strong proficiency in Python and libraries such as NumPy, pandas, scikit-learn, TensorFlow/PyTorch.- Expertise in SQL for data extraction, transformation, and analysis.- Proven experience with PySpark or distributed data processing frameworks.- Deep understanding of machine learning and ...
... We're the world's leading data, insights, and consulting company we shape the brands of tomorrow by better understanding people everywhere. We are looking for a Data Scientist who will work with a team of expert Data Analysts/Scientists/Engineers and a team of expert and dedicated developers. The key focus of this role is ...
... The position oversees modern cloud data platforms, with a particular focus on AWS and Snowflake environments. You will work closely with Product, Analytics, Data Governance, and other stakeholders to connect engineering initiatives with strategic business goals. The role offers significant influence over data architecture, ...
Big Data Processing: Design and manage scalable data pipelines to process massive datasets efficiently for model training and inference. Build, train, and fine-tune complex neural networks across text, audio, and visual modalities. Cloud Deployment: Architect and deploy models to cloud environments, leveraging distributed ...
... SLA responsibility in an ML/AI data operations environment - Knowledge of databases (SQL, MySQL) and Advanced Excel experience working with and analyzing large datasets to drive business improvements - Familiarity with ML data annotation workflows, labeling tools, or AI/ML program operations - Experience working in geographically ...
... Cleansing and Validation: Perform data cleansing and validation tasks to maintain high data quality. Identify and correct data anomalies and inconsistencies in datasets. Data Quality Metrics: Establish and monitor data quality metrics to track the performance and accuracy of data. Report on data quality issues and work with ...
... adopt AI tools and techniques to improve personal and team productivity - including AI-assisted coding, automated data validation, and GenAI-driven analysis workflows- Contribute to the development and testing of AI-powered product features, such as natural language interfaces, intelligent data exploration, or automated ...
... Product Engineering Role Overview We are looking for an experienced Data Engineer with 5–7 years of hands-on experience in ETL/ELT, database engineering, Big Data processing, and complex bi-directional data integrations . The candidate should have strong SQL and database expertise, experience working with large datasets, ...