... Responsibilities Develop, test, and maintain scalable Python applications. Work with large datasets to extract, transform, and analyze data. Build and optimize data pipelines and workflows. Perform data cleaning, validation, and preprocessing. Collaborate with cross-functional teams including data analysts and engineers. ...
... statements, and offering solutions in the domains of Retail, Pharma, Banking, Insurance, etc. - Contribute to internal product development initiatives related to data science. - Develop data science roadmap, and guide data scientist to meet their deliverables. - Handling end-to-end client AI & analytics programs. Your role ...
... platforms Design & deliver data warehousing solutions and key data engineering workstreams for any required solution. Support cross-functional teams across the data space. 3+ years of experience designing and building scalable distributed data pipelines and dimensional data models - 3+ years of experience in Python and SQL ...
... https://jobeax.com/link/MJvWAZuK638tHuBb RESPONSIBILITIES : Tech Stack & Skills : - Experience in model development using Python/PySpark libraries. Development on Databricks or Dataiku DSS is a plus.- Strong experience on Spark with Scala/Python/Java.- Proficiency in building, training, and evaluating state-of-the-art machine ...
... orchestrate datasets. Perform routine data monitoring, quality checks, and troubleshooting to maintain data integrity and workflow performance. Collaborate with data analysts, data scientists, and senior engineers to deliver data solutions. Document workflows, dataset schemas, and data engineering best practices. Required ...
We're looking for a Data Center Engineer to join our Infrastructure Engineering team in Mumbai. In this role, you'll help build, maintain, and scale the critical infrastructure that powers our trading operations across APAC. Working alongside experienced engineers and traders, you'll solve complex technical challenges in ...
... Ensure the reliability, scalability, and efficiency of data pipelines for ingestion, transformation, and storage. Work with cross-functional teams to understand data needs and deliver high-quality solutions. Troubleshoot and resolve data pipeline issues in production environments. Apply data quality and governance best practices, ...
... Requirement Gathering: Partner with business stakeholders across Analytics, Operations, GTM, and G&A to understand problem statements and translate them into clear data requirements.. - Data Modeling & Transformation: Design and build dimensional and analytical data models in Snowflake using dbt (Cloud/Core) that serve as the ...
... MAKE AN IMPACT Job Title: Data Engineer _DEPS Experience: 3–5 Years Location: Bangalore and Pune (Hybrid – Client Office) Job Summary: Seeking a skilled Senior Data DevOps Engineer having experience in Cloudera platforms, data engineering, and DevOps automation. The ideal candidate will manage and optimize Cloudera environments, ...
... efficient code using Scala and Spark - Work with Azure or on-premises Hadoop ecosystem for data processing - Solve real-world engineering problems related to data infrastructure - Collaborate with cross-functional teams to deliver data-driven solutions - Ensure data quality, reliability, and performance across data systems ...
... Python. Develop and optimize large-scale data transformations using PySpark. Ensure data quality, reliability, and performance of data pipelines. GCP BigQuery Data Ingestion & ETL/ELT Data Pipeline Orchestration Python PySpark Strong SQL and data engineering fundamentals hands-on experience in GCP-based data platforms.
... powering self-serve analytics. Write performant Spark/PySpark and SQL; optimize partitioning, storage formats, and query cost. Data Quality & Reliability (10%) Own data quality: validation, freshness/SLA monitoring, and observability so bad data is caught before it reaches consumers. Make the data layer debuggable: lineage, tests, ...
... deploying complex data architectures within Microsoft Fabric and the broader Azure ecosystem.- Advanced proficiency in Python and SQL for building sophisticated data processing pipelines and complex analytical models.- Strong background in data engineering principles, including ETL/ELT design, data modeling, and performance ...
... Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering experience in Banking Risk, Retail Products, Cards, Mortgage, Deposits, or Wealth Management. Assumed Requirements / Certifications: Databricks Certified Data Engineer, AWS Certified Data Analytics, or Azure Data Engineer Associate.
... workflows, dose analysis, and program management for radiation safety professionals. We are looking for a seasoned Data Engineer to take a lead role in shaping the data foundation across Landauer's growing portfolio of products. You will set the standards for how data is modeled, moved, and governed — owning the full data lifecycle ...
... of truth. Develops and operationalizes data pipelines to bring data into Costco’s GCP landscape for the delivery of certified data sets Works in tandem with Data Architects, Data Stewards, and Data Quality Engineers to design data pipelines and recommends ongoing optimization of data storage, data ingestion, data quality ...
... Data Engineer will design, build, and maintain scalable data pipelines and solutions, ensuring high performance and reliability across COSMOS DB and related data platforms. Daily responsibilities include modeling and structuring data, implementing ETL processes, optimizing data storage and retrieval, and collaborating ...
... framework (catalog, lineage, ownership)- Ensure data quality, consistency, and compliance- Work closely with business leaders to translate requirements into data solutions- Lead cross-functional teams (data engineers, analysts, data scientists)- Manage vendors, partners, and external technology providers- Drive roadmap ...
... solutions in Client Data Product Management . As a Data Product Manager within our dynamic team , you will be responsible for scaling the CDM(client data management) data product, strengthening operating models, and driving operational excellence through data-driven execution. Drive product development for the CDM data product ...
... hands-on experience in data science.- Strong proficiency in Python and libraries such as NumPy, pandas, scikit-learn, TensorFlow/PyTorch.- Expertise in SQL for data extraction, transformation, and analysis.- Proven experience with PySpark or distributed data processing frameworks.- Deep understanding of machine learning and ...