... https://jobeax.com/link/MJvWAZuK638tHuBb RESPONSIBILITIES : Tech Stack & Skills : - Experience in model development using Python/PySpark libraries. Development on Databricks or Dataiku DSS is a plus.- Strong experience on Spark with Scala/Python/Java.- Proficiency in building, training, and evaluating state-of-the-art machine ...
We are seeking a Data Engineer to join our team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL ...
Responsible for conducting data analysis to extract actionable insights, exploring datasets to uncover patterns and anomalies, analyzing historical data for trend identification and forecasting, investigating data discrepancies, providing user training and support on data analysis tools, communicating findings through compelling ...
... provide constructive feedback, and work closely with researchers and stakeholders to ensure alignment with project objectives. Develop and optimize Python-based data science solutions using public datasets. Create well-documented Python code using Jupyter notebooks. Bachelor's/Master's degree in Engineering, Computer Science, ...
... software applications using SQL, Python and .NET - Working knowledge of Azure infrastructure management and resource deployment, system networking and security and data visualization tools like PowerBI, Grafana, and Tableau - Preferred qualifications are a Master's or PhD degree in Computer Science or related Engineering areas ...
... building data pipelines, and writing efficient, scalable Python code. Key Responsibilities Develop, test, and maintain scalable Python applications. Work with large datasets to extract, transform, and analyze data. Build and optimize data pipelines and workflows. Perform data cleaning, validation, and preprocessing. Collaborate ...
... statements, and offering solutions in the domains of Retail, Pharma, Banking, Insurance, etc. - Contribute to internal product development initiatives related to data science. - Develop data science roadmap, and guide data scientist to meet their deliverables. - Handling end-to-end client AI & analytics programs. Your role ...
... dashboards. - Expert-level proficiency in SQL , including complex queries, joins, aggregations, subqueries, and data manipulation. - Strong understanding of data preparation, data modelling, reporting, and visualisation principles. - Experience working with financial, operational, risk, or business performance data. - ...
... https://jobeax.com/link/MJvWAZuK638tHuBb RESPONSIBILITIES : Tech Stack & Skills : - Experience in model development using Python/PySpark libraries. Development on Databricks or Dataiku DSS is a plus.- Strong experience on Spark with Scala/Python/Java.- Proficiency in building, training, and evaluating state-of-the-art machine ...
... orchestrate datasets. Perform routine data monitoring, quality checks, and troubleshooting to maintain data integrity and workflow performance. Collaborate with data analysts, data scientists, and senior engineers to deliver data solutions. Document workflows, dataset schemas, and data engineering best practices. Required ...
... Ensure the reliability, scalability, and efficiency of data pipelines for ingestion, transformation, and storage. Work with cross-functional teams to understand data needs and deliver high-quality solutions. Troubleshoot and resolve data pipeline issues in production environments. Apply data quality and governance best practices, ...
... MAKE AN IMPACT Job Title: Data Engineer _DEPS Experience: 3–5 Years Location: Bangalore and Pune (Hybrid – Client Office) Job Summary: Seeking a skilled Senior Data DevOps Engineer having experience in Cloudera platforms, data engineering, and DevOps automation. The ideal candidate will manage and optimize Cloudera environments, ...
... efficient code using Scala and Spark - Work with Azure or on-premises Hadoop ecosystem for data processing - Solve real-world engineering problems related to data infrastructure - Collaborate with cross-functional teams to deliver data-driven solutions - Ensure data quality, reliability, and performance across data systems ...
... Python. Develop and optimize large-scale data transformations using PySpark. Ensure data quality, reliability, and performance of data pipelines. GCP BigQuery Data Ingestion & ETL/ELT Data Pipeline Orchestration Python PySpark Strong SQL and data engineering fundamentals hands-on experience in GCP-based data platforms.
... deploying complex data architectures within Microsoft Fabric and the broader Azure ecosystem.- Advanced proficiency in Python and SQL for building sophisticated data processing pipelines and complex analytical models.- Strong background in data engineering principles, including ETL/ELT design, data modeling, and performance ...
... Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering experience in Banking Risk, Retail Products, Cards, Mortgage, Deposits, or Wealth Management. Assumed Requirements / Certifications: Databricks Certified Data Engineer, AWS Certified Data Analytics, or Azure Data Engineer Associate.
... of truth. Develops and operationalizes data pipelines to bring data into Costco’s GCP landscape for the delivery of certified data sets Works in tandem with Data Architects, Data Stewards, and Data Quality Engineers to design data pipelines and recommends ongoing optimization of data storage, data ingestion, data quality ...
... activities and site communications. Collaborate with the project team to address issues, resolve discrepancies, and ensure trial milestones are met. Support data management activities by ensuring timely and accurate data collection and entry. Participate in audits and inspections as required, providing necessary documentation ...
... Data Engineer will design, build, and maintain scalable data pipelines and solutions, ensuring high performance and reliability across COSMOS DB and related data platforms. Daily responsibilities include modeling and structuring data, implementing ETL processes, optimizing data storage and retrieval, and collaborating ...