... BigQuery, Cloud Functions, Composer, GCS Proficient hands-on programming experience in Spark/Scala (python/java) Proficient in building production level ETL/ELT data pipelines from data ingestion to consumption Data Engineering knowledge (such as Data Lake, Data warehouse - Redshift/Hive/Snowflake, Integration, Migration) ...
... https://jobeax.com/link/MJvWAZuK638tHuBb RESPONSIBILITIES : Tech Stack & Skills : - Experience in model development using Python/PySpark libraries. Development on Databricks or Dataiku DSS is a plus.- Strong experience on Spark with Scala/Python/Java.- Proficiency in building, training, and evaluating state-of-the-art machine ...
We are seeking a Data Engineer to join our team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL ...
... inquiries related to Addepar's portfolio data feeds and on general data product functionalities within established SLAs. Manage and complete requests from internal Data teams that require client outreach and/or action to resolve data verification issues. Investigate client reported bugs and data processing issues, and triage ...
Responsible for conducting data analysis to extract actionable insights, exploring datasets to uncover patterns and anomalies, analyzing historical data for trend identification and forecasting, investigating data discrepancies, providing user training and support on data analysis tools, communicating findings through compelling ...
... provide constructive feedback, and work closely with researchers and stakeholders to ensure alignment with project objectives. Develop and optimize Python-based data science solutions using public datasets. Create well-documented Python code using Jupyter notebooks. Bachelor's/Master's degree in Engineering, Computer Science, ...
... evaluating and implementing data-engineering and software technologies - In addition, you have experience in programming languages and frameworks: SQL, Python, Spark, Databricks (Delta Lake) - Experience in data storages - SQL and NoSQL databases, Azure Data Lake Storage- and in developing data solutions, models, API and software ...
... Posted 2 weeks ago Job Be an early applicant About the job Job Purpose: Design and implement scalable data engineering and data warehouse solutions while managing data schemas, SQL query tuning, and code reviews. Who You Are: - 5+ years of experience in Data Engineering, with strong knowledge of Data Platforms and Data Warehousing ...
... environments (preferably AWS). Document architecture, data flow, and support runbooks; continuously improve platform performance and resilience. Integrate with customer data platforms and pipelines, including bespoke data frameworks. 4–8 years of experience in data engineering or backend development in data-intensive environments. ...
... building data pipelines, and writing efficient, scalable Python code. Key Responsibilities Develop, test, and maintain scalable Python applications. Work with large datasets to extract, transform, and analyze data. Build and optimize data pipelines and workflows. Perform data cleaning, validation, and preprocessing. Collaborate ...
... validate data from multiple sources Perform exploratory data analysis (EDA) Develop dashboards and reports using BI tools Interpret trends and patterns in complex datasets Present insights to stakeholders in a clear and actionable manner Write SQL queries to extract and manipulate data Automate data processes where possible ...
... statements, and offering solutions in the domains of Retail, Pharma, Banking, Insurance, etc. - Contribute to internal product development initiatives related to data science. - Develop data science roadmap, and guide data scientist to meet their deliverables. - Handling end-to-end client AI & analytics programs. Your role ...
... platforms Design & deliver data warehousing solutions and key data engineering workstreams for any required solution. Support cross-functional teams across the data space. 3+ years of experience designing and building scalable distributed data pipelines and dimensional data models - 3+ years of experience in Python and SQL ...
... dashboards. - Expert-level proficiency in SQL , including complex queries, joins, aggregations, subqueries, and data manipulation. - Strong understanding of data preparation, data modelling, reporting, and visualisation principles. - Experience working with financial, operational, risk, or business performance data. - ...
... https://jobeax.com/link/MJvWAZuK638tHuBb RESPONSIBILITIES : Tech Stack & Skills : - Experience in model development using Python/PySpark libraries. Development on Databricks or Dataiku DSS is a plus.- Strong experience on Spark with Scala/Python/Java.- Proficiency in building, training, and evaluating state-of-the-art machine ...
... orchestrate datasets. Perform routine data monitoring, quality checks, and troubleshooting to maintain data integrity and workflow performance. Collaborate with data analysts, data scientists, and senior engineers to deliver data solutions. Document workflows, dataset schemas, and data engineering best practices. Required ...
... drives / memory / transceivers and other hardware related tasks, as required by our server and network teams. Assisting with planning equipment moves or new data center build-outs. Coordinating the packing, shipping, and logistics of equipment to and from remote colocation sites. Maintaining data center documentation. ...
... Ensure the reliability, scalability, and efficiency of data pipelines for ingestion, transformation, and storage. Work with cross-functional teams to understand data needs and deliver high-quality solutions. Troubleshoot and resolve data pipeline issues in production environments. Apply data quality and governance best practices, ...
... Own end-to-end delivery of data products—from raw source data ingestion through transformation to governed, consumption-ready datasets. Collaborate with the Data Platform team on pipeline integration, CI/CD workflows, and adherence to shared coding and deployment standards. - Data Quality & Analysis: Implement data quality ...