... BigQuery, Cloud Functions, Composer, GCS Proficient hands-on programming experience in Spark/Scala (python/java) Proficient in building production level ETL/ELT data pipelines from data ingestion to consumption Data Engineering knowledge (such as Data Lake, Data warehouse - Redshift/Hive/Snowflake, Integration, Migration) ...
We are seeking a Data Engineer to join our team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL ...
... inquiries related to Addepar's portfolio data feeds and on general data product functionalities within established SLAs. Manage and complete requests from internal Data teams that require client outreach and/or action to resolve data verification issues. Investigate client reported bugs and data processing issues, and triage ...
Responsible for conducting data analysis to extract actionable insights, exploring datasets to uncover patterns and anomalies, analyzing historical data for trend identification and forecasting, investigating data discrepancies, providing user training and support on data analysis tools, communicating findings through compelling ...
... evaluating and implementing data-engineering and software technologies - In addition, you have experience in programming languages and frameworks: SQL, Python, Spark, Databricks (Delta Lake) - Experience in data storages - SQL and NoSQL databases, Azure Data Lake Storage- and in developing data solutions, models, API and software ...
... Posted 2 weeks ago Job Be an early applicant About the job Job Purpose: Design and implement scalable data engineering and data warehouse solutions while managing data schemas, SQL query tuning, and code reviews. Who You Are: - 5+ years of experience in Data Engineering, with strong knowledge of Data Platforms and Data Warehousing ...
... environments (preferably AWS). Document architecture, data flow, and support runbooks; continuously improve platform performance and resilience. Integrate with customer data platforms and pipelines, including bespoke data frameworks. 4–8 years of experience in data engineering or backend development in data-intensive environments. ...
... building data pipelines, and writing efficient, scalable Python code. Key Responsibilities Develop, test, and maintain scalable Python applications. Work with large datasets to extract, transform, and analyze data. Build and optimize data pipelines and workflows. Perform data cleaning, validation, and preprocessing. Collaborate ...
... validate data from multiple sources Perform exploratory data analysis (EDA) Develop dashboards and reports using BI tools Interpret trends and patterns in complex datasets Present insights to stakeholders in a clear and actionable manner Write SQL queries to extract and manipulate data Automate data processes where possible ...
... platforms Design & deliver data warehousing solutions and key data engineering workstreams for any required solution. Support cross-functional teams across the data space. 3+ years of experience designing and building scalable distributed data pipelines and dimensional data models - 3+ years of experience in Python and SQL ...
... dashboards. - Expert-level proficiency in SQL , including complex queries, joins, aggregations, subqueries, and data manipulation. - Strong understanding of data preparation, data modelling, reporting, and visualisation principles. - Experience working with financial, operational, risk, or business performance data. - ...
... orchestrate datasets. Perform routine data monitoring, quality checks, and troubleshooting to maintain data integrity and workflow performance. Collaborate with data analysts, data scientists, and senior engineers to deliver data solutions. Document workflows, dataset schemas, and data engineering best practices. Required ...
... drives / memory / transceivers and other hardware related tasks, as required by our server and network teams. Assisting with planning equipment moves or new data center build-outs. Coordinating the packing, shipping, and logistics of equipment to and from remote colocation sites. Maintaining data center documentation. ...
... Ensure the reliability, scalability, and efficiency of data pipelines for ingestion, transformation, and storage. Work with cross-functional teams to understand data needs and deliver high-quality solutions. Troubleshoot and resolve data pipeline issues in production environments. Apply data quality and governance best practices, ...
... Own end-to-end delivery of data products—from raw source data ingestion through transformation to governed, consumption-ready datasets. Collaborate with the Data Platform team on pipeline integration, CI/CD workflows, and adherence to shared coding and deployment standards. - Data Quality & Analysis: Implement data quality ...
... of our transformation, we know that data is at the center of everything that we do. We seek a dedicated professional to help take us to the next level of our Data & Analytics journey. We depend on our Data Analysts to investigate data seeking opportunities to reduce duplicates, improve the quality and discover anomalies ...
... MAKE AN IMPACT Job Title: Data Engineer _DEPS Experience: 3–5 Years Location: Bangalore and Pune (Hybrid – Client Office) Job Summary: Seeking a skilled Senior Data DevOps Engineer having experience in Cloudera platforms, data engineering, and DevOps automation. The ideal candidate will manage and optimize Cloudera environments, ...
... efficient code using Scala and Spark - Work with Azure or on-premises Hadoop ecosystem for data processing - Solve real-world engineering problems related to data infrastructure - Collaborate with cross-functional teams to deliver data-driven solutions - Ensure data quality, reliability, and performance across data systems ...
... Python. Develop and optimize large-scale data transformations using PySpark. Ensure data quality, reliability, and performance of data pipelines. GCP BigQuery Data Ingestion & ETL/ELT Data Pipeline Orchestration Python PySpark Strong SQL and data engineering fundamentals hands-on experience in GCP-based data platforms.