... closely with cross-functional product teams, data scientists, and senior stakeholders to translate intricate business requirements into scalable, high-performance data pipelines. By architecting efficient storage and processing solutions, you will directly influence the speed and accuracy of decision-making across the organization, ...
... controls Identify gaps, defects, and technical debt across the platform and remediate where implementations are incorrect or sub-optimal Ensure correctness of data processing patterns including change data capture, slowly changing dimensions, deduplication, and business reconciliation Design and implement target-state architectures ...
Data Science Data Engineer Work Type: Full Time At VIDA we're building the future of digital identity. As a Data Engineer you'll work across the stack—from platform and infrastructure to data pipelines and end-user tooling. You'll help us modernize our batch ETLs and lead our push into streaming and real-time fraud detection. ...
... optimized SQL queries for data extraction, transformation, and analysis.- Build and manage ETL pipelines using AWS Glue.- Work with AWS Glue Data Catalog for metadata management and data discovery.- Design and implement serverless data processing solutions using AWS Lambda.- Use Amazon S3 for scalable data storage and data ...
... techniques. Hands-on experience integrating and optimizing Starburst/Trino, including connecting Starburst from Lambda and Glue ETL jobs. Experience with NoSQL databases such as DynamoDB, MongoDB. Experience working with data formats including Avro, Parquet, JSON, XML, and CSV. Comfortable challenging your peers and leadership ...
... details related to data engineering or a relevant https://jobeax.com/link/cwwOL3BLJkve6nkU Skills :- Azure Databricks, Python, pyspark, sql, ADF- Proficiency in data pipeline design and implementation- Strong understanding of data factory and orchestration toolsPreferred Skills:- Familiarity with advanced data processing techniques- ...
... (mandatory) for data engineering use cases PySpark / Sparkbased processing Building reusable ETL components, utilities, and data pipelines Strong understanding of data modeling, transformations, and performance optimization Data Processing & Engineering Proven handson experience with distributed processing frameworks such as ...
... using - Azure Data Factory, Azure Pipelines, Azure DevOps, and Git for version control. - Proven experience working with Azure Databricks and PySpark/Spark for data engineering tasks and Power BI as data analytics work - In-depth knowledge of performance optimization techniques for large-scale data processing, including code ...
... in Python and SQL - Experience with Unity Catalog, cluster administration, and at least one major cloud platform (Azure, AWS, or GCP) - Solid understanding of data warehousing and dimensional modeling Databricks Certified Data Engineer (Associate or Professional) Cloud data engineering certification have minimum 5 years ...
... magic! Model Development: Design, develop, and implement predictive models and algorithms using advanced statistical techniques and machine learning frameworks. Data Analysis: Conduct thorough data exploration, cleaning, and transformation to prepare datasets for modeling efforts, ensuring high data quality and accuracy. Collaboration: ...
... capture the most granular data. Align the engineering and product teams on the Data Granularity and accuracy requirements. Build Alerts to detect bugs in the data pipeline. Perform data analysis to test hypothesis and uncover factor influencing metric variations Build Feature Store for Data Science implementations Overall ...
... identify opportunities for improvement. Strong analytic skills related to working with unstructured data sets. Build processes supporting data transformation, data structures, metadata, dependency and workload management. Working knowledge of message queuing, stream processing, and highly scalable 'big data' data stores. ...
Role OverviewWe are seeking a high-caliber Data Migration Consultant Developer with proven expertise in Guidewire Data Migration methodology and tools . The ideal candidate must hold a Mammoth Certification (Data Migration Track) and bring extensive, hands-on experience in migrating client data from legacy Policy systems ...
... BigQuery, Cloud Functions, Composer, GCS Proficient hands-on programming experience in Spark/Scala (python/java) Proficient in building production level ETL/ELT data pipelines from data ingestion to consumption Data Engineering knowledge (such as Data Lake, Data warehouse - Redshift/Hive/Snowflake, Integration, Migration) ...
... https://jobeax.com/link/MJvWAZuK638tHuBb RESPONSIBILITIES : Tech Stack & Skills : - Experience in model development using Python/PySpark libraries. Development on Databricks or Dataiku DSS is a plus.- Strong experience on Spark with Scala/Python/Java.- Proficiency in building, training, and evaluating state-of-the-art machine ...
We are seeking a Data Engineer to join our team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL ...
... inquiries related to Addepar's portfolio data feeds and on general data product functionalities within established SLAs. Manage and complete requests from internal Data teams that require client outreach and/or action to resolve data verification issues. Investigate client reported bugs and data processing issues, and triage ...
Responsible for conducting data analysis to extract actionable insights, exploring datasets to uncover patterns and anomalies, analyzing historical data for trend identification and forecasting, investigating data discrepancies, providing user training and support on data analysis tools, communicating findings through compelling ...
... provide constructive feedback, and work closely with researchers and stakeholders to ensure alignment with project objectives. Develop and optimize Python-based data science solutions using public datasets. Create well-documented Python code using Jupyter notebooks. Bachelor's/Master's degree in Engineering, Computer Science, ...
... enterprise-wide data Build and manage robust data pipelines integrating sources like ERP systems, IoT devices, iHistorian, web services, and external APIs Leverage modern data engineering stacks (Python, SQL, Spark, Delta Lake, Databricks, Azure) to enable large-scale, high-performance processing Collaborate with stakeholders and IT ...