... ability to design, build, and operate Data Lakes or Data Warehouses. Proficiency with Data Orchestration tools (Airflow, Dagster, Prefect). Familiarity with Change Data Capture tools (Canal, Debezium, Maxwell). Strong command of at least one primary language (Scala, Python, etc.) and SQL. Experience with data catalog and metadata ...
... modelling depth — you can design a canonical model across messy sources and defend the decisions behind it. - Production pipeline engineering — Spark or equivalent, Python, strong SQL, and orchestration with Airflow or similar. - Data quality and governance in practice — validation frameworks, lineage and reconciliation. Not as ...
... others. Required qualifications and experience: - Bachelor's degree in a relevant field, such as Life Sciences or Healthcare. - Proven experience in clinical data management within the pharmaceutical or biotechnology industry. - Familiarity with data management software and systems (e.g., Medidata, Oracle RDC, or similar). ...
... and development Data quality validation and monitoring Schema design for analytics and reporting Performance optimization and scalability Preferred Experience Databricks Delta Lake experience Azure Synapse Analytics Python for data engineering Spark SQL optimization Real-time data streaming Data governance and metadata management ...
... markets, while owning the full lifecycle of a signal, forming a hypothesis about what market behavior predicts, engineering it into a feature, validating it with data, and shipping it into a production research platform. This is a research-focused data science role with meaningful hands-on coding and implementing signals within ...
... collaborative mindset with the ability to thrive in a hybrid work environment across Bangalore, Chennai, Pune, or Mumbai.- A Bachelors or Masters degree in Computer Science, Information Systems, or a related quantitative field is preferred.- Candidates must possess 8 to 12 years of relevant professional experience in data engineering ...
Data Science Data Engineer Work Type: Full Time At VIDA we're building the future of digital identity. As a Data Engineer you'll work across the stack—from platform and infrastructure to data pipelines and end-user tooling. You'll help us modernize our batch ETLs and lead our push into streaming and real-time fraud detection. ...
... data pipelines and ETL workflows on AWS.- Develop efficient data processing solutions using python and pySpark.- Write complex and optimized SQL queries for data extraction, transformation, and analysis.- Build and manage ETL pipelines using AWS Glue.- Work with AWS Glue Data Catalog for metadata management and data discovery.- ...
... techniques. Hands-on experience integrating and optimizing Starburst/Trino, including connecting Starburst from Lambda and Glue ETL jobs. Experience with NoSQL databases such as DynamoDB, MongoDB. Experience working with data formats including Avro, Parquet, JSON, XML, and CSV. Comfortable challenging your peers and leadership ...
... details related to data engineering or a relevant https://jobeax.com/link/cwwOL3BLJkve6nkU Skills :- Azure Databricks, Python, pyspark, sql, ADF- Proficiency in data pipeline design and implementation- Strong understanding of data factory and orchestration toolsPreferred Skills:- Familiarity with advanced data processing techniques- ...
... S3, AWS Glue, Athena, Redshift Experience designing cloudnative data lakes and data warehouse architectures on AWS Deep understanding of batch and streaming data pipelines Experience building scalable, faulttolerant data ingestion and transformation workflows SQL & Python (Mandatory) Strong SQL expertise Writing complex ...
... mapping, and change detection. - Ensure annotation accuracy, consistency, and adherence to defined quality standards (QA/QC frameworks) across all delivered datasets. - Define and continuously improve annotation guidelines, taxonomies, and labeling protocols in collaboration with Data Science and ML teams. - Stay current ...
... Data Migration. Proven expertise with Guidewire Data Migration Framework, Mammoth ETL, and related Guidewire tools. Strong knowledge of Guidewire PolicyCenter data model and its integration with external legacy systems. Proficiency in SQL, PL/SQL, and data transformation scripting (Python, Java, or Groovy preferred). Hands-on ...
We are seeking a Data Engineer to join our team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL ...
... inquiries related to Addepar's portfolio data feeds and on general data product functionalities within established SLAs. Manage and complete requests from internal Data teams that require client outreach and/or action to resolve data verification issues. Investigate client reported bugs and data processing issues, and triage ...
... Data Engineering, with strong knowledge of Data Platforms and Data Warehousing - 2+ years of hands-on experience with SSAS Tabular Models - Experience designing data ingestion and orchestration pipelines using Kafka, Snowflake, and Python - Hands-on experience with DBT for data modeling and pipeline development - Strong knowledge ...
... improve platform performance and resilience. Integrate with customer data platforms and pipelines, including bespoke data frameworks. 4–8 years of experience in data engineering or backend development in data-intensive environments. Proficient in Python and SQL; Strong experience with cloud-native data tools and services (S3, ...