... experience evaluating and implementing data-engineering and software technologies - In addition, you have experience in programming languages and frameworks: SQL, Python, Spark, Databricks (Delta Lake) - Experience in data storages - SQL and NoSQL databases, Azure Data Lake Storage- and in developing data solutions, models, API ...
... Data Engineering, with strong knowledge of Data Platforms and Data Warehousing - 2+ years of hands-on experience with SSAS Tabular Models - Experience designing data ingestion and orchestration pipelines using Kafka, Snowflake, and Python - Hands-on experience with DBT for data modeling and pipeline development - Strong knowledge ...
... improve platform performance and resilience. Integrate with customer data platforms and pipelines, including bespoke data frameworks. 4–8 years of experience in data engineering or backend development in data-intensive environments. Proficient in Python and SQL; Strong experience with cloud-native data tools and services (S3, ...
... members is maximized throughout the year Desired Skills: - Good level of proficiency in a structured programming language, e.g. Python, R. - Experience designing data science solutions to business problems - Deep understanding of ML algorithms for common use cases in both structured and unstructured data ecosystems. - Comfortable ...
... across the data space. 3+ years of experience designing and building scalable distributed data pipelines and dimensional data models - 3+ years of experience in Python and SQL - Experience using Databricks platform and PySpark is a must. - Extensive experience of Microsoft Azure data services – Data Factory, ADLS gen2, Event ...
... drives / memory / transceivers and other hardware related tasks, as required by our server and network teams. Assisting with planning equipment moves or new data center build-outs. Coordinating the packing, shipping, and logistics of equipment to and from remote colocation sites. Maintaining data center documentation. ...
... connectors or processing frameworks. - Deep understanding of Apache Airflow and ability to design and manage complex DAGs. - Solid SQL skills and familiarity with data warehouse platforms (e.g., Snowflake, Redshift, BigQuery). - Familiarity with version control tools (Git), CI/CD pipelines, and Agile methodologies. - Exposure ...
... flexibility. - SQL Mastery: Advanced SQL skills with the ability to write complex transformations, window functions, and optimization techniques in Snowflake. - Python : Proficiency in Python for data manipulation, pipeline scripting, and automation. - Business Acumen: You can sit in a discovery session with a business stakeholder, ...
... of our transformation, we know that data is at the center of everything that we do. We seek a dedicated professional to help take us to the next level of our Data & Analytics journey. We depend on our Data Analysts to investigate data seeking opportunities to reduce duplicates, improve the quality and discover anomalies ...
... Responsibilities Administer and support Cloudera CDP/CDH platforms, including HDFS, Hive, Spark, YARN, Hue, and CDE. Develop, deploy, and optimize PySpark and Python-based data processing solutions. Build and maintain CI/CD pipelines using Jenkins and GitHub/Bitbucket. Integrate security and code quality tools such as Checkmarx ...
... code using Python. Develop and optimize large-scale data transformations using PySpark. Ensure data quality, reliability, and performance of data pipelines. GCP BigQuery Data Ingestion & ETL/ELT Data Pipeline Orchestration Python PySpark Strong SQL and data engineering fundamentals hands-on experience in GCP-based data platforms.
... rely on you to drive. Who You Are Must-Have - 25 years building production data pipelines: real systems with real consumers, not just one-off scripts - Strong data engineering fundamentals : data modeling, batch vs. streaming, idempotency, incremental processing, partitioning. - Expert SQL and strong Python : query optimization, ...
Roles and Responsibilities Architect and maintain enterprise-grade ELT and ETL data pipelines using Python, PySpark, Kafka, and Databricks to manage large-scale risk data. Build and deploy GenAI agents utilizing Google ADK, Google Flash 2.5+ LLMs, and Model Context Protocol (MCP) integrated with Human-in-the-Loop workflows. ...
... architectures - Deep SQL expertise; hands-on experience designing schemas and tuning queries at production scale - Extensive, hands-on experience with AWS cloud data services for storage, compute, messaging, and eventing - Advanced Python proficiency for pipeline development; experience structuring code for maintainability ...
... metrics.- Support test cases, documentation and data validation for dashboard https://jobeax.com/link/uZNeF9ElGyYiAC23 We Expect From You :- Build strong SQL and data-hygiene habits early; accuracy matters more than speed at this stage.- Ask for help or clarification rather than guessing on ambiguous requests.- Document analysis ...
... with Data Architects, Data Stewards, and Data Quality Engineers to design data pipelines and recommends ongoing optimization of data storage, data ingestion, data quality and orchestration. Designs, develops, and implements ETL/ELT/CDC processes using Data Build Tool (DBT) and other native GCP Services (BigQuery Subscriptions, ...
... activities and site communications. Collaborate with the project team to address issues, resolve discrepancies, and ensure trial milestones are met. Support data management activities by ensuring timely and accurate data collection and entry. Participate in audits and inspections as required, providing necessary documentation ...
... Data Engineer will design, build, and maintain scalable data pipelines and solutions, ensuring high performance and reliability across COSMOS DB and related data platforms. Daily responsibilities include modeling and structuring data, implementing ETL processes, optimizing data storage and retrieval, and collaborating ...
Key responsibilities :Data Architecture & Platform Build :- Design and implement modern data architecture (Data Lake, Data Warehouse, Lakehouse)- Enable self-service analytics and reporting platforms- Lead implementation of BI tools and dashboards- Partner with business teams to define KPIs, metrics, and insights frameworks- ...