... Design, Performance Tuning.- Data Integration : Apache Kafka, AWS DMS, REST APIs, CDC.- Metadata & Governance : Apache Atlas (preferred), Metadata Management, Data Catalog, Business Glossary, Data Lineage, Data Stewardship, Data Quality.- Business Intelligence : Tableau, Semantic Layer, KPI Frameworks, Data Mart Design, ...
... machine learning infrastructure, or data science workflows. - Experience building data systems that support analytics, machine learning, or other high-volume data applications. Who Should Apply This role is suited to data engineers who have practical experience working with large-scale datasets and distributed data systems. ...
... machine learning infrastructure, or data science workflows. - Experience building data systems that support analytics, machine learning, or other high-volume data applications. Who Should Apply This role is suited to data engineers who have practical experience working with large-scale datasets and distributed data systems. ...
... logical data models (3NF) and their application within IT project and business data governance processes.- Defines data modelling standards and applies within all work products.- Owns and develops artefacts that exploit data models to understand data and data integrations. This may include modelling data flows, developing CRUD ...
... - Collect, clean, organize, and validate datasets. - Perform exploratory data analysis (EDA) to identify patterns, trends, and anomalies. - Analyze business data and identify meaningful insights. - Prepare analytical reports and summaries. - Create and maintain data visualizations and dashboards. - Work with spreadsheets ...
... creation and review process. Requirements - Professional experience in data entry, data validation, quality assurance, or a closely related field. - Experience working with high-accuracy or data-integrity-sensitive processes. - Strong ability to identify data inconsistencies, missing information, formatting issues, and silent ...
... test accuracy, error detection, and data reconciliation. This opportunity is suited to professionals who are meticulous, process-oriented, and experienced in working with complex datasets where accuracy and data integrity are critical. You’ll design realistic data entry and validation challenges, create datasets containing ...
... Delta Lake tables and Lakehouse solutions for analytics and reporting.- Implement governance and security using Unity Catalog.- Optimize Databricks jobs, Spark workloads, and SQL queries for performance and cost efficiency.- Develop and support real-time and batch data processing solutions.- Configure Databricks Notification ...
... datasets . Strong understanding of data accuracy, completeness, relevance, and integrity. Experience handling PII and sensitive information . Strong knowledge of data privacy practices and protocols. Experience with document review, data quality assurance, or audit-related work . Strong analytical and problem-solving skills. ...
... maintaining large-scale data pipelines.- Advanced proficiency in Python for data processing, automation, and integration.- Deep understanding of relational and NoSQL databases, including optimization and management techniques.- Experience with distributed data processing frameworks (e.g., Hadoop, Spark, Flink).- Strong foundation ...
... specifically designed to support AI/ML workflows, including feature engineering and model training datasets. Design and implement robust data pipelines, data lakes, and data warehousing solutions. Partner with data scientists and analysts to deploy machine learning models and analytics solutions on Databricks. Enhance Databricks cluster ...
... or structured analytical tasks - Background in data annotation or evaluation task development - Experience creating technical reports involving statistics or data science - Previous experience contributing to AI, machine-learning or data-driven evaluation projects - Bilingual proficiency in English and a second language ...
... or structured analytical tasks - Background in data annotation or evaluation task development - Experience creating technical reports involving statistics or data science - Previous experience contributing to AI, machine-learning or data-driven evaluation projects - Bilingual proficiency in English and a second language ...
... respond to infrastructure https://jobeax.com/link/5UI7zp9uZQKTl6Ox Stack & Requirements:- 6+ years of professional experience in cloud infrastructure, DevOps, and database engineering.- Hands-on experience with Microsoft Azure (VMs, Networking, Storage, DevOps pipelines, CLI, Load Balancer, Key Vault).- Strong working knowledge ...
... you an all-around data enthusiast with a knack for ETL We're hiring Data Engineers to help build and optimize the foundational architecture of our product's data. We've built a strong data engineering team to date, but have a lot of work ahead of us, Migrating from relational databases to a streaming and big data architecture, ...
... new technologies and approaches to innovate with increasingly large data sets. Drive automation and efficiency in Data ingestion, data movement and data access workflows by innovation and collaboration. Understand, implement and enforce Software development standards and engineering principles in the Big Data space. Work ...
... 15-minute schedules), including idempotency, backfills, and late-arriving data patterns - Experience with orchestration and transformation tooling (e.g., Strong data quality discipline: automated tests/expectations, monitoring/alerting, and clear SLAs for critical datasets - Comfortable working hybrid in Bengaluru (Whitefield) ...
... Bengaluru office. You will work on a broad range of cutting-edge data science and machine learning problems across a variety of industries. If you are passionate to work on complex unstructured business problems that can be solved using data science and machine learning we would like to talk to you. Function : Data Science and ...
... security, and scalability while enabling faster and more reliable software delivery. Key Responsibilities: - Hands on experience working on Databricks, creating data pipelines, transformation rules using Unity Catalog, PySpark, SparkSQL - Hands on Databricks Devs – Deployments, CICD, Performance Management - Hands on Databricks ...
... Implement efficient, distributed and scalable pipelines and integrate data from multiple sources to create data products Implement design patterns that support data ingestion, data movement, transformation, aggregation, and much more. Collaborate with Product managers, Software engineers, and Data Scientists, and work towards ...