... Must-have: Ability to identify data quality issues, investigate root causes, and apply fixes in line with business rules Experience in data mapping and in executing data remediation and cleansing activities as part of a platform migration SQL (and Python if possible) skills for querying, analysing, and remediating data at scale ...
Project Role : Data Engineer Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems. Data Engineering Minimum 5 ...
... data architecture, cloud platform design, or equivalent data-platform architecture.- Distributed Computing: 10+ years of experience working with distributed data processing or modern cloud data platforms, with strong hands-on exposure to Databricks, Spark, Snowflake, or equivalent technologies.- Data Architecture: Strong ...
... compelling and clear business focused insights. Can work with multiple databases and big data solutions. Work with the data engineering team to develop & maintain data pipelines. Identify, design, and implement internal process improvements and tools to automate data processing and ensure data integrity while meeting data security ...
... Maintain organized digital records while safeguarding confidential customer, payment, and financial information. Support the Controller and Accounting team with data reviews, record corrections, reporting, and general administrative duties as needed. Previous experience in data entry, order processing, administrative support, ...
... Intelligence Platform team, you will: Build out performant APIs to serve consistent data that serve as sources of truth for many systems in Abnormal. Establish data pipelines that ensure our data is updated regularly and reliably Enhance our frameworks to allow us to make changes to our APIs and datasets in an agile manner ...
... and real time data pipelines Build and maintain robust ETL ELT processes for ingesting transforming and loading data from multiple sources Develop and manage data lakes data warehouses and metadata frameworks Collaborate with cross functional teams to understand data requirements and deliver data solutions Ensure data quality ...
... stakeholders. Data Pipeline Design Operations: Design, develop, and support data pipelines to extract, transform, and load (ETL/ELT) data into the new harmonized data platform. Utilize Azure Data Factory (ADF) for orchestration (scheduling, workflow management) and Azure Databricks (Spark) for large-scale data transformation ...
... Science & Supply Chain talent. Collect business requirements from the various stakeholders (Data Scientists, Product Managers, IM) High interest in developing E2E data solutions DevOps/ Agile WoW Experience in designing and building advanced reports Programming languages: Python, PySpark, SQL Data enthusiast Big data frameworks: ...
... reliable, high-quality data, driving excellence in analytics and shaping outcomes that impact industries like H ealthcare and Financial Services . Design and build Data Pipelines using GCP services and BigQuery. Develop and optimize data solutions using Java or Python. Ensure data quality, governance, and compliance across projects. ...
... quality. Big Data - Big Data - Pyspark Database - Database Programming - SQL Data Science and Machine Learning - Data Science and Machine Learning - Apache Spark Data & AI - Data Engineering - Data Quality & Validation Data & AI - Data Engineering - Apache Kafka Big Data - Big Data - Azure databricks Data Science and Machine ...
... huge datasets and have experience with the organization and curation of data for analytics. You have a strategic and long term view on architecting advanced data eco systems. You are experienced in building efficient and scalable data services and have the ability to integrate data systems with AWS tools and services to ...
... (Pandas, Polars) and modern big data tools such as Databricks (Spark), Flink, or Kafka. - Strong understanding of data pipeline development, ETL frameworks, and data lakes/warehouses. - Extensive experience with cloud data platforms such as AWS (S3, Redshift, Lambda, Glue), GCP (BigQuery, Dataflow, Dataproc), or Snowflake. ...
... scalable data processing solutions using Big Data technologies and distributed system s Strong hands-on experience with Python and PySpar k for developing efficient data pipelines and applications Develop and optimize complex SQL querie s, data models, and database solutions Analyze and improve performance of Apache Spark workload ...
... understanding of data modeling, schema design, validation strategies, data transformation, and data quality management . - Experience designing solutions for data deduplication, entity resolution, data enrichment, classification, or record linkage . - Experience with both relational and document-oriented databases; practical ...
... processes to seamlessly integrate data from multiple sources into our data warehouse. Implement advanced data modeling techniques and design efficient, optimized databases to support complex business needs. Ensure the highest levels of data quality and integrity through rigorous data validation and cleansing processes. Collaborate ...
Job Description :We are seeking an experienced Data Science Manager to lead a team of data scientists and analysts in developing data-driven products, insights, and https://jobeax.com/link/5WM6jVJPxXxkaX5M ideal candidate will have a strong background in applied machine learning, data strategy, and stakeholder management, ...
... Kinesis, FireHose, Lambda, and IAM roles and permissions Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases) Experience with dbt (data build tool) for transformation layer management and semantic modeling. Familiarity with data lakehouse ...
... Qualifications Educational qualifications & Experience: Bachelor's degree in computer science, Information Security, or related field. 10–12 years of experience in data loss prevention, data discovery/classification, and database/data security. Experience with DSPM tools and integration in hybrid environments. Familiarity with ...
Data & AI ArchitectLocation: Hyderabad / Bangalore / PuneKey Technologies:Databricks | Microsoft Fabric | GenAI | LLM | RAG | PySpark | Data Lake | Lakehouse | Data Warehouse | Data ModellingRole Overview :We are looking for an experienced Data & AI Architect to design and lead enterprise - scale Data and AI architectures. ...