This contract role is for a Big Data Engineer experienced in designing and maintaining large-scale data infrastructure, pipelines, and processing systems.
You’ll apply your engineering expertise to build reliable data solutions and contribute high-quality technical input to the training and improvement of next-generation AI systems. The work involves large-scale data processing, Python development, distributed systems, databases, data quality, and governance.
What You’ll Do
Design, build, and maintain scalable data pipelines and architectures for large-scale data environments.
Translate business and technical requirements into reliable data engineering solutions.
Develop data integration, transformation, and processing workflows using Python and appropriate big data technologies.
Build, manage, and optimize distributed databases and storage systems for performance, reliability, and scalability.
Monitor data infrastructure, troubleshoot failures, and improve system availability and performance.
Implement data quality, security, and governance practices across data pipelines and platforms.
Apply sound data modeling, ETL, and data warehousing principles to support reliable data systems.
Collaborate with technical and non-technical stakeholders to understand requirements and communicate implementation decisions.
Document architectures, processes, technical solutions, and operational procedures clearly.
Requirements
Proven hands-on experience in big data engineering, including building and maintaining large-scale data pipelines.
Advanced proficiency in Python for data processing, automation, and system integration.
Strong understanding of relational and NoSQL databases, including database design, optimization, and administration.
Experience with distributed data processing frameworks such as Hadoop, Spark, or Flink.
Solid understanding of data modeling, ETL processes, and data warehousing principles.
Strong troubleshooting and analytical skills for diagnosing performance, reliability, and data-quality issues.
Ability to communicate complex technical concepts clearly to both technical and non-technical stakeholders.
Detail-oriented, proactive, and comfortable working independently in a remote environment.
Preferred Qualifications
Experience working in fast-paced or startup-like environments.
Experience collaborating with globally distributed teams.
Hands-on experience with cloud-based data platforms and services, including AWS, Google Cloud, or Microsoft Azure.
Familiarity with MLOps, machine learning infrastructure, or data science workflows.
Experience building data systems that support analytics, machine learning, or other high-volume data applications.
Who Should Apply
This role is suited to data engineers who have practical experience working with large-scale datasets and distributed data systems. You should be comfortable moving between pipeline development, database optimization, data processing, infrastructure reliability, and data governance.
Strong independent problem-solving and communication skills are important, particularly for engineers working remotely and across multidisciplinary teams.
Compensation
Annual compensation equivalent: $62,400–$166,400
Hourly rate: $30–$80/hour
Annual equivalent is calculated at 40 hours per week × 52 weeks per year.
Actual earnings will depend on hours worked and the duration of the contract.
Work Arrangement
Contract position
Fully remote
Work focused on large-scale data engineering, distributed processing, databases, pipelines, and data infrastructure
Collaboration with cross-functional technical and business stakeholders
... Engineer will own the end-to-end operationalisation of machine learning, large language model (LLM), and agentic AI workloads on the Bajaj Finance Enterprise Data Platform — a 5PB+ medallion lakehouse built on Azure Databricks and Unity Catalog. This role sits at the intersection of data engineering, model lifecycle management, ...
Job Summary We are looking for a skilled Data Engineer with a strong background in building scalable, high-performance data pipelines on Google Cloud Platform (GCP) . The ideal candidate will have hands-on experience with BigQuery , Airflow , and cloud-based ETL workflows , along with solid programming expertise in Java ...
... https://jobeax.com/link/fYGlpInxuIv5PeS1 & Skills :- 4 - 7 years in data engineering: SQL, Python, ETL/ELT orchestration.- Cloud data platform experience (Azure preferred): pipelines, storage, APIs.- Strong understanding of data lake and data architecture.- Data modelling for analytics (star schema) and data quality frameworks. (ref:hirist.tech)
Job Title: Data Engineer - ML Training Data Pipeline Notice period: 0-30 Days Experience : 5+ Years Location: Hyderabad OR Pune We are looking for Data Engineer - ML Training Data Pipeline who can Build and maintain the data pipeline that transforms raw production traces into high-quality training datasets for LLM fine-tuning-ingestion, ...
... in driving value for our customers by building data solutions. You'll be carrying out data engineering tasks to build, maintain, test and optimise a scalable data architecture, as well as carrying out data extractions, transforming data to make it usable to data analysts and scientists, and loading data into data platforms. ...
... working with big data and ETL development 3-5 years of experience in big data analytics technologies like Spark/Spark streaming, , Kafka streaming, Elasticsearch, Bigquery etc. Extensive years of experience in RDMBS/ NoSQL databases, Enterprise Messaging Application (Kafka) and Big data (Spark) Hands-on experience in designing ...
Consultant Data Engineer Databricks | Azure | PySpark | Spark SQL - Bengaluru / Hyderabad / GurugramLooking for an experienced Data Engineer with strong expertise in Azure Databricks, PySpark, Spark SQL, Azure Data Factory (ADF), ADLS, ETL, and SQL. You will design scalable data pipelines, build modern Lakehouse solutions, ...
Consultant Data Engineer Databricks | Azure | PySpark | Spark SQL - Bengaluru / Hyderabad / GurugramLooking for an experienced Data Engineer with strong expertise in Azure Databricks, PySpark, Spark SQL, Azure Data Factory (ADF), ADLS, ETL, and SQL. You will design scalable data pipelines, build modern Lakehouse solutions, ...
... Airflow workflows and production data science infrastructureIDEAL PROFILE:Looking for candidates with strong hands-on experience in:- Data Operations- Production Data Engineering- Data Platform Support- Data Pipeline Monitoring & Troubleshooting- Not looking for purely development-focused Data Engineers or DevOps/SRE https://jobeax.com/link/V6ZwWogXLIs5s3d8 ...
... Airflow workflows and production data science infrastructureIDEAL PROFILE:Looking for candidates with strong hands-on experience in:- Data Operations- Production Data Engineering- Data Platform Support- Data Pipeline Monitoring & Troubleshooting- Not looking for purely development-focused Data Engineers or DevOps/SRE https://jobeax.com/link/V6ZwWogXLIs5s3d8 ...
... and results-driven Data Engineer to join our growing Data Engineering team. You will be instrumental in ingesting, transforming, and delivering high-quality data to enable data-driven decision-making for our clients. This role is ideal for someone who thrives in a consulting environment, enjoys solving complex data challenges, ...
... role Ethos is seeking a Data Engineer to join our Data Platform team and build the internal infrastructure, services, and tooling that the rest of the company's data stack runs on. This is a software engineering role with a deep data background. You will design and operate distributed workflow systems on Temporal and Airflow, ...
... & Data Platform. You’ll design and maintain reliable, secure, and scalable data pipelines and models that power a new generation of reporting, insights, and data-driven products across the StarRez ecosystem. This role sits at the intersection of data, product, and engineering. You will work closely with Data Analysts, ...
Role Overview:Seeking a highly skilled Senior Data Engineer with deep expertise in Databricks, Real-Time Data Processing, Lakehouse Architecture, and AI-driven solutions. The role involves designing, developing, and supporting scalable data platforms and streaming pipelines, leveraging modern Databricks capabilities such ...
... architectures to ingest, process, and deliver high-quality data to business stakeholders. - Develops and optimizes batch-processing pipelines, ensuring collected data is structured and formatted for immediate, analytical consumption. - Adheres to and promotes data engineering standards and best practices within the immediate ...
Project Role : Data Engineer Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems. AWS AI Services Minimum 3 Year(s) ...
... Experience with system integration, implementation, upgrades, migrations, or technology modernization. - Good understanding of application, infrastructure, cloud, database, networking, or platform technologies relevant to the role. - Familiarity with automation, monitoring, DevOps, CI/CD, cloud platforms, or modern engineering ...
... members of IT, including business analysts, database administrators, and managers, to design and develop Cognos reporting capability- Exhibit a commitment to data quality by validating results against sources specified in the requirements- Develop and maintain stored procedures- Write and support queries against databases ...
... in data security and privacy engineering, this individual deploys and tunes data discovery, classification, DLP, encryption, and key management controls. The Data Protection Engineer partners closely with Engineering, IT, Compliance, Legal, and Privacy stakeholders to embed privacy-by-design principles into how data is ...
About the Role:KPI Partners is seeking a skilled Azure Data Engineer with expertise in data engineering, cloud solutions, and experience with data pipeline development using Azure services. You will play a key role in designing and implementing scalable data solutions that drive insights and business decisions for our https://jobeax.com/link/OBSYxKmspgkYWOcn ...