This contract role is for a Big Data Engineer experienced in designing and maintaining large-scale data infrastructure, pipelines, and processing systems.
You’ll apply your engineering expertise to build reliable data solutions and contribute high-quality technical input to the training and improvement of next-generation AI systems. The work involves large-scale data processing, Python development, distributed systems, databases, data quality, and governance.
What You’ll Do
Design, build, and maintain scalable data pipelines and architectures for large-scale data environments.
Translate business and technical requirements into reliable data engineering solutions.
Develop data integration, transformation, and processing workflows using Python and appropriate big data technologies.
Build, manage, and optimize distributed databases and storage systems for performance, reliability, and scalability.
Monitor data infrastructure, troubleshoot failures, and improve system availability and performance.
Implement data quality, security, and governance practices across data pipelines and platforms.
Apply sound data modeling, ETL, and data warehousing principles to support reliable data systems.
Collaborate with technical and non-technical stakeholders to understand requirements and communicate implementation decisions.
Document architectures, processes, technical solutions, and operational procedures clearly.
Requirements
Proven hands-on experience in big data engineering, including building and maintaining large-scale data pipelines.
Advanced proficiency in Python for data processing, automation, and system integration.
Strong understanding of relational and NoSQL databases, including database design, optimization, and administration.
Experience with distributed data processing frameworks such as Hadoop, Spark, or Flink.
Solid understanding of data modeling, ETL processes, and data warehousing principles.
Strong troubleshooting and analytical skills for diagnosing performance, reliability, and data-quality issues.
Ability to communicate complex technical concepts clearly to both technical and non-technical stakeholders.
Detail-oriented, proactive, and comfortable working independently in a remote environment.
Preferred Qualifications
Experience working in fast-paced or startup-like environments.
Experience collaborating with globally distributed teams.
Hands-on experience with cloud-based data platforms and services, including AWS, Google Cloud, or Microsoft Azure.
Familiarity with MLOps, machine learning infrastructure, or data science workflows.
Experience building data systems that support analytics, machine learning, or other high-volume data applications.
Who Should Apply
This role is suited to data engineers who have practical experience working with large-scale datasets and distributed data systems. You should be comfortable moving between pipeline development, database optimization, data processing, infrastructure reliability, and data governance.
Strong independent problem-solving and communication skills are important, particularly for engineers working remotely and across multidisciplinary teams.
Compensation
Annual compensation equivalent: $62,400–$166,400
Hourly rate: $30–$80/hour
Annual equivalent is calculated at 40 hours per week × 52 weeks per year.
Actual earnings will depend on hours worked and the duration of the contract.
Work Arrangement
Contract position
Fully remote
Work focused on large-scale data engineering, distributed processing, databases, pipelines, and data infrastructure
Collaboration with cross-functional technical and business stakeholders
... Product engineering team to influence designs with data, AI and analytics use cases in mind - In depth experience in System design involving large Petabytes of data with Databricks Lakehouse - Experience in modern AI/Data infrastructure patterns, Semantics layer Organizing data for AI agents (metadata, context) - AWS, GCP, ...
... Science, Engineering, or a related field. - Over 5 years of software development experience, with at least 2 years in a technical leadership role overseeing data engineering or data platform teams. - Strong proficiency in big data technologies, including Apache Spark/PySpark, Apache Airflow and Apache Kafka. - Experience ...
... enterprise, recognized as a global Microsoft and Databricks https://jobeax.com/link/z5IScpoZA6T2iqD7 bridge enterprise challenges with modern intelligence across Big Data, Cloud Infrastructure, Data Engineering, and Applied Artificial https://jobeax.com/link/Z4S3vFYKsGGU38on Overview :We are seeking an experienced Senior Data Scientist ...
... enterprise, recognized as a global Microsoft and Databricks https://jobeax.com/link/z5IScpoZA6T2iqD7 bridge enterprise challenges with modern intelligence across Big Data, Cloud Infrastructure, Data Engineering, and Applied Artificial https://jobeax.com/link/Z4S3vFYKsGGU38on Overview :We are seeking an experienced Senior Data Scientist ...
... cost.- Big-data and pipeline fluency: Advanced SQL plus distributed processing for large spatial workloads (Spark or Dask), and building reliable, repeatable data pipelines.- Productionizing models: Experience turning models into deployable, real-time APIs in collaboration with engineering - clean, tested, well-documented ...
... Experience with data quality tooling (SODA, Collibra, or similar) Exposure to cloud platforms (Azure, AWS, or GCP) Experience in a regulated or enterprise-scale data environment Prior experience mentoring or leading a small pod of engineers EXPERIENCE - 8+ years in data engineering, with at least 4+ years focused on Snowflake ...
... Azure Data Factory, Azure Data Lake, Microsoft Fabric, Logic Apps,Power BI, GitHub, and Data Modeling. Required Skills: - 10+ years of Software Engineering / Data Engineering experience. - Strong SQL Server and T-SQL expertise. - Strong experience working in the Azure environment. - Hands-on experience with Azure Data Factory. ...
... checks Contribute to and help shape company-wide data governance standards Collaborate with analytics, BI, and business teams to deliver trusted, well-modeled data 5–7 years of hands-on data engineering experience - Proven production data engineering experience at scale - Strong Python and SQL skills - Deep analytical warehouse ...
... production issues, work with Databricks to open incident tickets/case and follow up to close the issue and ensure long term fixes - Additional knowledge of other Databricks key features, like Genie, Apps, Models creations etc Required Qualifications: - Bachelor's degree in computer science, Information Technology, Engineering, ...
... Engineer will own the end-to-end operationalisation of machine learning, large language model (LLM), and agentic AI workloads on the Bajaj Finance Enterprise Data Platform — a 5PB+ medallion lakehouse built on Azure Databricks and Unity Catalog. This role sits at the intersection of data engineering, model lifecycle management, ...
Job Summary We are looking for a skilled Data Engineer with a strong background in building scalable, high-performance data pipelines on Google Cloud Platform (GCP) . The ideal candidate will have hands-on experience with BigQuery , Airflow , and cloud-based ETL workflows , along with solid programming expertise in Java ...
... https://jobeax.com/link/fYGlpInxuIv5PeS1 & Skills :- 4 - 7 years in data engineering: SQL, Python, ETL/ELT orchestration.- Cloud data platform experience (Azure preferred): pipelines, storage, APIs.- Strong understanding of data lake and data architecture.- Data modelling for analytics (star schema) and data quality frameworks. (ref:hirist.tech)
... in driving value for our customers by building data solutions. You'll be carrying out data engineering tasks to build, maintain, test and optimise a scalable data architecture, as well as carrying out data extractions, transforming data to make it usable to data analysts and scientists, and loading data into data platforms. ...
Consultant Data Engineer Databricks | Azure | PySpark | Spark SQL - Bengaluru / Hyderabad / GurugramLooking for an experienced Data Engineer with strong expertise in Azure Databricks, PySpark, Spark SQL, Azure Data Factory (ADF), ADLS, ETL, and SQL. You will design scalable data pipelines, build modern Lakehouse solutions, ...
... working with big data and ETL development 3-5 years of experience in big data analytics technologies like Spark/Spark streaming, , Kafka streaming, Elasticsearch, Bigquery etc. Extensive years of experience in RDMBS/ NoSQL databases, Enterprise Messaging Application (Kafka) and Big data (Spark) Hands-on experience in designing ...
... Airflow workflows and production data science infrastructureIDEAL PROFILE:Looking for candidates with strong hands-on experience in:- Data Operations- Production Data Engineering- Data Platform Support- Data Pipeline Monitoring & Troubleshooting- Not looking for purely development-focused Data Engineers or DevOps/SRE https://jobeax.com/link/V6ZwWogXLIs5s3d8 ...
... and results-driven Data Engineer to join our growing Data Engineering team. You will be instrumental in ingesting, transforming, and delivering high-quality data to enable data-driven decision-making for our clients. This role is ideal for someone who thrives in a consulting environment, enjoys solving complex data challenges, ...
... role Ethos is seeking a Data Engineer to join our Data Platform team and build the internal infrastructure, services, and tooling that the rest of the company's data stack runs on. This is a software engineering role with a deep data background. You will design and operate distributed workflow systems on Temporal and Airflow, ...
... & Data Platform. You’ll design and maintain reliable, secure, and scalable data pipelines and models that power a new generation of reporting, insights, and data-driven products across the StarRez ecosystem. This role sits at the intersection of data, product, and engineering. You will work closely with Data Analysts, ...
Role Overview:Seeking a highly skilled Senior Data Engineer with deep expertise in Databricks, Real-Time Data Processing, Lakehouse Architecture, and AI-driven solutions. The role involves designing, developing, and supporting scalable data platforms and streaming pipelines, leveraging modern Databricks capabilities such ...