Data Engineer (FrontEnd) in Mumbai, India - Jobeax
Vacancy description
Data Engineer (FrontEnd) in Mumbai, India
Skillyt
Rs 22 - 30 lakhs p.a.
India, Mumbai
Data Engineer (FrontEnd) in Mumbai, India is listed on Jobeax. Browse 30,000+ vacancies available.
The Role You own the data substrate and the batch runtime of a product we are building from scratch. This is a data-specialist seat, not a general backend seat. Routine application development is guided in weekly reviews; what cannot be substituted is your depth in data modeling, set-based computation, and pipeline correctness. You will work self-directed, writing your own specs and decisions, with your tests and documentation carrying the quality bar between reviews. Stack: Python, PostgreSQL, Django, server-rendered frontend (htmx). What You Will Own Core Responsibilities / Engineering Focus Deep third-party data ingestion Build reliable ingestion pipelines for external data sources Account for silently dropped webhooks through reconciliation with source-of-truth totals Large-scale batch computation Use set-based SQL for processing large datasets Ensure jobs complete within fixed nightly processing windows Multi-tenant reliability Maintain strict tenant isolation Ensure one tenant’s bad data does not impact another tenant’s processing Backfill & replay capabilities Design systems to support backfills and replays as standard capabilities Avoid relying on one-off emergency scripts Append-only & auditable data Maintain immutable, auditable records Ensure every automated decision can be reconstructed and traced Resilient vendor API integrations Build for failures and unreliable external systems Implement retries, circuit breakers and reconciliation mechanisms Operational reliability Build strong alerting and monitoring Handle database partitioning effectively Design and maintain idempotent jobs
What We Are Looking For Required Ideal Candidate Profile
5–8 years of experience building and operating production systems — years are a proxy; the skills below are the actual bar.
Strong PostgreSQL expertise
Schema design
Query planning and optimization
Batch performance at real scale
Production data pipeline ownership Ingestion Transformation Reconciliation Data serving Experience should go beyond application CRUD/ORM work. Strong SQL & data thinking Thinks in sets, not loops Multi-step SQL transformations Window functions Statistical aggregates, percentiles & distributions Incremental computation Experience building cohort, retention or funnel metrics from raw event/order data Production Python backend experience Experience with any Python backend framework is fine Able to keep business logic cleanly separated from framework-specific code Strong testing mindset Tests are part of the development process, not an afterthought Especially experienced with testing batch jobs and reconciliation systems where data errors can remain unnoticed Production problem-solving / “war stories” Has dealt with real incidents such as:
Queues backing up Webhooks silently dropping Data loss or inconsistencies Reconciliation catching issues that went unnoticed
Strong written communication Comfortable with documentation-driven development Can write clear specs, decisions and technical documentation High ownership Comfortable building systems from scratch Can independently own outcomes with minimal supervision
Good to have Production Django experience E-commerce / D2C platform experience Streaming & event-driven architecture Kafka or similar systems Familiarity with probabilistic data structures HyperLogLog Bloom Filters t-digest ~ Comfort with server-rendered applications/screens when required
Job Description – Senior Data Engineer / Platform Re-Engineering Lead (Azure Synapse & Databricks Migration) Role Overview We are looking for a highly experienced Senior Data Engineer to lead the re-engineering of an existing enterprise data platform built on Azure Synapse Analytics. The role requires deep technical seniority ...
Data Science Data Engineer Work Type: Full Time At VIDA we're building the future of digital identity. As a Data Engineer you'll work across the stack—from platform and infrastructure to data pipelines and end-user tooling. You'll help us modernize our batch ETLs and lead our push into streaming and real-time fraud detection. ...
About the Role:We are looking for an experienced AWS Data Engineer with 6 - 8 years of hands-on experience in designing, developing, and maintaining scalable data engineering solutions on AWS. The ideal candidate should have strong expertise in Python, PySpark, SQL, and AWS data services, with experience building robust ...
... details related to data engineering or a relevant https://jobeax.com/link/cwwOL3BLJkve6nkU Skills :- Azure Databricks, Python, pyspark, sql, ADF- Proficiency in data pipeline design and implementation- Strong understanding of data factory and orchestration toolsPreferred Skills:- Familiarity with advanced data processing techniques- ...
... work independently and collaboratively to solve complex data challenges and drive continuous improvements across our data engineering practices. 4+ years of data engineering experience, building and managing large-scale data pipelines - Strong proficiency in Python, PySpark, Spark, Databricks, Delta Lake, PowerBI and SQL, ...
Senior Data Engineer Starts Immediately Competitive salary Competitive salary Experience We are looking for a Senior Data Engineer to join our team. In this role, you will design, build, and optimize scalable data pipelines on the Databricks Lakehouse Platform. You will partner with data science, analytics, and business ...
... BigQuery, Cloud Functions, Composer, GCS Proficient hands-on programming experience in Spark/Scala (python/java) Proficient in building production level ETL/ELT data pipelines from data ingestion to consumption Data Engineering knowledge (such as Data Lake, Data warehouse - Redshift/Hive/Snowflake, Integration, Migration) ...
We are seeking a Data Engineer to join our team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL ...
... Posted 2 weeks ago Job Be an early applicant About the job Job Purpose: Design and implement scalable data engineering and data warehouse solutions while managing data schemas, SQL query tuning, and code reviews. Who You Are: - 5+ years of experience in Data Engineering, with strong knowledge of Data Platforms and Data Warehousing ...
... environments (preferably AWS). Document architecture, data flow, and support runbooks; continuously improve platform performance and resilience. Integrate with customer data platforms and pipelines, including bespoke data frameworks. 4–8 years of experience in data engineering or backend development in data-intensive environments. ...
... Responsibilities Develop, test, and maintain scalable Python applications. Work with large datasets to extract, transform, and analyze data. Build and optimize data pipelines and workflows. Perform data cleaning, validation, and preprocessing. Collaborate with cross-functional teams including data analysts and engineers. ...
... responsibilities will include estimating, designing, coding, testing, deploying, and ensuring scalability and performance on Azure using key technologies like Databricks. As a hands-on technologist with an extensive data engineering background using Databricks , you will be joining a group of Data Engineers who are passionate ...
We're looking for a Data Center Engineer to join our Infrastructure Engineering team in Mumbai. In this role, you'll help build, maintain, and scale the critical infrastructure that powers our trading operations across APAC. Working alongside experienced engineers and traders, you'll solve complex technical challenges in ...
Hello, We are hiring for 'Data Engineer with Java' for Pune Location. Exp :5+Years Loc : Pune Notice Period: Immediate joiners(notice period served or serving candidate) Mandatory Skills: Data Engineering Airflow Data Build Tool(DBT) Java ETL/ELT Pipeline CI/CD SQL NOTE: We are looking for Immediate joiners(notice period ...
... analysis to validate source data, understand distributions, and surface data anomalies during scoping and build phases. Work with stakeholders to define and monitor data SLAs, ensuring data products are trustworthy and consistently delivered. What You Bring - 4–6 years of experience in data engineering, analytics engineering, ...
... MAKE AN IMPACT Job Title: Data Engineer _DEPS Experience: 3–5 Years Location: Bangalore and Pune (Hybrid – Client Office) Job Summary: Seeking a skilled Senior Data DevOps Engineer having experience in Cloudera platforms, data engineering, and DevOps automation. The ideal candidate will manage and optimize Cloudera environments, ...
... efficient code using Scala and Spark - Work with Azure or on-premises Hadoop ecosystem for data processing - Solve real-world engineering problems related to data infrastructure - Collaborate with cross-functional teams to deliver data-driven solutions - Ensure data quality, reliability, and performance across data systems ...
... Python. Develop and optimize large-scale data transformations using PySpark. Ensure data quality, reliability, and performance of data pipelines. GCP BigQuery Data Ingestion & ETL/ELT Data Pipeline Orchestration Python PySpark Strong SQL and data engineering fundamentals hands-on experience in GCP-based data platforms.
... Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering experience in Banking Risk, Retail Products, Cards, Mortgage, Deposits, or Wealth Management. Assumed Requirements / Certifications: Databricks Certified Data Engineer, AWS Certified Data Analytics, or Azure Data Engineer Associate.
... workflows, dose analysis, and program management for radiation safety professionals. We are looking for a seasoned Data Engineer to take a lead role in shaping the data foundation across Landauer's growing portfolio of products. You will set the standards for how data is modeled, moved, and governed — owning the full data lifecycle ...