Lead Data Engineer SQL AWS in Bengaluru, India - Jobeax
Vacancy description
Lead Data Engineer SQL AWS in Bengaluru, India
Arrive
India, Bengaluru
Lead Data Engineer SQL AWS in Bengaluru, India is listed on Jobeax. Browse 30,000+ vacancies available.
As Arrive, we guide customers and communities towards brighter futures and more livable cities, it isn't a challenge just anyone could take on. Luckily, we have something to help us make it happen. Just as our entire brand is inspired by the North Star, the shining light leading travelers to their destinations since time began, our values guide us. They help us be at our best. For the cities and communities we serve. As a Lead Data Engineer at Arrive, you are a technical cornerstone of our new Bengaluru data hub and a founding member of the team. You don't just build models - you architect the patterns and standards that define how Arrive turns raw data into trusted sources of truth. Working within a Medallion architecture (Bronze, Silver, Gold) on top of a shared data platform (Snowflake, dbt, Fivetran Airflow, AWS), you will lead the evolution of our data models, transforming complex data streams from disparate source systems - many inherited through acquisitions - into unified, high-integrity datasets.
At Arrive, data is owned by the domain teams that produce it; By championing clarity over control and bringing genuine curiosity to legacy systems, you ensure Arrive's data is the strategic catalyst for the future of urban.
Architect the Medallion: Lead the design and evolution of our global data models (Bronze, Silver, Gold). You define the dbt patterns, macros, and testing suites that ensure our Gold layer is the definitive global source of truth.
Mentor and Elevate: Act as a technical mentor for Data Engineers in Bengaluru. You lead by example through high-quality code reviews, pair programming, and fostering shared knowledge - embodying 'Arrive Together' in how you work.
Modernize with Intent: Take ownership of complex brownfield data models inherited from global acquisitions. You unravel legacy transformation logic and refactor it into scalable, well-documented models without disrupting downstream consumers.
Enable Data Producers: Work with domain engineering teams to improve data quality at the source. You help producers understand how their data flows through the Medallion layers and what 'good' looks like.
Drive Production Excellence: Define the standards for 'Always-On'data modeling. You implement automated testing, data quality checks, and CI/CD best practices that set the bar for the entire team.
Pioneer AI-Assisted Engineering: Lead the practical adoption of AI tools and agentic workflows to accelerate data documentation, quality checks, and modeling productivity.
Collaborate with the Platform Team: Work closely with the Data Platform team (who operate Snowflake, Airflow, Fivetran, and AWS infrastructure) to ensure your modeling work leverages the platform effectively and feeds back requirements
Experience: 8+ years of professional experience in data engineering, with a focus on data modelling and transformation.
Programming: Strong proficiency in SQL and Python for data processing and pipeline development.
Data Platforms: Hands-on experience with Snowflake, dbt, and modern cloud data warehouses.
Data Modelling: Deep experience with dbt and modern data modelling practices. Data Platforms: Hands-on experience working with Snowflake or comparable cloud data warehouses
Experience working with Apache Airflow or similar orchestration tools.
Version Control & CI/CD: Experience using GitHub for version control, CI/CD workflows, and automated testing of data models and pipelines.
Cloud Platforms: Familiarity with AWS services such as S3 and cloud-based data infrastructure.
Arrive, including brands like EasyPark, Flowbird, RingGo, ParkMobile and Parkopedia, is a leading global mobility platform. Present in over 90 countries and 20,000 cities, the company helps people and decision-makers make smarter decisions about urban mobility and ease the experience of travel worldwide. Arrive delivers a unique combination of the core ingredients to make cities more livable: from smart payments and optimized car parks to data-driven traffic reduction and support for reinvestment in public transport and green space. It's about more than function, it's about saving time and simplifying the experience of travel for everyone. Travel is more than a journey, it's how you Arrive.
... Redshift, S3, Glue, EMR, Kinesis, Firehose, Lambda, IAM. Workflow orchestration tools such as Airflow and AWS Step Functions. Experience with NoSQL databases and data stores. Understanding of SDLC, CI/CD, code reviews, testing, and software architecture best practices. Preferred Candidate Profile: A hands-on Data Engineer with ...
... experiences. Contribute to AI-enabled features and use AI development tools responsibly to improve coding, testing, research, documentation, and delivery. Apply sound engineering practices across architecture, performance, security, testing, debugging, and maintainability. Deploy and operate services using AWS and CI/CD workflows, ...
... architects, analysts, developers, and business stakeholders to understand data requirements.- Ensure data quality, security, governance, and best practices across data engineering https://jobeax.com/link/Gpaa4oyRTNYXQU0V Stack:- Python, PySpark, SQL, AWS Glue, AWS Glue Data Catalog, AWS Lambda, Amazon S3, AWS Step Functions, ...
... BigQuery, Cloud Functions, Composer, GCS Proficient hands-on programming experience in Spark/Scala (python/java) Proficient in building production level ETL/ELT data pipelines from data ingestion to consumption Data Engineering knowledge (such as Data Lake, Data warehouse - Redshift/Hive/Snowflake, Integration, Migration) ...
Hello, We are hiring for 'Data Engineer with Java' for Pune Location. Exp :5+Years Loc : Pune Notice Period: Immediate joiners(notice period served or serving candidate) Mandatory Skills: Data Engineering Airflow Data Build Tool(DBT) Java ETL/ELT Pipeline CI/CD SQL NOTE: We are looking for Immediate joiners(notice period ...
... expectations are met through disciplined implementation (access controls, lineage, and quality standards) 5+ years building and operating production data pipelines (data engineering—not primarily BI/analytics or data science) - Strong SQL plus strong data modeling skills (dimensional and/or lakehouse modeling) - Hands-on Databricks ...
Job Summary :The Data Integration Engineer will design, develop, and support data integration solutions that extract financial and operational data from enterprise source systems and deliver trusted data to analytics platforms and external vendors. The role emphasizes source-to-target mapping, transformation, validation, ...
... experience in Data Engineering.- Strong Python + SQL expertise.- Strong Data Architecture and System Design skills.- End-to-end production pipeline ownership.- Expert SQL including Window Functions, CTEs, CASE logic.- Strong experience with MySQL / PostgreSQL.- Experience designing Transaction / Fact Tables.- Strong data quality, ...
... with enterprises to design, build, and operate cloud-native, AI-first systems that generate measurable business outcomes. Searce is a Google Cloud Global MSP, AWS Advanced Consulting Partner, Databricks Partner, and Anthropic Claude Partner. As a Lead SRE within the CSRE Delivery team, you will anchor reliability, scalability, ...
... Prometheus, Grafana, ELK, Datadog Agile Tools: Jira, Confluence Additional Expectations The candidate should be a strong team player leading a team with an ownership mindset, capable of independently managing AWS environments, automation frameworks, deployments, and collaborating with global engineering/customer teams.
... frameworks, or extensible platform systems. Backend https://jobeax.com/link/JSd7UYCSg572jJDY, Python, Java, Ruby Frontend React, TypeScript, Ant Design System, Figma AWS, Kubernetes (EKS), S3 Data MySQL, Snowflake, Aerospike, DynamoDB, ScyllaDB AI Tooling CI/CD, TDD, Agile, You Build It You Run It AI-native engineering culture ...
... proven track record of building and operating large-scale, distributed, API-driven systems in production. Expertise in observability, alerting, and reliability engineering , using tools such as Prometheus, Grafana, Datadog, Splunk, ELK, or equivalent ecosystems. Strong command of cloud platforms and open systems ( AWS, Azure, ...
... Department: Data Engineering / Product Engineering Role Overview We are looking for an experienced Data Engineer with 5–7 years of hands-on experience in ETL/ELT, database engineering, Big Data processing, and complex bi-directional data integrations . The candidate should have strong SQL and database expertise, experience working ...
... operations at ELGi. Data Engineering & Architecture Design scalable, production-grade ETL/ELT pipelines across Azure, AWS, and GCP. Develop modular, reusable data engineering frameworks for ingestion, transformation, and orchestration. Build Spark, SQL, and Python workflows in Databricks. Optimize compute, storage, and ...
... least one certification • Databricks Certified Data Engineer Associate OR Databricks Certified Data Engineer Professional Additional Certifications (Preferred) - Databricks Certified Associate Developer for Apache Spark - Cloud platform certifications (Azure Data Engineer Associate, AWS Certified Data Analytics, or Google Cloud ...
... maintain scalable data pipelines and distributed systems. The ideal candidate should have hands-on experience with ETL/ELT pipelines, workflow orchestration, AWS cloud services, data visualization, distributed architectures, and modern data engineering https://jobeax.com/link/TtkK7BHPbancvnFq Responsibilities:- Design, ...
... using Python and PySpark to ingest data from files, APIs, RDBMS, NoSQL, and message queues into AWS data stores Develop and optimize PySpark jobs for large-scale data processing, transformation, and aggregation on AWS EMR or Databricks Build and manage AWS-native data workflows using S3, Glue, Lambda, Redshift, and Step Functions; ...
... Senior Data Engineer with strong hands-on experience building and supporting data solutions on AWS . The role requires solid experience with Amazon Redshift, S3, AWS Glue, Amazon AppFlow, Lambda, and SQL . You will work on data ingestion, transformation, data pipelines, data warehouse solutions, and integrations across different ...
... Data Factory, Databricks, Synapse, dbt (any two – Mandatory). Data Warehousing: Azure SQL Server/Redshift/Big Query/Databricks/Snowflake (Anyone - Mandatory). Data Visualization: Looker, Power BI, Tableau (Basic understanding to support stakeholder queries). Cloud: Azure (Mandatory), AWS or GCP (Good to have). SQL and Scripting: ...
... entity resolution and record linkage. Strong SQL and Python programming skills. Experience with ETL/ELT pipeline development. Familiarity with cloud platforms like AWS, Azure, or GCP. Knowledge of data engineering best practices and scalable architectures. Experience with Apache Spark or PySpark. Knowledge of Docker, Kubernetes, ...