Data Engineer - Python, AWS, SQL in Bengaluru, India
India, Bengaluru
Data Engineer - Python, AWS, SQL in Bengaluru, India is listed on Jobeax. Browse 30,000+ vacancies available.
Position Overview:We are looking for a Data Engineer with hands-on experience in Apache Airflow, Python, and Jupyter Notebook to build, orchestrate, and maintain robust data pipelines that power our analytics and business decisions. In this role, you will work closely with data teams to clean, transform, and move data reliably across our platforms while ensuring pipeline performance and data integrity at https://jobeax.com/link/AwOVwl2FkAY2iw53 Architecture:- Design, build, and maintain scalable ETL/ELT data pipelines using Apache Airflow and https://jobeax.com/link/9BZSgic8TwoqWitR & Workflow:- Schedule, monitor, optimize, and troubleshoot Airflow DAGs to ensure reliable, high-availability data https://jobeax.com/link/zruxWYQz3eqyaZpe Processing:- Perform data extraction, cleaning, transformation, and validation across structured and unstructured multi-source https://jobeax.com/link/sluZEnCJKJyUJkcP & Analysis:- Use Jupyter Notebook for exploratory data analysis, pipeline testing, data profiling, and rapid prototyping.Cross-Team Collaboration:- Partner with data analysts, data scientists, and product teams to translate business requirements into reliable data https://jobeax.com/link/slfm2H5M70Q62wJ5 & Performance:- Ensure high data quality, strict data integrity, and pipeline execution speed at scale.Documentation:- Maintain clean, clear documentation for all data pipelines, lineage, and operational https://jobeax.com/link/iF5NRBBcXKPJeYbp Qualifications:Experience:- 1 to 2 years of hands-on, production experience in data https://jobeax.com/link/lptHFLDlMUnekis0 Proficiency:- Strong proficiency in Python for data manipulation, ETL scripting, and https://jobeax.com/link/DH9RkYsfBywYaX8K Orchestration:- Working experience with Apache Airflow (DAG creation, task dependencies, custom operators, and scheduling).Prototyping Tools:- High comfort level working in Jupyter Notebook for data prototyping and interactive https://jobeax.com/link/zlk2beUnqej1rgqa Foundations:- Solid understanding of SQL, relational/non-relational databases, and data modeling https://jobeax.com/link/SKciAM5XjIOOw9Ja Qualifications:Education:- https://jobeax.com/link/2DENLGgQM7whDhaE / B.E. in Computer Science, Data Science, or a related field, preferably from a premier institution (NIT or IIT).Cloud Infrastructure:- Familiarity with major cloud environments (AWS, GCP, or Azure).Big Data Technologies:- Exposure to distributed computing frameworks like Apache Spark, streaming platforms like Kafka, or cloud data warehouses (e.g., Snowflake, BigQuery, Redshift).Shift Timing:- 10:00 AM 7:00 PM IST (ref:hirist.
... architects, analysts, developers, and business stakeholders to understand data requirements.- Ensure data quality, security, governance, and best practices across data engineering https://jobeax.com/link/Gpaa4oyRTNYXQU0V Stack:- Python, PySpark, SQL, AWS Glue, AWS Glue Data Catalog, AWS Lambda, Amazon S3, AWS Step Functions, ...
... understanding of event-driven architecture and real-time data streaming patterns Familiarity with Avro/JSON schema management and Schema Registry Experience integrating data between legacy and modern platforms Strong SQL/Python skills and experience working with relational databases Understanding of data mapping, data quality, and ...
... and scalable data infrastructure and data products in complex environments Development experience in one or more object-oriented programming languages (e.g. Python, Scala, Java, C#) Experience with SQL and noSQL database fundamentals, query structures and design best practices, including scalability, readability, and reliability ...
... and development Data quality validation and monitoring Schema design for analytics and reporting Performance optimization and scalability Preferred Experience Databricks Delta Lake experience Azure Synapse Analytics Python for data engineering Spark SQL optimization Real-time data streaming Data governance and metadata management ...
... Posted 2 weeks ago Job Be an early applicant About the job Job Purpose: Design and implement scalable data engineering and data warehouse solutions while managing data schemas, SQL query tuning, and code reviews. Who You Are: - 5+ years of experience in Data Engineering, with strong knowledge of Data Platforms and Data Warehousing ...
... a related field (Master's degree preferred). Core Technical Skills: Python, PySpark, Databricks, Google ADK, LLMs, FastAPI, Spring Boot, Microservices, Kafka, SQL, Data Mesh, Starburst. Infrastructure and Cloud: Kubernetes, OpenShift, Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering ...
... scalable, and high-performance full-stack web applications using Python (FastAPI) on the backend and https://jobeax.com/link/6WcBICn0LM558SP0 on the frontend. Database Design & Management: Architect and optimize both relational ( SQL ) and non-relational ( NoSQL ) database schemas to handle complex data structures efficiently. ...
Job Description: We are looking for a skilled Python SQL and ETL Developer to design build and maintain scalable data pipelines and data integration solutions The ideal candidate will work closely with data engineers analysts and business stakeholders to ensure reliable high quality data delivery for reporting analytics ...
Data Engineer: PySpark+AWS Glue 4-8 years in Data Engineering with hands-on expertise in Snowflake, AWS Glue, AWS (S3/Lambda/CloudWatch), Python, PySpark, and SQL. Strong experience in ETL/ELT, Data Warehousing, Data Pipelines, Performance Tuning, and Cloud Data Platforms. Exposure to AI/ML or GenAI solutions is a plus ...
... governance. Experience : Minimum 6 years of hands-on experience in data engineering, with a proven track record in complex pipeline development and cloud-based data migration projects. Bachelor’s or higher degree in Computer Science, Data Engineering, or a related field. Skills : Proficiency in Spark, SQL, Python, and other ...
... scalability. Required Skills: - 7+ years of Data Engineering experience. - Strong experience in Data Modeling, Data Warehousing, and ETL development. - Proficiency in Python, Java, Scala, or NodeJS. - 3+ years of SQL experience. - Experience building and operating distributed data systems. Preferred Skills: AWS technologies: Redshift, ...
... experienced engineers across the development lifecycle. Design and build scalable backend services, REST APIs, integrations, and reusable software components using Python and, where relevant TypeScript. Develop data ingestion, transformation, service-to-service communication, and automation capabilities that support reliable product ...
... in Python and SQL - Experience with Unity Catalog, cluster administration, and at least one major cloud platform (Azure, AWS, or GCP) - Solid understanding of data warehousing and dimensional modeling Databricks Certified Data Engineer (Associate or Professional) Cloud data engineering certification have minimum 5 years ...
... least four GCP services among Data Flow, Data Proc, Pub Sub, BigQuery, Cloud Functions, Composer, GCS Proficient hands-on programming experience in Spark/Scala (python/java) Proficient in building production level ETL/ELT data pipelines from data ingestion to consumption Data Engineering knowledge (such as Data Lake, Data warehouse ...
Hello, We are hiring for 'Data Engineer with Java' for Pune Location. Exp :5+Years Loc : Pune Notice Period: Immediate joiners(notice period served or serving candidate) Mandatory Skills: Data Engineering Airflow Data Build Tool(DBT) Java ETL/ELT Pipeline CI/CD SQL NOTE: We are looking for Immediate joiners(notice period ...
... using Python and PySpark to ingest data from files, APIs, RDBMS, NoSQL, and message queues into AWS data stores Develop and optimize PySpark jobs for large-scale data processing, transformation, and aggregation on AWS EMR or Databricks Build and manage AWS-native data workflows using S3, Glue, Lambda, Redshift, and Step Functions; ...
... JavaScript. - Knowledge of database systems (e.g., PostgreSQL, MySQL, MongoDB). - Familiarity with version control tools like Git. - Understanding of RESTful APIs and web services integration. - Experience with cloud services (e.g., AWS, Azure) is a plus. - Strong analytical and problem-solving skills, with attention to detail.