Data Engineer (AWS, SQL) in Bengaluru, India is listed on Jobeax. Browse 30,000+ vacancies available.
Data Science Data Engineer Work Type: Full Time At VIDA we're building the future of digital identity. As a Data Engineer you'll work across the stack—from platform and infrastructure to data pipelines and end-user tooling. You'll help us modernize our batch ETLs and lead our push into streaming and real-time fraud detection. We run on AWS and make extensive use of Scala Spark, dbt, and Terraform. Build and expose the data lake/lakehouse so teams can understand business and product performance and make better decisions. Design and evolve a data platform that enables scientists and analysts to query, model, and ship features and product metrics. Maintain and fix existing ETLs, improving reliability, observability, and cost while planning safe backfills and replays. Model curated layers and maintain transformation logic (e.g., dbt), including testing, documentation, and data quality. Terraform) and contribute to CI/CD and deployment practices. Partner with Engineering, and Data Science to deliver low-latency signals and features for real-time decisioning. Solid engineering foundation with 5+ years of relevant experience in data engineering and/or distributed systems. Hands-on work building data pipelines (batch and/or streaming) using at least one distributed framework (e.g., Comfortable with SQL and a transformation framework (e.g., Familiarity with AWS services for data (any mix of S3, EMR/Glue, Kinesis/MSK, Lambda, Redshift/Athena) or equivalent cloud experience. Experience using Terraform (or similar IaC) to manage data/platform resources. We believe this must change—and that digital identity should be trustworthy, private, and accessible to everyone. At VIDA, we're creating a frictionless digital identity system that meets the needs of our time and works anywhere, for everyone. Global initiatives like the UN and World Bank's ID4D aim to provide legal identity for all by 2030. We expect digital identity to become a basic right, and we want VIDA to help lead that future. We're a mission-driven, pragmatic team motivated by real-world impact—from preventing fraud and corruption to unlocking fair access to essential services. Phone Social Network and Web Links
... and detail-oriented Associate Data Engineer to join our growing data team. If you are a recent graduate passionate about data and eager to build a career in data engineering, this is the perfect opportunity for you! In this role, you will work closely with senior engineers to build, maintain, and optimize data pipelines. ...
... managing this platform, utilizing leading technologies like Databricks, Delta Lake, Spark, PySpark, Scala, and the AWS suite. We are actively seeking skilled Data Engineers to join our team and contribute to scaling our platform across the organization. Create and manage robust data ingestion pipelines leveraging Databricks, ...
... 3-6 years of overall experience in software development and data analysis. - Hands-on experience in Python for application development. - Strong proficiency in SQL for data analysis and query generation. - Experience with hosted environments and cloud platforms such as AWS, Azure, or similar (preferred). - Ability to work ...
... AWS Glue, Lambda, and DBT. Work with PySpark and SQL to transform and cleanse data. Configure and manage Airflow workflows and Amazon EMR clusters. Integrate data services with API Gateway and REST/Soap APIs. Required qualifications - 3 to 5 years of experience in data engineering. - Proficient in AWS S3, Glue, Lambda, ...
... teams including data analysts and engineers. Write reusable, efficient, and well-documented code. Integrate APIs and third-party services. Troubleshoot and debug data-related issues. Required Skills Strong proficiency in Python. Experience with data libraries like Pandas, NumPy. Hands-on experience with SQL and relational databases. ...
... data quality, reliability, and performance across data systems Required Skills: - Scala (must-have), python - Apache Spark - Strong coding fundamentals and engineering principles - Azure or on-premises Hadoop ecosystem experience - 5+ years of experience in Data Engineering - Python * SQL * AWS / Azure / GCP Preferred (Bonus) ...
... dbt, Apache Airflow, Fivetran, AWS, Git, and Looker Strong SQL development skills with experience in query optimization and performance tuning Experience in data ingestion techniques using custom pipelines or SaaS tools like Fivetran Expertise in data modeling and optimization of existing or new data models Experience ...
... Bachelor/Master's Degree with 4+ years of experience across Data Engineering (Data Pipelining, Warehousing, ETL Tools etc.) Extensive hands-on experience with data engineering techniques and Python & SQL Strong working knowledge of Snowflake, Airflow and dbt You are comfortable and have expertise in data engineering tooling ...
... Skills: Must To Have Skills: Proficiency in Python (Programming Language). Strong understanding of data structures and algorithms. Experience with ETL tools and data integration techniques. Familiarity with cloud platforms such as AWS or Azure. Knowledge of database management systems and SQL. Additional Information: The candidate ...
... and development Data quality validation and monitoring Schema design for analytics and reporting Performance optimization and scalability Preferred Experience Databricks Delta Lake experience Azure Synapse Analytics Python for data engineering Spark SQL optimization Real-time data streaming Data governance and metadata management ...
... Posted 2 weeks ago Job Be an early applicant About the job Job Purpose: Design and implement scalable data engineering and data warehouse solutions while managing data schemas, SQL query tuning, and code reviews. Who You Are: - 5+ years of experience in Data Engineering, with strong knowledge of Data Platforms and Data Warehousing ...
... Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering experience in Banking Risk, Retail Products, Cards, Mortgage, Deposits, or Wealth Management. Assumed Requirements / Certifications: Databricks Certified Data Engineer, AWS Certified Data Analytics, or Azure Data Engineer Associate.
... Python, SQL, and Spark/PySpark.- Hands-on experience with Azure Data Factory, Azure Databricks, and/or Synapse Analytics.- Experience with ETL/ELT frameworks and data integration tools.- Good understanding of data engineering concepts and data pipeline architecture.- Experience with data transformation, processing, and optimization.- ...
... experience with Python, SQL, and PySpark. Strong understanding of unit testing, test automation, debugging, and code quality best practices. Good understanding of data engineering fundamentals and advanced concepts. Experience building scalable and fault-tolerant data pipelines. Strong knowledge of CI/CD, automated testing, ...
... governance. Experience : Minimum 6 years of hands-on experience in data engineering, with a proven track record in complex pipeline development and cloud-based data migration projects. Bachelor’s or higher degree in Computer Science, Data Engineering, or a related field. Skills : Proficiency in Spark, SQL, Python, and other ...
... experiences. Contribute to AI-enabled features and use AI development tools responsibly to improve coding, testing, research, documentation, and delivery. Apply sound engineering practices across architecture, performance, security, testing, debugging, and maintainability. Deploy and operate services using AWS and CI/CD workflows, ...
... architects, analysts, developers, and business stakeholders to understand data requirements.- Ensure data quality, security, governance, and best practices across data engineering https://jobeax.com/link/Gpaa4oyRTNYXQU0V Stack:- Python, PySpark, SQL, AWS Glue, AWS Glue Data Catalog, AWS Lambda, Amazon S3, AWS Step Functions, ...
... BigQuery, Cloud Functions, Composer, GCS Proficient hands-on programming experience in Spark/Scala (python/java) Proficient in building production level ETL/ELT data pipelines from data ingestion to consumption Data Engineering knowledge (such as Data Lake, Data warehouse - Redshift/Hive/Snowflake, Integration, Migration) ...