Data Engineer - AWS Platform in Pune, India is listed on Jobeax. Browse 30,000+ vacancies available.
About the Role:We are looking for an experienced AWS Data Engineer with 6 - 8 years of hands-on experience in designing, developing, and maintaining scalable data engineering solutions on AWS. The ideal candidate should have strong expertise in Python, PySpark, SQL, and AWS data services, with experience building robust ETL pipelines and Cloud-based data https://jobeax.com/link/JcPnJNS8HEL7JdHn Responsibilities:- Design, develop, and maintain scalable data pipelines and ETL workflows on AWS.- Develop efficient data processing solutions using python and pySpark.- Write complex and optimized SQL queries for data extraction, transformation, and analysis.- Build and manage ETL pipelines using AWS Glue.- Work with AWS Glue Data Catalog for metadata management and data discovery.- Design and implement serverless data processing solutions using AWS Lambda.- Use Amazon S3 for scalable data storage and data lake solutions.- Orchestrate and automate workflows using AWS Step Functions.- Implement event-driven architectures using Amazon EventBridge.- Use Amazon Athena for querying and analyzing data stored in Amazon S3.- Monitor, troubleshoot, and optimize data pipelines for performance, reliability, and scalability.- Collaborate with data architects, analysts, developers, and business stakeholders to understand data requirements.- Ensure data quality, security, governance, and best practices across data engineering https://jobeax.com/link/Gpaa4oyRTNYXQU0V Stack:- Python, PySpark, SQL, AWS Glue, AWS Glue Data Catalog, AWS Lambda, Amazon S3, AWS Step Functions, Amazon EventBridge, Amazon https://jobeax.com/link/TvfOgTReNTFVgNXo Skills:- Experience working with AWS Data Lake architectures.- Knowledge of data partitioning, file formats such as Parquet/Avro, and data optimization techniques.- Experience with CI/CD and DevOps practices for data pipelines.- Understanding of data security, IAM, monitoring, and governance on AWS.- Good problem-solving and communication https://jobeax.com/link/LGmb3dSeAFrVWm6i Requirements:- Notice Period: Immediate to 15 Days.- Experience: 6 - 8 Years.- Work Mode: Hybrid. (ref:hirist.tech)
... now. We are currently seeking a AWS Data & AI Platform Engineer to join our team in Hyderabad, Karnātaka (IN-KA), India (IN). We are seeking an experienced AI Platform Engineer to build and operate secure, scalable AI platforms on AWS for regulated financial services environments. This role focuses on enabling generative AI ...
... https://jobeax.com/link/uquJtnmzNQbyMgrx and Kubernetes are FactSet's primary containerized compute platforms, running thousands of apps with large-scale clustered compute resources in FactSet data centers and cloud hyperscalers. You will manage a team of Compute Platform Ops Engineers to provide first-tier support, review and implement component upgrades ...
... in DevOps practices, methodologies, and tools, such as CI/CD pipelines, configuration management, and containerization. Experience with cloud platforms (e.g., AWS, Azure, GCP) and infrastructure as code (e.g., AWS Experience along with AI/ML Basics is a plus Enthusiastically follow technology trends, software engineering ...
... and perspective to help EY become even better, too. Join us and build an exceptional experience for yourself, and a better working world for all. SIEM SOAR/Platform Engineer The ideal candidate will have extensive experience with Palo Alto Cortex XSOAR (formerly Demisto) and a strong background in security automation and ...
Company : Very big MNCRole : GCP Data EngineerExperience : 8 - 15 yrsNotice Period : 30 daysLocation : PAN INDIATech Stack :- GCP data engineer- Pyspark- Dataflow- Dataproc- BigQuery- AirflowKey Responsibilities :- Design, develop, and optimize endtoend data pipelines using PySpark, Dataflow, and Dataproc- Implement data ingestion, ...
Company : Very big MNCRole : GCP Data EngineerExperience : 8 - 15 yrsNotice Period : 30 daysLocation : PAN INDIATech Stack :- GCP data engineer- Pyspark- Dataflow- Dataproc- BigQuery- AirflowKey Responsibilities :- Design, develop, and optimize endtoend data pipelines using PySpark, Dataflow, and Dataproc- Implement data ingestion, ...
Role Overview : We are looking for a Principal Engineer to own the semantic model architecture at the heart of our data platform. This role sits at the intersection of data modeling, platform engineering, and product https://jobeax.com/link/1gXrtZ4TydNBW4pC You will Do : - Design and evolve the semantic modeling layer that ...
... Work Type: Full Time At VIDA we're building the future of digital identity. As a Data Engineer you'll work across the stack—from platform and infrastructure to data pipelines and end-user tooling. You'll help us modernize our batch ETLs and lead our push into streaming and real-time fraud detection. We run on AWS and make ...
... of overall experience in software development and data analysis. - Hands-on experience in Python for application development. - Strong proficiency in SQL for data analysis and query generation. - Experience with hosted environments and cloud platforms such as AWS, Azure, or similar (preferred). - Ability to work independently ...
... on-premises to cloud Spark code optimization and Medallion Architecture. Strong knowledge of Databricks and its components, including Delta Live Table (DLT) pipeline implementations. Familiarity with AWS services (experience with additional cloud platforms like GCP or Azure is a plus). AWS/GCP/Azure Data Engineer Certification.
... BigQuery, Cloud Functions, Composer, GCS Proficient hands-on programming experience in Spark/Scala (python/java) Proficient in building production level ETL/ELT data pipelines from data ingestion to consumption Data Engineering knowledge (such as Data Lake, Data warehouse - Redshift/Hive/Snowflake, Integration, Migration) ...
We are seeking a Data Engineer to join our team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL ...
... similar). Experience in building data pipelines or ETL processes. Familiarity with REST APIs and backend frameworks (Flask/Django is a plus). Understanding of data structures and algorithms. Good To Have Experience with big data tools (Spark, Hadoop). Knowledge of cloud platforms (AWS, Azure, or GCP). Exposure to machine ...
... understanding of Apache Airflow and ability to design and manage complex DAGs. - Solid SQL skills and familiarity with data warehouse platforms (e.g., Snowflake, Redshift, BigQuery). - Familiarity with version control tools (Git), CI/CD pipelines, and Agile methodologies. - Exposure to cloud environments like AWS, GCP, or Azure
We are looking for a Data Engineer to join our team and play a key role in building, optimizing, and managing data pipelines. You will work closely with Data Analysts and Business Functions to ensure seamless data processing, reporting, and dashboarding. Our tech stack includes BigQuery, Snowflake, Airflow, Stitch/Fivetran, ...
... take risks when looking for novel solutions to complex problems. Python for data pipelining and automation. Airbyte for ETL purpose Google Cloud Platform (GCP)/AWS, Snowflake, Terraform, Kubernetes, Cloud SQL, Cloud Functions, BigQuery, DataStore, and more: we keep adopting new tools as we grow! Airflow and dbt for data ...
... Skills: Must To Have Skills: Proficiency in Python (Programming Language). Strong understanding of data structures and algorithms. Experience with ETL tools and data integration techniques. Familiarity with cloud platforms such as AWS or Azure. Knowledge of database management systems and SQL. Additional Information: The candidate ...
... experience in data engineering. - Proficiency in: Programming languages - Python, Java, SQL, Spark SQL. - Data technologies - Hadoop, PySpark, NoSQL databases. - Data visualization tools - Qliksense, Tableau, Power BI - Cloud platorms - Azure Data Factory, Azure Databricks, AWS NovintiX is a fast-growing engineering and digital ...