Data Engineer (AI, db) in Pune, India is listed on Jobeax. Browse 30,000+ vacancies available.
Note: This job role is part of MetLife's Hack4Job India (a hiring hackathon). Only shortlisted candidates will be invited. Department: Global Technology
Role Overview MetLife is seeking an experienced Data Engineer to drive our digital and AI transformation journey. This role focuses on building modern data platforms, enhancing data storage and access, and ensuring seamless data consumption through APIs. The ideal candidate will work with Azure Cloud technologies to build robust data pipelines, data lakes, and marts to support business analysts and data scientists.
Key Responsibilities Modern Data Platform Development: Build data lake components on cloud-based platforms Design and develop data marts for business analysts and data scientists Data Engineering & Pipelines: Design data pipelines to integrate structured, semi-structured, and unstructured data from multiple sources Implement ETL/ELT processes to transform and cleanse data Ensure data quality and transformation rules align with Enterprise standards Work with Medallion architecture and implement best practices for data modeling Agile & DevOps Practices: Deliver solutions using Agile methodologies in a CI/CD-driven environment Work on containerized solutions (Azure Kubernetes) and scheduling tools like Azure Scheduler Follow secure coding practices and authentication/authorization protocols
Candidate Qualifications
Education: Bachelor's degree in computer science or equivalent
Experience: 4 - 8 years of experience in data engineering or data application development (ETL/ELT/BI)
2+ years of experience in cloud-based data platform development
Expertise in building Azure-based data pipelines, including:
Azure Data Factory / Synapse
DataBricks / Synapse Spark Pool
Cosmos DB
Azure Data Lake Storage (ADLS)
Dedicated SQL Pool / Azure SQL
Azure Logic Apps
Hands-on experience with data transformation and cleansing using Spark, Python, R, SQL
Strong understanding of CI/CD, test-driven development, and domain-driven design
Skills & Competencies Technical Expertise: Proficiency in Python, SQL, Spark, Azure Data Factory, and ETL processes Experience in secure coding, authentication, and monitoring tools like Veracode, MS Entra, PingOne Working knowledge of Azure Kubernetes, Azure DevOps, SonarQube, and Azure AppInsights Soft Skills: Strong communication and collaboration in a global, multi-cultural environments (experience in a Japanese work environment is a plus) ~ Able to work in a high-paced, diverse environment with a can-do attitude
Language : Business proficiency in English; Japanese language is a plus
This is a great opportunity to be part of MetLife's technology transformation journey.
Role Overview:Seeking a highly skilled Senior Data Engineer with deep expertise in Databricks, Real-Time Data Processing, Lakehouse Architecture, and AI-driven solutions. The role involves designing, developing, and supporting scalable data platforms and streaming pipelines, leveraging modern Databricks capabilities such ...
... Generative AI solutions across business functions.- Analyze large datasets, derive actionable insights, and build scalable analytical workflows.- Design and implement AI-powered applications using LLMs, RAG frameworks, AI Agents, and Prompt Engineering techniques.- Collaborate with stakeholders to understand business requirements ...
... comfortable using LLMs/AI copilots as part of daily workflow (analysis, documentation, communication, process design), not just as a novelty. - Experience working with AI/ML teams and understanding of how annotated data feeds into model training. - Strong understanding of quality frameworks (QA/QC) for spatial data and annotation ...
... assistants)- Ensure responsible AI practices including model governance and explainability- Establish data governance framework (catalog, lineage, ownership)- Ensure data quality, consistency, and compliance- Work closely with business leaders to translate requirements into data solutions- Lead cross-functional teams (data engineers, ...
... of our mission-critical data infrastructure. In this role, you will serve as the primary custodian of our database ecosystem, working closely with software engineering teams, DevOps engineers, and product stakeholders to ensure seamless data availability. By architecting robust database solutions and optimizing complex ...
... Studio, Azure AI Foundry, Azure OpenAI, or AWS Bedrock.- Hands-on experience with AI orchestration frameworks such as Semantic Kernel, LangChain, LangGraph, LlamaIndex, AutoGen, or CrewAI.- Experience with relational, NoSQL, and vector databases, including SQL Server, PostgreSQL, MongoDB, Cosmos DB, and Azure AI Search or ...
... Studio, Azure AI Foundry, Azure OpenAI, or AWS Bedrock.- Hands-on experience with AI orchestration frameworks such as Semantic Kernel, LangChain, LangGraph, LlamaIndex, AutoGen, or CrewAI.- Experience with relational, NoSQL, and vector databases, including SQL Server, PostgreSQL, MongoDB, Cosmos DB, and Azure AI Search or ...
... Studio, Azure AI Foundry, Azure OpenAI, or AWS Bedrock.- Hands-on experience with AI orchestration frameworks such as Semantic Kernel, LangChain, LangGraph, LlamaIndex, AutoGen, or CrewAI.- Experience with relational, NoSQL, and vector databases, including SQL Server, PostgreSQL, MongoDB, Cosmos DB, and Azure AI Search or ...
Role Overview : Join a high-impact team building scalable AI systems that power conversational interfaces, speech recognition, and large-scale data transformation https://jobeax.com/link/Hn2b60846aifVrwk You'll Do : - Architect and operate end-to-end AI systems for conversational interfaces and speech recognition.- Drive ...
... Container Apps, API Management, and Azure Cosmos DB.- Proven experience in scaling applications on Azure to handle production-level workloads, ensuring high availability and performance.- Proven experience in designing and implementing secure applications, including identity management (Azure AD, OAuth, etc.) and data protection ...
... may be invited to take part in online projects such as rating AI-generated content, evaluating factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based ...
... Develop production-ready code and perform data analysis with Python and SQL Create workflows for model development and apply feature engineering methods Use Azure AI Search to make data and models easier to consume for business needs Coordinate with developers and project managers using GitLab and Jira Refine data pipelines ...
... statistics, ML classification, time-series modeling). - Integrate outputs with downstream systems (dashboards, APIs, ML models) to support digital agronomy and sustainability reporting. - Maintain data quality, document methodologies, and support junior geospatial analysts. - Promote best practices in model development and code ...
... may be invited to take part in online projects such as rating AI-generated content, evaluating factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based ...
... applications for accuracy, latency, and cost. Implement data preprocessing, feature engineering, and ML model training workflows. Work with structured and unstructured datasets to solve business problems. Collaborate with Product Managers, Software Engineers, and Subject Matter Experts to deliver AI-driven features. Monitor model ...
... may be invited to take part in online projects such as rating AI-generated content, evaluating factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based ...
... automation, machine learning or generative AI, rather than applying AI by default. Commitment to continuous learning in data, AI, cloud engineering and the maritime domain. Reliable, maintainable and well-tested data and AI workflows that improve product quality and team efficiency. Clear evidence that AI-assisted outputs are evaluated, ...
... generation (RAG) systems Develop and optimize prompts, evaluation frameworks, and guardrails for LLM-powered applications Engineer scalable data and ML pipelines in Databricks using PySpark, Delta Lake, and MLflow Deploy, monitor, and maintain models in production on Azure (Azure AI Foundry, Azure OpenAI, Azure Functions, AKS), ...
... Strategists to translate model outputs into actionable credit limits, cutoff thresholds, and loss-forecasting simulations. - Platform Architecture: Collaborate with Data Platform Engineers to build reliable pipelines and high-throughput feature extraction. - Raise the technical bar through peer reviews, reusable tooling, and pro-active ...