Data Engineer - ML & AI in Bengaluru, India is listed on Jobeax. Browse 30,000+ vacancies available.
iMerit is a leading AI data solutions company specializing in transforming unstructured data into structured intelligence for advanced machine learning and analytics applications. Our clients span autonomous mobility, medical AI, agriculture, and more—powering next-generation AI systems with high-quality data services. We are seeking a skilled Data Engineer to help scale and enhance our internal data observability and analytics platform. This platform integrates with data annotation tools and ML pipelines to provide visibility, insights, and automation across large-scale data operations. You will design and optimize robust data pipelines, build integrations with internal platforms (e.g., Design and build scalable batch and real-time data pipelines across structured and unstructured sources. Integrate analytics and observability services with upstream annotation tools and downstream ML validation systems to enable full-cycle traceability. Collaborate with product, platform, and analytics teams to define event models, metrics, and data contracts. Develop ETL/ELT workflows using tools like AWS Glue, PySpark, or Airflow; ensure data quality, lineage, and reconciliation. annotation throughput, quality KPIs, latency). Build data models and queries to power dashboards and insights via tools like Athena, QuickSight, or Redash. Contribute to infrastructure-as-code and CI/CD practices for deployment across cloud environments (preferably AWS). Document architecture, data flow, and support runbooks; continuously improve platform performance and resilience. Integrate with customer data platforms and pipelines, including bespoke data frameworks.
4–8 years of experience in data engineering or backend development in data-intensive environments. Proficient in Python and SQL; Strong experience with cloud-native data tools and services (S3, Lambda, Glue, Kinesis, Firehose, RDS). Experience with data lake and warehouse patterns (e.g., Delta Lake, Redshift, Snowflake). Solid understanding of data modeling, schema design, and versioned datasets. Data Governance and Security: Understanding and implementing data governanc policies and security measures. Proven experience in building resilient, production-grade pipelines and troubleshooting live systems. Good working knowledge of Database fundamentals, relational databases and SQL
Experience with observability/monitoring systems (e.g., Familiarity with data governance, RBAC, PII redaction, or compliance in analytics platforms. Exposure to annotation/ML workflow tools or ML model validation platforms. Comfort working in Agile, distributed teams using tools like Git, JIRA, and Slack.
You'll work at the intersection of AI, data infrastructure, and impact—contributing to platforms that ensure AI is explainable, auditable, and ethical at scale. Join a team building the next generation of intelligent data operations.
... query optimization and data access strategies for APIs serving real-time compliance dashboards, establishing benchmarks and SLOs Lead data infrastructure work for AI/ML features, including dataset curation, feature engineering, and pipeline design supporting cloud-native AI capabilities Define and enforce data quality standards, ...
... Pipeline, ETL/ELT, and Data Warehousing. - Effective use of AI and/or LLMs to increase efficiency in deliverables. - Extensive experience working with various data sources (DB2, SQL,Oracle, flat files (csv, delimited), APIs, XML, JSON). - Experience implementing data integration techniques such as event/message based integration ...
... ground up with new age technology to simplify the consumption of data for our customers in various industry verticals. We are looking for a skilled and motivated Data Engineer to design, build, and maintain scalable data pipelines and data infrastructure. You will work closely with Data Science, Engineering, and Product teams ...
... address their data needs Utilize databases and tools including and not limited to, Postgres, Redshift, Airflow, and MongoDB to support our data ecosystem. Leverage AI frameworks and libraries to integrate advanced analytics into our solutions. Qualifications Experience : Minimum of 3 years of experience in data engineering, ...
... comfortable using LLMs/AI copilots as part of daily workflow (analysis, documentation, communication, process design), not just as a novelty. - Experience working with AI/ML teams and understanding of how annotated data feeds into model training. - Strong understanding of quality frameworks (QA/QC) for spatial data and annotation ...
... and backend engineering, test-driven development, microservices and architecture design principles - Must have experience on AI/ML implementation such as: Langchain, Langgraph and ML models such as Gradient Boost, Random Forest etc. - AI & LLM Frameworks: Experience with foundational models and APIs (e.g. OpenAI, Claude, ...
... scalable solutions Generate actionable insights through data analysis and modeling Deploy, monitor, and improve model performance Qualifications - Bachelor's/Master's in Computer Science, Data Science, Engineering, or related field - 3–8 years of relevant experience Skills: data scientist,genai,python,sql,pyspark,aiml,ml
... candidate will leverage generative AI and large language models (LLM) to enhance our data-driven decision-making processes. Key responsibilities : Design and develop data models using generative AI and LLM technologies. Implement retrieval-augmented generation techniques to improve data accessibility. Analyze complex datasets to ...
... End‑to‑End ML Project Lifecycle Python/R Programming Skills Software Engineering & Agile Framework Preferred Experience Prior experience with oil & gas, commercial domain, supply chain, production systems, wells or subsurface domain is highly desirable. Experience working with Azure Databricks or other data science frameworks. ...
... equivalent drawing standards is highly desirable. Education Bachelor's or Master's degree in Computer Science, AI/ML, Data Science, Electronics, Mechanical Engineering, Mathematics, Statistics or a related discipline. Master's degree is preferred. Candidate Profile - Hands-on senior engineer who can code, train/evaluate ...
... opportunity with the role currently based in Bengaluru, India. We use an AI-powered matching process to ensure your application is reviewed quickly, objectively, and fairly against the role's core requirements. Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to ...
Technology, Digital and Data Your Work Shapes the World at Caterpillar Inc. When you join Caterpillar, you're joining a global team who cares not just about the work we do – but also about each other. We are the makers, problem solvers, and future world builders who are creating stronger, more sustainable communities. We ...
... scalable AI/ML solutions and take models from experimentation through production. Requirements KEY RESPONSIBILITIES - Design, develop, test, and deploy scalable AI/ML models and solutions . - Work with large and complex datasets to perform data preparation, transformation, feature engineering, and analysis. - Develop and ...
... expertise in backend development, API integration, and scalable application design. The ideal candidate should have hands-on experience in Python frameworks, database management, cloud technologies, and software development best practices. Develop, test, and maintain scalable Python applications. Design and build RESTful ...
As an AIML Engineer, you will be responsible for designing, developing, and deploying AI-powered applications and agentic AI solutions leveraging Python, LangChain, LangGraph, LLMs, RAG, Oracle Database 23c/26ai, and Oracle APEX. You will work closely with cross-functional teams to build scalable AI solutions that integrate ...
As an AIML Engineer, you will be responsible for designing, developing, and deploying AI-powered applications and agentic AI solutions leveraging Python, LangChain, LangGraph, LLMs, RAG, Oracle Database 23c/26ai, and Oracle APEX. You will work closely with cross-functional teams to build scalable AI solutions that integrate ...
As an AIML Engineer, you will be responsible for designing, developing, and deploying AI-powered applications and agentic AI solutions leveraging Python, LangChain, LangGraph, LLMs, RAG, Oracle Database 23c/26ai, and Oracle APEX. You will work closely with cross-functional teams to build scalable AI solutions that integrate ...
Project Role : Data Engineer Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems. https://jobeax.com/link/6WcBICn0LM558SP0 ...
Role Summary : The Senior Data Scientist will lead advanced analytics initiatives, leveraging AI/ML, Generative AI (GenAI), and Agentic AI to deliver decision intelligence solutions. This role involves designing solutions, collaborating with cross-functional teams, and mentoring junior data https://jobeax.com/link/jOXlfQ8JM0UfgeL3 ...