Data Engineer (AI, db) in Bengaluru, India is listed on Jobeax. Browse 30,000+ vacancies available.
Antern Technologies is a product and engineering services company with over a decade of experience building scalable digital platforms and enterprise solutions. The company specializes in AI-driven products and custom technology solutions that help organizations modernize systems, improve efficiency, and achieve measurable business outcomes. Experience: 8–13 Year's
Candidates should be comfortable working 100% remotely.
Preference for professionals who are currently working from home or in a hybrid work environment and are comfortable transitioning to a new remote project.
Candidates should be adaptable and interested in contributing to a new project and evolving technical environment.
Job Description This is a full-time, remote Data Engineer role candidates based in Bengaluru, focused on COSMOS DB for experienced professionals with 8–13 years of relevant background. Candidates who are already comfortable working in a remote or hybrid environment and are looking for a new project opportunity are encouraged to apply.
The Data Engineer will design, build, and maintain scalable data pipelines and solutions, ensuring high performance and reliability across COSMOS DB and related data platforms. Daily responsibilities include modeling and structuring data, implementing ETL processes, optimizing data storage and retrieval, and collaborating with product, engineering, and analytics teams to support business reporting and AI-driven initiatives.
Bulk historical load from the systems of record (Aderant matters and parties via the published field maps): 300k-400k matters, 1-3M party rows.
the search-service performance work.
Incremental sync for same-business-day freshness (including the platform's own write-back and the ledger reconciliation).
Large-scale data / ETL engineering; SQL Server; performance tuning, Graph / vector stores, record linkage; graph and vector databases; SQL Server (incl. large-scale data engineering
Nice-to-have. Experience in Legal or party data at scale.
Experience. Data migration experience, Min 2 years of experience in Knowledge graph and Vector Data migration.
... portfolio companies across North America, the UK, EU, and https://jobeax.com/link/YIrO88Y2oCG2cGci :We are seeking a Senior AI & Data Engineer to join our Tech & Data Strategy team in Chennai. You excel in your understanding of AI: how large language models work, how AI agents reason and use tools, and how AI automation transforms ...
Role Overview:Seeking a highly skilled Senior Data Engineer with deep expertise in Databricks, Real-Time Data Processing, Lakehouse Architecture, and AI-driven solutions. The role involves designing, developing, and supporting scalable data platforms and streaming pipelines, leveraging modern Databricks capabilities such ...
... Generative AI solutions across business functions.- Analyze large datasets, derive actionable insights, and build scalable analytical workflows.- Design and implement AI-powered applications using LLMs, RAG frameworks, AI Agents, and Prompt Engineering techniques.- Collaborate with stakeholders to understand business requirements ...
... comfortable using LLMs/AI copilots as part of daily workflow (analysis, documentation, communication, process design), not just as a novelty. - Experience working with AI/ML teams and understanding of how annotated data feeds into model training. - Strong understanding of quality frameworks (QA/QC) for spatial data and annotation ...
... assistants)- Ensure responsible AI practices including model governance and explainability- Establish data governance framework (catalog, lineage, ownership)- Ensure data quality, consistency, and compliance- Work closely with business leaders to translate requirements into data solutions- Lead cross-functional teams (data engineers, ...
... of our mission-critical data infrastructure. In this role, you will serve as the primary custodian of our database ecosystem, working closely with software engineering teams, DevOps engineers, and product stakeholders to ensure seamless data availability. By architecting robust database solutions and optimizing complex ...
... Studio, Azure AI Foundry, Azure OpenAI, or AWS Bedrock.- Hands-on experience with AI orchestration frameworks such as Semantic Kernel, LangChain, LangGraph, LlamaIndex, AutoGen, or CrewAI.- Experience with relational, NoSQL, and vector databases, including SQL Server, PostgreSQL, MongoDB, Cosmos DB, and Azure AI Search or ...
... Studio, Azure AI Foundry, Azure OpenAI, or AWS Bedrock.- Hands-on experience with AI orchestration frameworks such as Semantic Kernel, LangChain, LangGraph, LlamaIndex, AutoGen, or CrewAI.- Experience with relational, NoSQL, and vector databases, including SQL Server, PostgreSQL, MongoDB, Cosmos DB, and Azure AI Search or ...
... Studio, Azure AI Foundry, Azure OpenAI, or AWS Bedrock.- Hands-on experience with AI orchestration frameworks such as Semantic Kernel, LangChain, LangGraph, LlamaIndex, AutoGen, or CrewAI.- Experience with relational, NoSQL, and vector databases, including SQL Server, PostgreSQL, MongoDB, Cosmos DB, and Azure AI Search or ...
Role Overview : Join a high-impact team building scalable AI systems that power conversational interfaces, speech recognition, and large-scale data transformation https://jobeax.com/link/Hn2b60846aifVrwk You'll Do : - Architect and operate end-to-end AI systems for conversational interfaces and speech recognition.- Drive ...
... Container Apps, API Management, and Azure Cosmos DB.- Proven experience in scaling applications on Azure to handle production-level workloads, ensuring high availability and performance.- Proven experience in designing and implementing secure applications, including identity management (Azure AD, OAuth, etc.) and data protection ...
... may be invited to take part in online projects such as rating AI-generated content, evaluating factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based ...
... Develop production-ready code and perform data analysis with Python and SQL Create workflows for model development and apply feature engineering methods Use Azure AI Search to make data and models easier to consume for business needs Coordinate with developers and project managers using GitLab and Jira Refine data pipelines ...
... statistics, ML classification, time-series modeling). - Integrate outputs with downstream systems (dashboards, APIs, ML models) to support digital agronomy and sustainability reporting. - Maintain data quality, document methodologies, and support junior geospatial analysts. - Promote best practices in model development and code ...
... may be invited to take part in online projects such as rating AI-generated content, evaluating factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based ...
... applications for accuracy, latency, and cost. Implement data preprocessing, feature engineering, and ML model training workflows. Work with structured and unstructured datasets to solve business problems. Collaborate with Product Managers, Software Engineers, and Subject Matter Experts to deliver AI-driven features. Monitor model ...
... may be invited to take part in online projects such as rating AI-generated content, evaluating factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based ...
... automation, machine learning or generative AI, rather than applying AI by default. Commitment to continuous learning in data, AI, cloud engineering and the maritime domain. Reliable, maintainable and well-tested data and AI workflows that improve product quality and team efficiency. Clear evidence that AI-assisted outputs are evaluated, ...
... generation (RAG) systems Develop and optimize prompts, evaluation frameworks, and guardrails for LLM-powered applications Engineer scalable data and ML pipelines in Databricks using PySpark, Delta Lake, and MLflow Deploy, monitor, and maintain models in production on Azure (Azure AI Foundry, Azure OpenAI, Azure Functions, AKS), ...