Data Engineer - ML & AI in Hyderabad, India - Jobeax
Vacancy description
Data Engineer - ML & AI in Hyderabad, India
Costco IT
India, Hyderabad
Data Engineer - ML & AI in Hyderabad, India is listed on Jobeax. Browse 30,000+ vacancies available.
About Costco Wholesale: Costco Wholesale is a multi-billion-dollar global retailer with warehouse club operations in eleven countries. They provide a wide selection of quality merchandise, plus the convenience of specialty departments and exclusive member services, all designed to make shopping a pleasurable experience for their members.
At Costco Wholesale India, we foster a collaborative space, working to support Costco Wholesale in developing innovative solutions that improve members’ experiences and make employees’ jobs easier. Our employees play a key role in driving and delivering innovation to establish IT as a core competitive advantage for Costco Wholesale.
Data Engineer
The Data Engineer is responsible for developing data pipelines and/or data integrations of for Costco’s enterprise certified data sets that are used for business critical data consumption use cases (i.e. Reporting, Data Science/Machine Learning, Data APIs, etc.). The Data Engineer will partner with product owners, data architects, and data platform teams to design, build, test, and automate data pipelines that are relied upon across the company as the single source of truth.
Develops and operationalizes data pipelines to bring data into Costco’s GCP landscape for the delivery of certified data sets
Works in tandem with Data Architects, Data Stewards, and Data Quality Engineers to design data pipelines and recommends ongoing optimization of data storage, data ingestion, data quality and orchestration.
Designs, develops, and implements ETL/ELT/CDC processes using Data Build Tool (DBT) and other native GCP Services (BigQuery Subscriptions, Dataproc, Dataflow, etc.)
Uses GCP services such as BigQuery, AlloyDB, Spanner, Dataplex, Pub/Sub, Cloud Storage, etc. to improve and speed delivery of our data products and services.
Identifies, designs, and implements internal process improvements: automating manual processes, optimizing pipeline delivery and support.
Identifies ways to improve data reliability, efficiency, and quality of data management
Participate in off hours 24/7 on call support on a rotational basis
5-8+ years’ experience engineering and operationalizing data pipelines with large and complex datasets.
5-8+ years’ hands-on experience with Data Build Tool (DBT), Dataflow & Dataproc (Spark)
5-8+ years’ experience working with BigQuery, Google Cloud Storage, AlloyDB and Spanner
5-8+ years’ experience with Data Pipeline, ETL/ELT, and Data Warehousing.
Effective use of AI and/or LLMs to increase efficiency in deliverables.
Extensive experience working with various data sources (DB2, SQL,Oracle, flat files (csv, delimited), APIs, XML, JSON).
Experience implementing data integration techniques such as event/message based integration (Kafka, Google Pub/Sub).
Advanced SQL skills; solid understanding of relational databases and business data; ability to write complex SQL queries against a variety of data sources.
Strong understanding of database storage concepts (data lake, relational databases, NoSQL, Graph, data warehousing).
Experience with Git / Azure DevOps / Jira.
CDMP (Certified Data Management Professional) Certification.
Role Summary : The Senior Data Scientist will lead advanced analytics initiatives, leveraging AI/ML, Generative AI (GenAI), and Agentic AI to deliver decision intelligence solutions. This role involves designing solutions, collaborating with cross-functional teams, and mentoring junior data https://jobeax.com/link/jOXlfQ8JM0UfgeL3 ...
... Azure, GCP) and MLOps tools. - Familiarity with vector databases, RAG pipelines, and LangChain or similar frameworks. - Proven track record of delivering Gen AI solutions in production. - Master's in Computer Science, AI/ML, or related field. Experience in domain-specific Gen AI applications (e.g., Experience - 5 + Years ...
... expertise in NLP, Fundamental machine learning, deep learning, transformer, state space-based architecture.- Azure ML and/or AWS.- Strong in Python coding, SQL and database queries, data preparation, and analysis.- Exploratory Data Analysis (EDA).- Experience with PyTorch.- LLM training and fine-tuning (e.g., GPT, LLaMA, Mistral, ...
... Extensive experience in building, consuming, and optimizing RESTful APIs, with proficiency in tools like Swagger, Postman, or similar. - Strong knowledge of SQL databases and querying languages - Demonstrated experience in building and maintaining robust CI/CD pipelines using tools such as Jenkins or GitLab CI. - Exceptional ...
... the Fortune 50, achieve discoveries, insights, and business outcomes faster and more sustainably. We're passionate about solving our customers' most complex data challenges to accelerate intelligent innovation and business value. As a Senior/Staff Kernel Engineer at Weka, your primary responsibility will be collaborating ...
... Celery or other queue mechanism is Must Good to Have: - Experience working with Docker and cloud platforms (AWS/GCP/Azure) is good to have - Familiarity with ML model serving is Nice to have. - Exposure to GenAI (e.g., llama, OpenAI, HuggingFace, LangChain, Agentic AI etc) or Data Engineering (e.g., Data platforms, pandas, ...
Shift Timings: Design, build, and maintain robust ETL/ELT pipelines feeding a Snowflake-based data platform Build and manage integrations using SnapLogic to connect source systems, APIs, and downstream consumers Develop and maintain data models and transformations in dbt, including tests, documentation, and CI/CD-based ...
Job Title: Senior Data Engineer Eperience - 10+ to 18 yrs Location: Remote Type: Contract( Comfortable with a 6-month contractual role) Requires strong hands-E xperience in AI and Data Engineering with strong expertise in Azure OpenAI, Azure AI Foundry, Agentic AI, RAG, LLMs, AutoGen, OCR, Databricks, PySpark, Python, Azure ...
... performant Use AI tooling within DE workflows, including code generation, pipeline automation, and data quality checks Contribute to and help shape company-wide data governance standards Collaborate with analytics, BI, and business teams to deliver trusted, well-modeled data 5–7 years of hands-on data engineering experience ...
... Engineering, or related field. - 8+ years of experience in Devops and 3+ years in DBX. - Strong hands-on experience in Databricks (Spark, Delta Lake, PySpark, MLflow). - Proficiency in SQL and programming languages like Python or Scala. - Experience with Azure cloud data services. - Solid understanding of data modeling, ...
... exceptional candidates from other institutions with demonstrable production AI/ML experience will be considered. Work Experience: 2–4 years of total experience in data/AI engineering, with a minimum of 1 years of hands-on MLOps or LLMOps experience in a production environment. Demonstrated experience deploying and monitoring ML ...
... manage schema evolution, and maintain high data availability . Implement data governance , version control, and CI/CD best practices. Monitor and troubleshoot data pipelines for continuous reliability and efficiency improvements . Why Join KANINI? Join KANINI’s award-winning Data Engineering Team, recognized as the "Outstanding ...
... Experience with NLP techniques and working with LLMs (e.g., Familiarity with prompt engineering, fine-tuning, and model deployment on Azure. - Experience with vector databases (e.g., Azure AI Search, FAISS). MLflow, Azure DevOps). Experience with data visualization tools (e.g., Power BI). Certifications in Azure AI or Data Science. ...
... years of experience in AI/ML, including model development, data preprocessing, EDA, training, and evaluation. - 2+ years of hands‑on experience in Generative AI (LLMs, embeddings, RAG, LLM‑based apps). - 6+ months of hands‑on experience with Agentic AI frameworks (CrewAI / AutoGen / LangGraph / LangChain Agents). - Strong ...
... and responsible AI practices Collaborate with product, data, and platform teams to deliver business solutions Bachelor's or master's degree in computer science, AI, Data Science, or related field Experience 3–5 years of experience in software/ML engineering At least 1–2 years of hands-on experience in Generative AI / LLM-based ...
Role –Gen AI Engineer Location: PAN India Exp: 5+ years Mode Of Interview - F2F Job Description Collect and prepare data for training and evaluating multimodal foundation models. This may involve cleaning and processing text data or creating synthetic data. Develop and optimize large-scale language models like GANs (Generative ...
... Glue, SageMaker), Azure, or GCP. - Strong background in data engineering: ETL/ELT pipelines, data modeling, SQL, and data warehouse technologies (Snowflake, Databricks preferred). - Experience with REST APIs, GraphQL, and enterprise system integrations (CRM, ERP, data marketplaces). - Familiarity with containerization ...
... probability, data mining and data cleaning . - Strong experience with SQL and Python ML packages such as Scikit-learn, XGBoost and Keras. - Experience in building data pipelines and working with at least one cloud platform . - Good knowledge of Git, GitLab and CI/CD . - Familiarity with AI ethics and Responsible AI practices ...
... and implement effective solutions.- Participate in design discussions, code reviews, technical documentation, and architecture decisions.- Work closely with Data Scientists, AI/ML Engineers, Architects, QA, DevOps, and Product teams.- Follow Agile/Scrum practices and contribute across the complete software development ...
... practical Java experience - 2 years of experience with prompt engineering and prompt/agent scripting - 2 years of hands-on experience integrating third-party AI/ML services and platform APIs (e.g., OpenAI, Anthropic, Salesforce and other SaaS) - Proven ability to design and manage domain-specific languages, prompt templates, ...