Data Scientist (Python, ML, SQL) in Noida, India - Jobeax
Vacancy description
Data Scientist (Python, ML, SQL) in Noida, India
Exl
Rs 5 - 10 lakhs p.a.
India, Noida
Data Scientist (Python, ML, SQL) in Noida, India is listed on Jobeax. Browse 30,000+ vacancies available.
Model Development: Research, design, and develop state-of-the-art generative models such as GPT, GANs, VAEs, or diffusion models for tasks like text generation, summarization, reasoning, Q&A, or predictive analytics.
Data Preparation: Clean, preprocess, and structure large datasets to train, validate, and test generative AI models.
Deployment: Implement scalable solutions for deploying generative AI models in production environments using tools like Docker, Kubernetes, or cloud platforms (AWS, GCP, Azure).
Collaboration : Work closely with cross-functional teams, including product managers, engineers, and stakeholders, to align AI solutions with business goals.
Evaluation: Develop metrics and benchmarks to evaluate model performance and ensure quality outputs.
Research: Stay up to date with the latest advancements in AI/ML, particularly in generative AI techniques, and propose innovative solutions.
Ethical AI Practices: Experience of architecting AI systems to solve complex business problems.
Build advanced RAG pipelines, text chunking, and retrieval, LLM Prompt Engineering, using Vector Databases. Implement right LLM selection based on use cases and client criteria (GPT-4, Llama2, Mistral, Claude, Gemini, Flan, BERT) while managing trifecta of accuracy, cost, and latency/scale.
Develop, fine-tune, context tune and implement state-of-the-art NLP models including Large Language Models like GPT, https://jobeax.com/link/10NB10tyPdNbii5N AI systems & architectures, while considering good Responsible AI standards and AI Governance.
Build Agentic AI, AutoGen, Muti-agents use cases, AI Autonomous Agents and LLM orchestration architectures for enterprise data at scale in Python.
Hands on experience with complementary technologies around LLMs like embedders, vector databases (chroma, weaviate etc.), Collaborate with cross-functional teams to integrate generative AI solutions into real-world applications.
Stay up-to-date with the latest advancements in deep learning and generative models and apply them to enhance our AI capabilities.
Document research findings, prepare technical reports, and contribute to whitepaper/scientific publications.
Provide deep leadership and coaching in the project delivery lifecycle. Qualifications Education and Experience:
Masters of Science or PhD in computer science, data science, statistics, Natural Language Processing
2-3 years of experience in GenAI and Large Language Models.
Strong programming skills in Python, including experience with libraries like TensorFlow, PyTorch, Hugging Face Transformers, or similar. Solid understanding of optimization techniques for training deep neural networks, regularization methods, and hyperparameter/fine tuning.
Hands-on experience with generative AI models, such as Llama3, GPT4, Claude
Knowledge of NLP, computer vision, or multimodal AI techniques.
Proficiency in data manipulation and analysis tools (e.g., Pandas, NumPy, SQL).
Experience in Generative AI Models and LLMs, finetuning LLMs, prompt engineering and experience with LLM orchestration frameworks like Langchain, LlamaIndex, RAGAS, etc.
Strong software engineering skills for rapid and accurate development of AI models and systems.
Understanding of Agentic AI, Autonomous Agents, AI Agents, AutoGen, https://jobeax.com/link/x5OJTgUk5BytQUwi, Langchain, and workflow steps design. Google AI Agents.
Provide business-oriented solution with ability to communicate effectively, both verbally and in writing, with technical and non-technical stakeholders.
Experience working in a collaborative environment, contributing to multidisciplinary teams and projects.
Experience in deploying ML models in cloud environments (AWS SageMaker, GCP AI Platform, or Azure ML).
Excellent communication skills, both technical and non-technical.
5 years of experience in AI, NLP including transformer architecture and LLMs, Computer Vision and related technologies
Ability to explain GenAI to non-technical audiences across many different industries. Candidate is also hands on developer while people managing Data Scientists, ML Engineers etc.
Experience in ML Engineering and MLOps, MLFlow
Strong understanding of statistical and machine learning concepts
Experience with deep learning frameworks such as TensorFlow and PyTorch
Familiarity with key concepts and techniques used in generative models, such as variational autoencoders (VAEs), generative adversarial networks (GANs), and flow-based models.
Strong programming skills in languages such as Python, along with experience working with popular deep learning frameworks like PyTorch and TensorFlow.
Understanding of Graph Database and/or Vector Database along with knowledge of cloud services (e.g., AWS, Azure, GCP).
Experience with deploying AI models in production environments.
Familiarity with domain-specific applications of generative AI
Leveraged both Azure and AWS for model inferencing. For finetuning, he has worked more on AWS SageMaker.
... improve platform performance and resilience. Integrate with customer data platforms and pipelines, including bespoke data frameworks. 4–8 years of experience in data engineering or backend development in data-intensive environments. Proficient in Python and SQL; Strong experience with cloud-native data tools and services (S3, ...
... technical authority on internal data contracts and API schemas in cross-functional discussions with backend, product, and AI/ML teams 5+ years of experience in data engineering, with a proven track record delivering production-grade pipelines, data models, and storage architectures - Deep SQL expertise; hands-on experience ...
... delimited), APIs, XML, JSON). - Experience implementing data integration techniques such as event/message based integration (Kafka, Google Pub/Sub). - Advanced SQL skills; solid understanding of relational databases and business data; ability to write complex SQL queries against a variety of data sources. - Strong understanding ...
Principal Data Scientist - 7+ Years - BangaloreWe are hiring a Principal - Data Scientist with GenAI for a leading global AI and analytics organization. The role offers an opportunity to work on complex business problems and develop, implement and deploy Machine Learning and Generative AI solutions across https://jobeax.com/link/iWvc74qOaz1N3AKd ...
... in driving value for our customers by building data solutions. You'll be carrying out data engineering tasks to build, maintain, test and optimise a scalable data architecture, as well as carrying out data extractions, transforming data to make it usable to data analysts and scientists, and loading data into data platforms. ...
... https://jobeax.com/link/fYGlpInxuIv5PeS1 & Skills :- 4 - 7 years in data engineering: SQL, Python, ETL/ELT orchestration.- Cloud data platform experience (Azure preferred): pipelines, storage, APIs.- Strong understanding of data lake and data architecture.- Data modelling for analytics (star schema) and data quality frameworks. (ref:hirist.tech)
... Requirements : - 3+ years of professional AI/ML engineering experience with a track record of delivering production-grade AI systems.- Strong programming skills in Python and SQL, with hands-on experience in ML libraries (scikit-learn, pandas, numpy) and deep-learning frameworks (PyTorch or TensorFlow).- Hands-on experience building ...
... Requirements : - 3+ years of professional AI/ML engineering experience with a track record of delivering production-grade AI systems.- Strong programming skills in Python and SQL, with hands-on experience in ML libraries (scikit-learn, pandas, numpy) and deep-learning frameworks (PyTorch or TensorFlow).- Hands-on experience building ...
Senior Data Scientist – Agentic AI / Gen AI Job Type: Full-time Experience Required: 6–11 Years Location: Chennai (Work From Office – Pallavaram) Job Summary We are looking for a highly skilled and experienced Senior Computer Vision & Data Scientist with strong expertise in Computer Vision, Generative AI, Agentic AI systems, ...
... Bachelor's degree (preferably in Computer Science, Mathematics, Data Science, or a related field). - 3-6 years of overall experience in software development and data analysis. - Hands-on experience in Python for application development. - Strong proficiency in SQL for data analysis and query generation. - Experience with hosted ...
... Computer Science, IT, Data Science, Mathematics, or a related field (Recent graduates from batch [Year] are welcome). Python: Solid foundational knowledge of Python programming (data structures, loops, functions, basic libraries like Pandas). PostgreSQL (SQL): Basic understanding of relational databases, table design, and ...
... ability to design, build, and operate Data Lakes or Data Warehouses. Proficiency with Data Orchestration tools (Airflow, Dagster, Prefect). Familiarity with Change Data Capture tools (Canal, Debezium, Maxwell). Strong command of at least one primary language (Scala, Python, etc.) and SQL. Experience with data catalog and metadata ...
... and development Data quality validation and monitoring Schema design for analytics and reporting Performance optimization and scalability Preferred Experience Databricks Delta Lake experience Azure Synapse Analytics Python for data engineering Spark SQL optimization Real-time data streaming Data governance and metadata management ...
... and lead our push into streaming and real-time fraud detection. We run on AWS and make extensive use of Scala Spark, dbt, and Terraform. Build and expose the data lake/lakehouse so teams can understand business and product performance and make better decisions. Design and evolve a data platform that enables scientists and ...
... team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL to transform and cleanse data. Configure and manage ...
... inquiries related to Addepar's portfolio data feeds and on general data product functionalities within established SLAs. Manage and complete requests from internal Data teams that require client outreach and/or action to resolve data verification issues. Investigate client reported bugs and data processing issues, and triage ...
... Warehousing - 2+ years of hands-on experience with SSAS Tabular Models - Experience designing data ingestion and orchestration pipelines using Kafka, Snowflake, and Python - Hands-on experience with DBT for data modeling and pipeline development - Strong knowledge of SQL, ETL/ELT, Azure Data Factory, APIs, Azure Functions, and ...
... Troubleshoot and debug data-related issues. Required Skills Strong proficiency in Python. Experience with data libraries like Pandas, NumPy. Hands-on experience with SQL and relational databases. Knowledge of data visualization tools (Matplotlib, Seaborn, or similar). Experience in building data pipelines or ETL processes. Familiarity ...
... demonstrate technical and non-technical information, ideas, procedures, and processes Prior experience in an SAP and/or Informatica driven environment Ability to use SQL to analyze data Bachelor's degree related to Information Systems, Business or other relevant academic discipline with a minimum 4 years job experience in IT or ...