Data Scientist (Python, ML, SQL) in Noida, India - Jobeax
Vacancy description
Data Scientist (Python, ML, SQL) in Noida, India
Exl
Rs 5 - 10 lakhs p.a.
India, Noida
Data Scientist (Python, ML, SQL) in Noida, India is listed on Jobeax. Browse 30,000+ vacancies available.
Model Development: Research, design, and develop state-of-the-art generative models such as GPT, GANs, VAEs, or diffusion models for tasks like text generation, summarization, reasoning, Q&A, or predictive analytics.
Data Preparation: Clean, preprocess, and structure large datasets to train, validate, and test generative AI models.
Deployment: Implement scalable solutions for deploying generative AI models in production environments using tools like Docker, Kubernetes, or cloud platforms (AWS, GCP, Azure).
Collaboration : Work closely with cross-functional teams, including product managers, engineers, and stakeholders, to align AI solutions with business goals.
Evaluation: Develop metrics and benchmarks to evaluate model performance and ensure quality outputs.
Research: Stay up to date with the latest advancements in AI/ML, particularly in generative AI techniques, and propose innovative solutions.
Ethical AI Practices: Experience of architecting AI systems to solve complex business problems.
Build advanced RAG pipelines, text chunking, and retrieval, LLM Prompt Engineering, using Vector Databases. Implement right LLM selection based on use cases and client criteria (GPT-4, Llama2, Mistral, Claude, Gemini, Flan, BERT) while managing trifecta of accuracy, cost, and latency/scale.
Develop, fine-tune, context tune and implement state-of-the-art NLP models including Large Language Models like GPT, https://jobeax.com/link/10NB10tyPdNbii5N AI systems & architectures, while considering good Responsible AI standards and AI Governance.
Build Agentic AI, AutoGen, Muti-agents use cases, AI Autonomous Agents and LLM orchestration architectures for enterprise data at scale in Python.
Hands on experience with complementary technologies around LLMs like embedders, vector databases (chroma, weaviate etc.), Collaborate with cross-functional teams to integrate generative AI solutions into real-world applications.
Stay up-to-date with the latest advancements in deep learning and generative models and apply them to enhance our AI capabilities.
Document research findings, prepare technical reports, and contribute to whitepaper/scientific publications.
Provide deep leadership and coaching in the project delivery lifecycle. Qualifications Education and Experience:
Masters of Science or PhD in computer science, data science, statistics, Natural Language Processing
2-3 years of experience in GenAI and Large Language Models.
Strong programming skills in Python, including experience with libraries like TensorFlow, PyTorch, Hugging Face Transformers, or similar. Solid understanding of optimization techniques for training deep neural networks, regularization methods, and hyperparameter/fine tuning.
Hands-on experience with generative AI models, such as Llama3, GPT4, Claude
Knowledge of NLP, computer vision, or multimodal AI techniques.
Proficiency in data manipulation and analysis tools (e.g., Pandas, NumPy, SQL).
Experience in Generative AI Models and LLMs, finetuning LLMs, prompt engineering and experience with LLM orchestration frameworks like Langchain, LlamaIndex, RAGAS, etc.
Strong software engineering skills for rapid and accurate development of AI models and systems.
Understanding of Agentic AI, Autonomous Agents, AI Agents, AutoGen, https://jobeax.com/link/x5OJTgUk5BytQUwi, Langchain, and workflow steps design. Google AI Agents.
Provide business-oriented solution with ability to communicate effectively, both verbally and in writing, with technical and non-technical stakeholders.
Experience working in a collaborative environment, contributing to multidisciplinary teams and projects.
Experience in deploying ML models in cloud environments (AWS SageMaker, GCP AI Platform, or Azure ML).
Excellent communication skills, both technical and non-technical.
5 years of experience in AI, NLP including transformer architecture and LLMs, Computer Vision and related technologies
Ability to explain GenAI to non-technical audiences across many different industries. Candidate is also hands on developer while people managing Data Scientists, ML Engineers etc.
Experience in ML Engineering and MLOps, MLFlow
Strong understanding of statistical and machine learning concepts
Experience with deep learning frameworks such as TensorFlow and PyTorch
Familiarity with key concepts and techniques used in generative models, such as variational autoencoders (VAEs), generative adversarial networks (GANs), and flow-based models.
Strong programming skills in languages such as Python, along with experience working with popular deep learning frameworks like PyTorch and TensorFlow.
Understanding of Graph Database and/or Vector Database along with knowledge of cloud services (e.g., AWS, Azure, GCP).
Experience with deploying AI models in production environments.
Familiarity with domain-specific applications of generative AI
Leveraged both Azure and AWS for model inferencing. For finetuning, he has worked more on AWS SageMaker.
Role Overview :The Senior Data Scientist, Product Analytics is a hands-on technical contributor and task manager within a cross-functional product team. This role sits at the intersection of deep technical execution and emerging AI https://jobeax.com/link/f9LyLs7glaUomOJP And Qualifications :- 5 - 7 years of experience ...
Role Overview : As a commercially savvy Lead Data Scientist, you will lead your team on client briefs whilst collaborating closely with consultants and Trade Partner stakeholders. You'll bring fresh ideas and innovative thinking to push the boundaries of what's possible with consumer analytics. You will drive AI excellence ...
... https://jobeax.com/link/6WcBICn0LM558SP0 (including Hooks, state management solutions like Redux/Zustand, and modern UI libraries). Databases: Strong hands-on experience with both SQL (e.g., PostgreSQL, MySQL) and NoSQL databases (e.g., MongoDB, DynamoDB, Redis). AI / ML Expertise: Hands-on experience integrating AI/ML libraries and models ...
... platform ecosystem including Workbench, Connect, and integration patterns with R and Python environments Proficient in distributed computing technologies and big data analytics, including hands-on expertise with Python, Spark, SQL, and data transformation pipelines Strong understanding of MLOps/ModelOps principles, practices, ...
... AI/ML-focused solutions. 3+ years of direct team management experience, leading Data Scientists, ML Engineers, or similar technical roles. - Hands-on experience in Python with strong proficiency in libraries such as Pandas, NumPy, SciPy, Matplotlib, and regular expressions. Advanced SQL querying and optimization for large datasets. ...
... Large Language Models (LLMs). - Expert-level proficiency in Python, SQL and its data science libraries (e.g., Proven track record of leading complex, end-to-end data science projects that have delivered significant business impact. Experience with cloud-based ML platforms / ML ops (e.g., AWS SageMaker, MLflow) and their generative ...
... AI/ML-focused solutions. 3+ years of direct team management experience, leading Data Scientists, ML Engineers, or similar technical roles. - Hands-on experience in Python with strong proficiency in libraries such as Pandas, NumPy, SciPy, Matplotlib, and regular expressions. Advanced SQL querying and optimization for large datasets. ...
... YOLO or other object-detection models - Dataset annotation tools - REST APIs and Postman - pytest or other automated testing frameworks - Linux/Ubuntu - Docker - SQL/databases - Simulation environments - Robotics, drones, autonomous systems, or sensor data - MAVLink / ROS / ROS 2 You do not need to know everything listed above ...
... in:- Data Operations- Production Data Engineering- Data Platform Support- Data Pipeline Monitoring & Troubleshooting- Not looking for purely development-focused Data Engineers or DevOps/SRE https://jobeax.com/link/V6ZwWogXLIs5s3d8 Stack:Snowflake, dbt, Apache Airflow, Fivetran, Cloud Storage, SQL, Python, Bash, CI/CD, Git, ...
... in:- Data Operations- Production Data Engineering- Data Platform Support- Data Pipeline Monitoring & Troubleshooting- Not looking for purely development-focused Data Engineers or DevOps/SRE https://jobeax.com/link/V6ZwWogXLIs5s3d8 Stack:Snowflake, dbt, Apache Airflow, Fivetran, Cloud Storage, SQL, Python, Bash, CI/CD, Git, ...
... expertise in Python and SQL, along with experience in cloud platforms (AWS, GCP, or Azure) and containerization (Docker, Kubernetes) - Experience with real-time data processing, anomaly detection, and time-series forecasting in production - Experience working with large datasets and big data technologies like Spark and Kafka ...
... end-to-end data architecture on Azure covering ingestion, storage, transformation, governance, and consumption layers.- Define reference architectures using Azure Data Lake, Azure Synapse Analytics, Microsoft Fabric, Azure Databricks, and Azure SQL/Cosmos DB.- Architect AI/ML and Generative AI solutions using Azure Machine Learning, ...
... CI/CD tools and supporting ML productionization while mentoring fellow team members. 5+ years of experience in data engineering with significant hands-on work in Databricks - Strong proficiency in PySpark, Spark SQL, Delta Lake, and Delta Live Tables - Advanced skills in Python and SQL - Experience with Unity Catalog, cluster ...
... Strong experience with SQL and working with APIs is required. Familiarity with data warehousing concepts and tools such as Azure Data Factory or equivalent. Data Handling: Expertise in data engineering, including data cleaning, transformation, and preparation for analysis. Knowledge of data modelling, database design, ...
... transformational change initiatives, including data governance, data quality, or analytics transformation https://jobeax.com/link/LsX0MHwADPrlj8n9 technical skills in data profiling, analysis, and data management using modern tools and environments (Python, R, SQL, Spark, cloud platforms) Understanding of data lineage concepts and ...
... calculated tables/columns, and commonly used DAX functions. Ability to assess existing Tableau dashboards and translate their business logic, calculations, data models, and visualizations into Power BI. Proficiency in SQL for data extraction, transformation, validation, and analysis. - Ability to perform QC checks and ...
... proficiency in SQL and Python for data processing and pipeline development. Data Platforms: Hands-on experience with Snowflake, dbt, and modern cloud data warehouses. Data Modelling: Deep experience with dbt and modern data modelling practices. Data Platforms: Hands-on experience working with Snowflake or comparable cloud data warehouses ...
... designs through production operations. - Support the ongoing maintenance and operations of the data platform. Net / C# - Python - AWS or similar - Snowflake ~ SQL - Experience - 8+ years of software engineering experience with a strong data focus, including designing, building, and operating production data systems in fast-paced ...
... expectations are met through disciplined implementation (access controls, lineage, and quality standards) 5+ years building and operating production data pipelines (data engineering—not primarily BI/analytics or data science) - Strong SQL plus strong data modeling skills (dimensional and/or lakehouse modeling) - Hands-on Databricks ...
Role Overview : We are seeking a seasoned Data/SQL Business Analyst to join our high-growth team in Delhi NCR/Gurgaon. In this role, you will serve as the bridge between complex raw data and actionable business strategy, translating technical insights into intuitive dashboards that drive decision-making for senior leadership. ...