Senior Data Scientist - Big Data in Rāniganj, India - Jobeax
Vacancy description
Senior Data Scientist - Big Data in Rāniganj, India
hireflex
India, Rāniganj
Senior Data Scientist - Big Data in Rāniganj, India is listed on Jobeax. Browse 30,000+ vacancies available.
We are seeking a highly skilled Senior Data Scientist to lead advanced analytics projects and deliver end-to-end data solutions. The ideal candidate will have extensive experience turning complex datasets into actionable insights through robust statistical and machine learning models. You will work closely with cross-functional teams (data engineers, analysts, product managers, etc.) to align data science initiatives with business objectives and deliver meaningful, data-driven insights. This role requires strong business acumen and exceptional technical skills, especially in Python programming and cloud-based deployment. Key Responsibilities
End-to-End Model Development: Design, implement, and maintain full-stack data science workflows. This includes data ingestion (ETL/ELT), rigorous data preprocessing and feature engineering, model training, validation, and deployment into production. You will refactor and harden data preprocessing pipelines to improve stability and robustness, ensuring reliable data quality and consistency over time.
Advanced Analytics & Machine Learning: Develop and optimize predictive models (regression, classification, forecasting, etc.) using state-of-the-art techniques. Continuously improve model accuracy by implementing advanced evaluation metrics such as weighted R², adjusted R², correlation coefficients, RMSE, and MAE. Enhance feature selection processes and hyperparameter optimization routines to extract maximum value from data.
Model Stability & Validation : Perform comprehensive model stability testing and validation. Use fixed random seeds for reproducibility, bootstrap resampling to assess data variability, and temporal hold-out windows to evaluate model decay over time. Analyse feature stability and hyperparameter sensitivity to ensure model robustness across different customer segments, brands, packaging, or markets. Automate the baseline-model validation process to establish repeatable benchmarks and speed up iteration on new models.
Productionization & Automation : Productionize helper functions and data pipelines. This includes automating calculations (e.g. conversion/redemption rates) and standardizing feature computation so models can be retrained quickly and reliably. Collaborate with data engineers to implement scalable pipelines on Azure (Data Factory, Data Lake, Databricks) and deploy models via CI/CD processes (e.g. using Azure DevOps or similar tools). Ensure solutions adhere to secure coding practices and data governance standards.
Technical Leadership (Individual Contributor): Act as a subject-matter expert and individual contributor on data science best practices. Write clean, maintainable code in Python (following OOP principles and design patterns) and document workflows in Jupyter Notebooks or VS Code. Stay current with emerging tools and techniques to continuously improve our data infrastructure and methodologies.
Communication & Visualization : Translate technical results into clear business insights. Create dashboards or reports (e.g. using Power BI, Streamlit, or similar) to present findings to non-technical stakeholders. Provide concise documentation and explanation of model assumptions, limitations, and impacts on business decisions.
Required Qualifications and Skills
Programming & Technical Skill s: Expert proficiency in Python (including OOP and design patterns) for data science. Comfortable coding in Jupyter Notebook and VS Code. Hands-on experience with ML libraries (scikit-learn, TensorFlow, PyTorch) and data libraries (pandas, NumPy). Familiarity with big data tools (Spark, Dask, etc.) and version control (Git).
Cloud & Deployment: Demonstrated expertise with Azure cloud services for end-toend data solutions. This includes Azure Machine Learning (for model development and deployment), Azure Data Factory (ETL/ELT pipelines), Azure Databricks (data processing), and related services. Ability to architect scalable data pipelines and CI/CD processes in Azure is crucial. Experience with containerization (Docker) and orchestration (Kubernetes) is a plus.
Statistical & Analytical Skills: Deep understanding of statistical modelling, experimental design, and machine learning. Ability to select and apply appropriate evaluation metrics and rigorously validate models to ensure high predictive performance. Familiarity with advanced techniques (e.g. Bayesian methods, time series analysis) is beneficial.
Business Acumen & Communication: Strong business sense to translate analytical outputs into actionable strategies. Proven ability to collaborate with stakeholders and clearly communicate complex technical concepts to non-technical audiences.
Soft Skills: Highly organized, proactive problem-solver who thrives in a fast-paced environment. Self-motivated individual contributor with a continuous learning mindset. Attention to detail and commitment to quality in all deliverables.
... proficiency in Python, Python-spark, shell scripting and SQL, with hands-on experience in building large-scale ETL/ELT pipelines Good to have exposure in the Big Data Ecosystem like Hive, Spark and HDFS Good to have exposure with modern data stack components (Airflow, Iceberg). Strong understanding of APIs, microservices, ...
... designing scalable solutions for Data Warehouses, Data Lakes, batch processing, and real-time streaming . Ideal Candidate The ideal candidate is a hands-on Senior Data Engineer who can independently design and deliver scalable data solutions using Python, SQL, PySpark, and GCP , with strong expertise in BigQuery, Dataproc, ...
... transformation, and loading of data from a wide variety of data sources using SQL and AWS technologies u2022 Develop, maintain and optimize ETLs to increase data accuracy, data stability, data availability and pipeline performance u2022 Strong knowledge in the Big data tools such as AWS Redshift, Hadoop, map reduce etc. ...
... you an all-around data enthusiast with a knack for ETL We're hiring Data Engineers to help build and optimize the foundational architecture of our product's data. We've built a strong data engineering team to date, but have a lot of work ahead of us, Migrating from relational databases to a streaming and big data architecture, ...
... improved efficiency, and stronger operational performance. Role: Senior Data Scientist Location: Bangalore- Karnataka (Hybrid) Experience: 5+ years Department: Data Science Role Overview We are looking for individuals who sit at the intersection of Data Science and Physical Sciences . You understand that industrial data isn't ...
... architecture, and delivery of large-scale data engineering solutions across cloud and hybrid environments. Drive end-to-end data platform implementations encompassing data ingestion, transformation, storage, processing, governance, and consumption. Architect batch and real-time data pipelines leveraging modern big data frameworks ...
... reporting and product development teams 3 to 6 years of experience in the field of data engineering - Graduate degree or higher with courses in programming, data analytics - Good programming skill in Python/scala or any scripting language - Good knowledge of SQL is a required - Basic knowledge and understanding of big ...
... looking for a Big Data Engineer to drive our mission to unlock potential of data assets by consistently innovating, eliminating friction in how users access data from its Big Data repositories and enforce standards and principles in the Big Data space. The candidate will be part of an exciting, fast paced environment developing ...
... statements, and offering solutions in the domains of Retail, Pharma, Banking, Insurance, etc. - Contribute to internal product development initiatives related to data science. - Develop data science roadmap, and guide data scientist to meet their deliverables. - Handling end-to-end client AI & analytics programs. Your role ...
Experience Level: 2-5 years Location: Bengaluru Education: Bachelor's or Master's degree in Computer Science, Data Science, Mathematics, Statistics, or related fields. A strong academic pedigree is highly preferred. Overview: We are seeking a talented and motivated Data Scientist to join our team. The ideal candidate will ...
The RoleWe are looking for a Data Scientist who wants to work on some of the hardest and most interesting problems in healthcare https://jobeax.com/link/80R4u3Mod2ymK61H will work with large-scale, longitudinal EHR data to develop models across areas such as:Clinical prediction and risk stratificationPatient journey and ...
Role Summary We are hiring a hands-on Senior Data Scientist / AI Engineer to design and productionize AI capabilities for automated understanding and validation of complex 2D engineering drawings. The role combines Computer Vision, OCR/Document AI, multimodal LLMs, RAG/Graph-RAG, Knowledge Graphs and deterministic engineering ...
Our partner is looking for a Senior Data Scientist – Adversarial based in India. Join a global web security environment where advanced data science directly contributes to protecting major digital platforms and brands. As a Senior Data Scientist, you will investigate sophisticated attacks, analyze large-scale traffic data, ...
... innovation here – we make it happen, with our customers, where we work and live. Together, we are building a better world, so we can all enjoy living in it. As a Senior Data Scientist, you will lead analytics programs, provide technical direction, and partner with cross functional stakeholders to deliver high impact solutions ...
... features and data contracts. Build and ship models to: Enhance customer experiences and personalization Boost revenue via pricing/discount optimization Detect and block fraud/risk in real time Collaborate with Engineering to productionize via APIs/CI/CD/Docker on AWS. Build monitoring for model/data drift and business KPIs
... a multi-layer multi-objective problem space which has a unique impact on Global consumer base and powers assisted AI for our associates. The team consists of data scientists, application engineers, big data geeks and product visionaries all working together to design, prototype and build technology-driven products and experiences ...
... build, and maintain scalable data pipelines and architectures for large-scale data environments. - Translate business and technical requirements into reliable data engineering solutions. - Develop data integration, transformation, and processing workflows using Python and appropriate big data technologies. - Build, manage, ...
... build, and maintain scalable data pipelines and architectures for large-scale data environments. - Translate business and technical requirements into reliable data engineering solutions. - Develop data integration, transformation, and processing workflows using Python and appropriate big data technologies. - Build, manage, ...
... efficiently: spatial indexing (R-tree / GiST), optimised spatial joins, partitioning, query-plan diagnosis, and geometry simplification - to control runtime and cost.- Big-data and pipeline fluency: Advanced SQL plus distributed processing for large spatial workloads (Spark or Dask), and building reliable, repeatable data pipelines.- ...