Data Scientist - Python/ SQL- Fully Remote in Pune, India - Jobeax
Vacancy description
Data Scientist - Python/ SQL- Fully Remote in Pune, India
Fusemachines
RemoteWork from anywhere
Full-timeStandard weekly hours
India, Pune
Data Scientist - Python/ SQL- Fully Remote in Pune, India is listed on Jobeax. Browse 30,000+ vacancies available.
Fusemachines is a leading AI strategy, talent, and education services provider. Adjunct Associate Professor at Columbia University, Fusemachines has a core mission of democratizing AI. With a presence in 4 countries (Nepal, the United States, Canada, and the Dominican Republic) and more than 450 full-time employees, Fusemachines brings global AI expertise to transform companies worldwide. Founded in 2013, Fusemachines is a global provider of enterprise AI products and services, on a mission to democratize AI. Leveraging proprietary AI Studio and AI Engines, the company helps drive the clients' AI Enterprise Transformation, regardless of where they are in their Digital AI journeys. With offices in North America, Asia, and Latin America, Fusemachines provides a suite of enterprise AI offerings and specialty services that allow organizations of any size to implement and scale AI. Fusemachines continues to actively pursue the mission of democratizing AI for the masses by providing high-quality AI education in underserved communities and helping organizations achieve their full potential with AI. Type: Remote, Full-time You will build and validate the models that define, expand, and score the audiences advertisers activate across TV, CTV, digital, social, and programmatic. Working across survey, purchase, and cross-platform media data, you turn messy multi-source signals into accurate, explainable, privacy-safe segments, and you help shape the product itself with ideas for new audience capabilities. You will partner closely with ML Engineers and the US product and analytics team. Build statistical and ML models to create, expand, and score audience segments from survey, panel, purchase, and media-exposure data Lead data fusion work, combining deterministic and probabilistic sources into one representative consumer view while correcting for bias Partner with ML Engineering to move models into reproducible, monitored production, under privacy-by-design principles 5+ years in applied data science, with models that reached production or client delivery
Strong Python (pandas, scikit-learn, statsmodels) and SQL against large databases
Degree in a quantitative field (Statistics, Data Science, Economics, CS, Math, or similar)
Audience/identity or ad tech exposure (DSP/SSP, DMP/CDP, clean rooms, identity graphs) Familiarity with privacy-preserving methods and GDPR/CCPA in data collaboration Fusemachines is an Equal Opportunities Employer, committed to diversity and inclusion. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or any other characteristic protected by applicable federal, state, or local laws.
... and development Data quality validation and monitoring Schema design for analytics and reporting Performance optimization and scalability Preferred Experience Databricks Delta Lake experience Azure Synapse Analytics Python for data engineering Spark SQL optimization Real-time data streaming Data governance and metadata management ...
... and lead our push into streaming and real-time fraud detection. We run on AWS and make extensive use of Scala Spark, dbt, and Terraform. Build and expose the data lake/lakehouse so teams can understand business and product performance and make better decisions. Design and evolve a data platform that enables scientists and ...
... team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL to transform and cleanse data. Configure and manage ...
... inquiries related to Addepar's portfolio data feeds and on general data product functionalities within established SLAs. Manage and complete requests from internal Data teams that require client outreach and/or action to resolve data verification issues. Investigate client reported bugs and data processing issues, and triage ...
... Warehousing - 2+ years of hands-on experience with SSAS Tabular Models - Experience designing data ingestion and orchestration pipelines using Kafka, Snowflake, and Python - Hands-on experience with DBT for data modeling and pipeline development - Strong knowledge of SQL, ETL/ELT, Azure Data Factory, APIs, Azure Functions, and ...
... Troubleshoot and debug data-related issues. Required Skills Strong proficiency in Python. Experience with data libraries like Pandas, NumPy. Hands-on experience with SQL and relational databases. Knowledge of data visualization tools (Matplotlib, Seaborn, or similar). Experience in building data pipelines or ETL processes. Familiarity ...
... demonstrate technical and non-technical information, ideas, procedures, and processes Prior experience in an SAP and/or Informatica driven environment Ability to use SQL to analyze data Bachelor's degree related to Information Systems, Business or other relevant academic discipline with a minimum 4 years job experience in IT or ...
... a related field (Master's degree preferred). Core Technical Skills: Python, PySpark, Databricks, Google ADK, LLMs, FastAPI, Spring Boot, Microservices, Kafka, SQL, Data Mesh, Starburst. Infrastructure and Cloud: Kubernetes, OpenShift, Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering ...
... models - Hands-on experience working with large language models, building agents, prompt engineering, or model fine-tuning techniques. - Strong proficiency in Python & Data Science Toolkit - Extensive experience with Python-based data science tools and machine learning frameworks. Benefits & Perks - Excellent group health ...
Principal Data Scientist - 7+ Years - BangaloreWe are hiring a Principal - Data Scientist with GenAI for a leading global AI and analytics organization. The role offers an opportunity to work on complex business problems and develop, implement and deploy Machine Learning and Generative AI solutions across https://jobeax.com/link/iWvc74qOaz1N3AKd ...
Role Overview :The Senior Data Scientist, Product Analytics is a hands-on technical contributor and task manager within a cross-functional product team. This role sits at the intersection of deep technical execution and emerging AI https://jobeax.com/link/f9LyLs7glaUomOJP And Qualifications :- 5 - 7 years of experience ...
Role Overview : As a commercially savvy Lead Data Scientist, you will lead your team on client briefs whilst collaborating closely with consultants and Trade Partner stakeholders. You'll bring fresh ideas and innovative thinking to push the boundaries of what's possible with consumer analytics. You will drive AI excellence ...
... observable and human-governed AI-assisted workflows, evaluating their quality, and integrating them responsibly into existing data pipelines and business processes. Data Science, Engineering & Quality Design, build and maintain scalable data-processing and analytical workflows using Python, PySpark and SQL; Collect, clean, reconcile ...
Kroll is hiring a Senior Data Scientist to join its Enterprise Data Group. This role is designed for an experienced practitioner who can lead end-to-end ML initiatives, mentor junior team members, and partner with business and engineering stakeholders to translate complex problems into production-grade data science solutions. ...
... 2–4 years of Data Science experience, ideally including 3+ years focused on Credit Risk, Fraud Analytics, Lending, or Fintech. - Production-grade fluency in Python and advanced, highly analytical SQL. - Modern Stack Experience: Hands-on experience scaling data workflows over frameworks like Snowflake, Databricks, Spark, ...
... applications, or hybrid modeling approaches (physics-informed ML, combining domain knowledge with ML) - Familiarity with Agile development methodologies - Knowledge of data engineering, data warehousing, and SQL - Experience using data science techniques to solve real-world problems across multiple business domains and communicate ...
... bottom-line and top-line of India business. 2+ years of data scientist experience - 3+ years of data querying languages (e.g. SQL), scripting languages (e.g. Python) or statistical/mathematical software (e.g. R, SAS, Matlab, etc.) experience - 3+ years of machine learning/statistical modeling data analysis tools and techniques, ...