Requirements Strong programming skills in Python and SQL. Experience with data processing leveraging programming skills in Python and Spark. In-depth understanding of the Kafka platform for (real-time) data ingestion and processing of high-volume data. Design and architect data flows and data management in a Cloud environment ...
... factual accuracy, or comparing responses — when projects are available. While each project involves unique tasks, contributors may: - Carefully review provided data (text, images, or videos); - Label or classify content based on project guidelines; - Identify and flag factually incorrect, sensitive, inappropriate, or unclear ...
... shape. Senior Data Scientist We have an opportunity for a Senior Data Scientist to join our Data Science team in Bangalore, reporting to the Senior Director, Data Science. In this pivotal role, you will drive advanced analytics and machine learning solutions that power our security, networking, and cloud data products. ...
... Product, Data and Design teams to assist with data-related technical issues and support their data infrastructure needs. Create data tools for analytics and data scientist team members that assist them in building and optimizing our product into an innovative industry leader. Work with data and analytics experts to strive ...
... work experience. - Experience in building scalable products with preferably big data. - Excellent Python coding skills (Mandatory) - Experience in Apache spark, Data Lake and other Big data technologies. - Experience in either Data Warehouses or Relational Database is mandatory. - Experience in AWS cloud Mandatory Skills : ...
... statistical modelling, and explainability libraries. Data Platforms & Engineering: Experience working with Snowflake, Snowpark, Spark/PySpark, MLflow, cloud-based data platforms, and scalable analytical environments. Healthcare Data Expertise: Experience working with de-identified patient-level data including claims, specialty ...
... governance principles, data quality checks on each data layer. - A successful history of manipulating, processing and extracting value from large disconnected datasets. - Working knowledge of message queuing, stream processing, and highly scalable big data data stores. - Strong project management and organizational skills. ...
... and mitigate data risks throughout the data lifecycle, including protection, retention, storage, use, and quality l Partner with technology teams to capture data sources, formats, and data flows so that data can be validated for downstream analytics and reporting l Investigate and document potential data quality issues, ...
... including generative AI and LLM-based approaches — using enterprise business data, knowledge graphs, business process intelligence, and structured and unstructured data assets. - Learn SAP's deep data and process context — data models, metadata structures, and business process semantics across Order-to-Cash, Procure-to-Pay, Record-to-Report, ...
... vector-based data pipelines for AI and GenAI use cases Develop scalable batch and streaming data pipelines using cloud-based data platforms (e.g., Snowflake or Databricks) Create and evolve semantic data models that transform raw data into analytics-ready, trustworthy datasets Build data preprocessing, validation, and quality-assurance ...
... Install, configure, and update database software and tools - Monitor and optimize database performance and capacity - Perform backup and recovery operations- Ensure database security and compliance with policies and standards - Diagnose and resolve database issues and errors - Support database development and testing activities ...
... Perform data cleaning and preprocessing to prepare datasets for analysis. Write complex SQL queries to extract, manipulate, and analyze data from relational databases. Work with cloud platforms (e.g., AWS, Azure, Google Cloud) to store, manage, and analyze data. Use statistical techniques and software to analyze datasets, ...
... reliable, well-documented data products that drive operational and strategic decision-making. The ideal candidate is comfortable operating independently across the data stack, takes ownership of their work, and communicates proactively with team members across time zones. Core Responsibilities ETL / Data Pipeline Development ...
... with LangChain , RAG , and agent workflows Translate business goals into production-ready AI applications Work with prompt engineering , embeddings, and vector databases Collaborate with engineering, product, and data teams on performance optimization Deploy AI solutions on AWS, Azure, or GCP and ensure cloud efficiency Requirements: ...
... and internal partners to develop, test and operationalize data science and analytical solutions as well as ensure its adoption by businesses. Work toleveragedata (Transactional / Big Data, External/ Internal Data, Structured/ Unstructured Data) and analytics methodsto develop analytics solutions 6 - 8 years of experience, ...
... LangChain, LlamaIndex, and Hugging Face. Orchestration: LangGraph and Multi-Agent Systems (MAS). Development: Python, FastAPI, and Asynchronous Programming. RAG & Data: PostgreSQL, Vector Databases, and Advanced Retrieval strategies. ML/DL: PyTorch, TensorFlow, and Model Fine-tuning. Deployment: Docker, Production API management, ...
... degree in data science, Mathematics, Statistics, Computer Science, or related field AND 5+ years related experience (e.g., managing structured and unstructured data, applying statistical techniques and reporting results) OR master's degree in data science, Mathematics, Statistics, Computer Science, or related field AND 4+ ...
... ownership, OR - 2-3 years as Data/Backend Engineer with PM work (PRDs, roadmaps, user research) Technical Stack Depth Deep understanding: ETL/ELT pipelines, data warehousing, data lakes/lakehouses Advanced SQL (query optimization, not just SELECT statements) Hands-on with ONE of: Hive/Spark/Hadoop/HDFS, Snowflake/Databricks/BigQuery, ...
Principal Data Scientist with GenAI - 8+ years - BangaloreSummary :We are looking for an experienced Principal Data Scientist / GenAI professional who can independently develop, implement, and deploy Machine Learning and Generative AI use cases across functions. The role involves working closely with internal teams and ...