... and real time data pipelines Build and maintain robust ETL ELT processes for ingesting transforming and loading data from multiple sources Develop and manage data lakes data warehouses and metadata frameworks Collaborate with cross functional teams to understand data requirements and deliver data solutions Ensure data quality ...
... stakeholders. Data Pipeline Design Operations: Design, develop, and support data pipelines to extract, transform, and load (ETL/ELT) data into the new harmonized data platform. Utilize Azure Data Factory (ADF) for orchestration (scheduling, workflow management) and Azure Databricks (Spark) for large-scale data transformation ...
Data Entry Executive Start Date Starts Immediately CTC (ANNUAL) ₹ 2,00,000 - 2,16,000 ₹ 2,00,000 - 2,16,000 /year Experience No experience required No experience required Apply By 19 Sep' 26 Posted few hours ago Fresher Job 130 applicants About the job Job description: Job Title: Data Entry Executive Location: Work From ...
... requirements from the various stakeholders (Data Scientists, Product Managers, IM) High interest in developing E2E data solutions DevOps/ Agile WoW Experience in designing and building advanced reports Programming languages: Python, PySpark, SQL Data enthusiast Big data frameworks: Experience in supply chain Familiar with Agile
... and professionally. Curious what it's like to work at Iris Head to this video for an inside look at the people, the passion, and the possibilities. Databricks Workflows,PySpark,Apache Kafka,Snowflake Design scalable data engineering solutions using PySpark and modern distributed data processing frameworks. Define data ingestion, ...
... experience in Data engineering. Strong expertise in GCP and BigQuery. Hands-on experience with Airflow for workflow orchestration. Proficiency in Java or Python for data processing and automation. Solid understanding of Data modeling, ETL, and Database optimization. Experience working in Agile delivery environments. We weave diversity ...
... (BigQuery, Dataflow, Dataproc), or Snowflake. - Proven ability to lead data engineering teams and mentor junior engineers. It Would Be Great If You Have Experience working with IoT or real-time streaming data. Knowledge of data governance frameworks and best practices. Exposure to MLOps workflows and integrating data with machine ...
... quality and integrity through rigorous data validation and cleansing processes. Collaborate closely with data scientists and analysts to provide clean, structured data for advanced analytics and machine learning models. Oversee and manage big data technologies and frameworks, including Hadoop, Spark, and Apache Airflow, to handle ...
... You'll work at the intersection of engineering and real business impact, turning raw data into something reliable, fast, and meaningful. Key Responsibilities: Data Architecture and Design: Design scalable, reliable, and secure data solutions that meet real business needs Develop data models, data flows, and integration architectures ...
... Databricks or equivalent Spark-based platforms. - Strong proficiency in Python, SQL, PySpark, and Delta Live Tables. - Experience building API integrations and working with financial data feeds (custodian files, market data APIs, fund administrator reports). - Understanding of security master management and corporate action ...
... and Real-Time data pipelines Understand business requirements and translate them into scalable data architecture and technical solution s Work with cloud-based data platforms and services, preferably Google Cloud Platform (GCP ) Hands-on experience with GCP managed services such as: Google Cloud Storage, Dataproc, Dataflow, ...
... schemas, staging models, validation processes, and data refinement workflows. Design and operate data storage and search solutions supporting large-scale healthcare datasets. Build AI-powered data tooling that improves the accuracy, automation, and intelligence of data processing workflows. Use modern AI development tools as an ...
Role Overview : We are seeking a high-caliber Data Scientist and AI/ML Engineer to join our innovation hub in Gurgaon. In this role, you will bridge the gap between complex data architecture and cutting-edge Generative AI applications, working closely with cross-functional product teams and business stakeholders to solve ...
... Python‑based inference and orchestration workloads Build serverless AI workflows using AWS Lambda, API Gateway, and event‑driven architectures Implement secure cloud networking and access controls using VPC, IAM, private endpoints, and service networking Integrate data services (S3, DynamoDB, messaging/streaming) to support AI pipelines ...
... 14 years of overall experience, with at least 3 - 5 years leading data science or analytics teams.- Strong understanding of machine learning, statistics, and data modeling.- Hands-on experience with Python, R, SQL, and modern ML frameworks (e.g., scikit-learn, TensorFlow, PyTorch).- Proven ability to manage projects end ...
... model training, fine-tuning, and deployment strategies. Design and implement scalable data pipelines to curate, cleanse, and prepare structured and unstructured data for machine learning and foundation model training. Build and optimize complex agentic AI workflows using modern orchestration frameworks to drive intelligent ...
... (Terraform) and CI/CD frameworks (Tekton).- Cloud Optimization : Actively monitor and optimize cost and compute resources for high-intensity processes (Cloud Run, Dataflow, BigQuery) to ensure platform efficiency.- Data Governance & Security : Implement robust data encryption, masking techniques, and governance frameworks to ...
We are seeking a detail-oriented Data Associate to join our growing team. As a Data Associate, you will work closely with our product managers, financial analysts, and engineers to ensure high-quality data integration, analysis, and maintenance. The ideal candidate will have a passion for working with large datasets, a ...
... tenant data warehouse layer that underpins multi-tenant AI Agent workloads. - Build and govern the Data Agent layer that supplies on-prem/customer-controlled data to cloud LLM reasoning in the Hybrid AI zone, with staged, explainable data outputs and human review checkpoints. - Define data classification and routing rules ...
... and provide data- driven solutions. Manage multiple disparate data sources and integrate them into a cohesive data platform and dashboards. Develop and implement validation processes to ensure data accuracy and consistency. Utilize pyspark, spark, sql, and Databricks to optimize data processing and querying performance.