... and real time data pipelines Build and maintain robust ETL ELT processes for ingesting transforming and loading data from multiple sources Develop and manage data lakes data warehouses and metadata frameworks Collaborate with cross functional teams to understand data requirements and deliver data solutions Ensure data quality ...
... stakeholders. Data Pipeline Design Operations: Design, develop, and support data pipelines to extract, transform, and load (ETL/ELT) data into the new harmonized data platform. Utilize Azure Data Factory (ADF) for orchestration (scheduling, workflow management) and Azure Databricks (Spark) for large-scale data transformation ...
... Kinesis, FireHose, Lambda, and IAM roles and permissions Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases) Experience with dbt (data build tool) for transformation layer management and semantic modeling. Familiarity with data lakehouse ...
... Qualifications Educational qualifications & Experience: Bachelor's degree in computer science, Information Security, or related field. 10–12 years of experience in data loss prevention, data discovery/classification, and database/data security. Experience with DSPM tools and integration in hybrid environments. Familiarity with ...
Data & AI ArchitectLocation: Hyderabad / Bangalore / PuneKey Technologies:Databricks | Microsoft Fabric | GenAI | LLM | RAG | PySpark | Data Lake | Lakehouse | Data Warehouse | Data ModellingRole Overview :We are looking for an experienced Data & AI Architect to design and lead enterprise - scale Data and AI architectures. ...
... SQL. Develop and manage Vector Database pipelines for embedding storage and retrieval. Implement text normalization, sentence segmentation, deduplication, and data quality processes. Design and implement data masking, classification, and categorization solutions. Collaborate with AI/ML engineers to prepare datasets for model ...
... positive changes in an increasingly virtual world and it drives us beyond generational gaps and disruptions of the future. We are looking forward to hire Azure Data Lake (ADL) Professionals in the following areas : Role - Azure Data Engineer Azure Data Lake, Data Factory, Pipelines, Synapse (mandatory) SQL, Python, PySpark ...
... office environment. Duties may require extended periods of sitting and sustained visual concentration on a computer monitor or on numbers and other detailed data. Repetitive manual movements (e.g., data entry, using a computer mouse, using a calculator, etc.) are frequently required. Education & Experience - Bachelor's ...
... Science & Supply Chain talent. Collect business requirements from the various stakeholders (Data Scientists, Product Managers, IM) High interest in developing E2E data solutions DevOps/ Agile WoW Experience in designing and building advanced reports Programming languages: Python, PySpark, SQL Data enthusiast Big data frameworks: ...
... reliable, high-quality data, driving excellence in analytics and shaping outcomes that impact industries like H ealthcare and Financial Services . Design and build Data Pipelines using GCP services and BigQuery. Develop and optimize data solutions using Java or Python. Ensure data quality, governance, and compliance across projects. ...
... quality. Big Data - Big Data - Pyspark Database - Database Programming - SQL Data Science and Machine Learning - Data Science and Machine Learning - Apache Spark Data & AI - Data Engineering - Data Quality & Validation Data & AI - Data Engineering - Apache Kafka Big Data - Big Data - Azure databricks Data Science and Machine ...
Experience: 3 -5 Years (In Data Analytics) Mode: Hybrid - Fulltime Location: Bangalore Must Have Skills: Advanced SQL, Complex queries, Business Analysis, Experimentation Prefer: Immediate Joiners/Serving Notice Key Skills & Qualifications: Advanced SQL expertise with strong proficiency in writing and optimizing complex ...
... Machine Learning, Scikit-learn, XgBoost, SQL GeoIQ (a Lenskart subsidiary) is India's leading hyperlocal location AI platform, helping businesses make precise, data-driven decisions using street-level intelligence across demographics, income, infrastructure, and commercial activity. As a Data Scientist, you will be a core ...
... including generative AI and LLM-based approaches - using enterprise business data, knowledge graphs, business process intelligence, and structured and unstructured data assets. - Learn SAP's deep data and process context - data models, metadata structures, and business process semantics across Order-to-Cash, Procure-to-Pay, Record-to-Report, ...
... change, applying first-principles problem solving, and using data to learn and adapt along the way. Enterprise Data & AI Mission Enterprise Data & AI provides the data and systems to enable our enterprise transformation from human-powered operations to data-driven, AI-powered, and human-led agentic operations. This data and ...
Data Scientist – Mortgage Analytics & AI Experience – 5 to 10 years About the Role We are looking for an experienced Data Scientist to join our team and help build the next generation of data-driven solutions for the mortgage and real estate industry. In this role, you will work at the intersection of data science, machine ...
Job Role: Senior Platform Data Engineer (Databricks) Architect and operate large-scale distributed data platforms using Databricks and Spark. Architect Databricks deployments Drive performance standards 8+ years of data or cloud engineering experience - Production ownership experience - Responsible for the end-to-end administration, ...
... (Pandas, Polars) and modern big data tools such as Databricks (Spark), Flink, or Kafka. - Strong understanding of data pipeline development, ETL frameworks, and data lakes/warehouses. - Extensive experience with cloud data platforms such as AWS (S3, Redshift, Lambda, Glue), GCP (BigQuery, Dataflow, Dataproc), or Snowflake. ...
... processes to seamlessly integrate data from multiple sources into our data warehouse. Implement advanced data modeling techniques and design efficient, optimized databases to support complex business needs. Ensure the highest levels of data quality and integrity through rigorous data validation and cleansing processes. Collaborate ...
... scalable data processing solutions using Big Data technologies and distributed system s Strong hands-on experience with Python and PySpar k for developing efficient data pipelines and applications Develop and optimize complex SQL querie s, data models, and database solutions Analyze and improve performance of Apache Spark workload ...