... SQL. Develop and manage Vector Database pipelines for embedding storage and retrieval. Implement text normalization, sentence segmentation, deduplication, and data quality processes. Design and implement data masking, classification, and categorization solutions. Collaborate with AI/ML engineers to prepare datasets for model ...
... disruptions of the future. We are looking forward to hire Azure Data Lake (ADL) Professionals in the following areas : Role - Azure Data Engineer Azure Data Lake, Data Factory, Pipelines, Synapse (mandatory) SQL, Python, PySpark / Spark-SQL Web tech, relational DBs, data warehousing, ETL design Microsoft Certified Data Engineer ...
... (Pandas, Polars) and modern big data tools such as Databricks (Spark), Flink, or Kafka. - Strong understanding of data pipeline development, ETL frameworks, and data lakes/warehouses. - Extensive experience with cloud data platforms such as AWS (S3, Redshift, Lambda, Glue), GCP (BigQuery, Dataflow, Dataproc), or Snowflake. ...
... scalable data processing solutions using Big Data technologies and distributed system s Strong hands-on experience with Python and PySpar k for developing efficient data pipelines and applications Develop and optimize complex SQL querie s, data models, and database solutions Analyze and improve performance of Apache Spark workload ...
... processes to seamlessly integrate data from multiple sources into our data warehouse. Implement advanced data modeling techniques and design efficient, optimized databases to support complex business needs. Ensure the highest levels of data quality and integrity through rigorous data validation and cleansing processes. Collaborate ...
... understanding of data modeling, schema design, validation strategies, data transformation, and data quality management . - Experience designing solutions for data deduplication, entity resolution, data enrichment, classification, or record linkage . - Experience with both relational and document-oriented databases; practical ...
Experience: 3 -5 Years (In Data Analytics) Mode: Hybrid - Fulltime Location: Bangalore Must Have Skills: Advanced SQL, Complex queries, Business Analysis, Experimentation Prefer: Immediate Joiners/Serving Notice Key Skills & Qualifications: Advanced SQL expertise with strong proficiency in writing and optimizing complex ...
... Machine Learning, Scikit-learn, XgBoost, SQL GeoIQ (a Lenskart subsidiary) is India's leading hyperlocal location AI platform, helping businesses make precise, data-driven decisions using street-level intelligence across demographics, income, infrastructure, and commercial activity. As a Data Scientist, you will be a core ...
We help the world run better At SAP, we keep it simple: you bring your best to us, and we'll bring out the best in you. We're builders touching over 20 industries and 80% of global commerce, and we need your unique talents to help shape what's next. Data and Applied Science The context engine that makes AI enterprise ready. ...
... change, applying first-principles problem solving, and using data to learn and adapt along the way. Enterprise Data & AI Mission Enterprise Data & AI provides the data and systems to enable our enterprise transformation from human-powered operations to data-driven, AI-powered, and human-led agentic operations. This data and ...
Data Scientist – Mortgage Analytics & AI Experience – 5 to 10 years About the Role We are looking for an experienced Data Scientist to join our team and help build the next generation of data-driven solutions for the mortgage and real estate industry. In this role, you will work at the intersection of data science, machine ...
Job Role: Senior Platform Data Engineer (Databricks) Architect and operate large-scale distributed data platforms using Databricks and Spark. Architect Databricks deployments Drive performance standards 8+ years of data or cloud engineering experience - Production ownership experience - Responsible for the end-to-end administration, ...
... Integration: Build and manage data pipelines to ingest data from various sources, including third-party and RESTful APIs. Build ETL/ELT processes using ADF, Fabric Data Flows , and Notebooks. Embed MDM solutions into pipelines, ensuring seamless data sync across systems. Data Storage and Management: Design and implement data ...
Purpose of the role Build and maintain all data pipelines feeding the Databricks lakehouse, ensuring clean, timely, and governed data flows from every source across the Investment Division’s five portfolio clusters and, in later phases, the Group’s operating divisions. Own the bronze-to-silver-to-gold transformation logic, ...
... ability to design and fine-tune Deep Learning architectures and NLP models to solve real-world language processing challenges.- Strong proficiency in SQL and Big Data technologies, with the ability to manage and query large-scale databases to support data-intensive applications.- Exceptional communication skills, with the ability ...
... pipelines for enterprise data platforms, ensuring data quality, integrity, and security. Implement data engineering solutions utilizing Microsoft Fabric, Azure Data Factory, Databricks, PySpark, Spark SQL, and Python to process structured, semi-structured, and unstructured data. Build data pipelines that integrate data from ...
... events, depending on business needs. At NTT DATA, we are committed to staying flexible and meeting the evolving needs of both our clients and employees. NTT DATA recruiters will never ask for payment or banking information and will only use @nttdata.com and @talent.nttdataservices.com email addresses. If you are requested ...
... development of AI and machine learning solutions, including data preparation, model training, fine-tuning, and deployment strategies. Design and implement scalable data pipelines to curate, cleanse, and prepare structured and unstructured data for machine learning and foundation model training. Build and optimize complex agentic ...
... tools. Data Modeling & Architecture: Skilled in dimensional and semantic modeling, database design, and building performant, governed analytics-ready structures. Data Quality, Metadata & Curation: Capable of cleansing, transforming, and validating data; familiar with lineage, cataloging, and data quality frameworks. Experience ...
Job Description :We are seeking an experienced Data Science Manager to lead a team of data scientists and analysts in developing data-driven products, insights, and https://jobeax.com/link/5WM6jVJPxXxkaX5M ideal candidate will have a strong background in applied machine learning, data strategy, and stakeholder management, ...