... data processing solutions using frameworks such as Apache Beam (Data Flow) and Kafka. Work with GCP services like BigQuery, Dataflow, and Spanner to optimize data pipelines and workflows. Design and model data structures for efficient data storage and retrieval. Oversee ETL processes to ensure seamless data extraction, ...
... https://jobeax.com/link/0a2QmNkVhUZFALjk Candidate will have :- Experience with Agentic AI, Generative AI, LLMs, and prompt engineering.- Experience with big data technology stack (Hadoop, Spark, HDFS, EMR, Glue).- Experience with AWS, Azure or GCP.- Experience with Databricks/SageMaker/DataRobot, MLFlow or other ML and ...
... daysMandatory Skills : - Python- PySpark- SQL- Oracle EBS- Databricks / Azure Data Factory (ADF)Key Responsibilities : - Design and develop scalable data pipelines using Databricks and PySpark.- Develop and optimize complex SQL queries and data processing workflows.- Work with Oracle EBS data structures and support data integration ...
... EMR, and other AWS services.- Develop PySpark-based data processing applications for large-scale distributed data workloads.- Write optimized SQL queries for data transformation, analysis, and reporting requirements.- Implement data ingestion frameworks to process data from databases, APIs, files, applications, and other ...
... in Data Science, Computer Science, Statistics, Business Analytics, or related field, or equivalent experience Business Acumen, Collaboration, Communication, Data Analysis, Data-Driven Business Intelligence, Data Visualization, Market/Industry Dynamics, Relationship Management, Statistics, Story Telling _____________________________________________________________________________________________________ ...
... deploying complex data architectures within Microsoft Fabric and the broader Azure ecosystem.- Advanced proficiency in Python and SQL for building sophisticated data processing pipelines and complex analytical models.- Strong background in data engineering principles, including ETL/ELT design, data modeling, and performance ...
... with vendors and other IT personnel for problem resolutio Proven hands-on network engineering experience of 6 – 8 Yr Cloud Networking – AWS & Azu reCisco Nexus datacenter Family switche network analytics too ob:You are proficient at managing networks. We have vast networks as a large technology company that specializes in ...
... Data Platform & Orchestration is seeking a Lead Data Engineer to design and build next-generation, cloud-native data platforms supporting Mastercard's global data ecosystem. In this role, you will lead the development of scalable batch and real-time data pipelines, enabling efficient data processing across Data Lakes and ...
... ETL Tool ONLY amongst Talend Datastage Informatica ISS Cloud version Data Fusion Data Flow Data Proc Mandatory SQL PLSQL Scripting Experience Expert in Python Data Flow Pubsub Big Query CICD Must have good Experienceknowledge on GCP components like GCS BigQuery AirFlow Cloud SQL PubSubKafka DataFlow and Google Cloud SDK ...
... in regulated industries (Pharma, Life Sciences, Banking). Design, build, and maintain scalable data pipelines supporting AI and GenAI use cases Work on modern data platforms (data lakes, lakehouse, streaming, Event hub like Kafka, containers, K8s) Enable high-quality, compliant data for analytics, ML, and GenAI workloads ...
... position is for a KPO Network Operations/Project Engineer. The project team will be responsible for supporting a multi-vendor network environment, including Data Center (DC), Load Balancers (LB), and Firewalls, along with F5 and Juniper technologies within the Defra infrastructure. Experienced in hardware and infrastructure ...
... complex and large-scale datasets. Strong experience in data analysis, data profiling, reconciliation, and Data Quality (DQ) validation. Ability to explore large datasets and translate analysis into clear findings, recommendations, and actionable insights. Experience working with large volumes of data and understanding of performance ...
Litmus is building the data foundation that powers industrial AI. AI doesn’t work without real-world, contextualized data - Litmus makes that data usable. As AI adoption accelerates, most industrial environments still can’t access or use their operational data. We’re a growth-stage software company helping manufacturers ...
... Python LLM / NLP EDA (Exploratory Data Analysis) Cloud Platforms (AWS, Azure, GCP) Explainable AI At least 4+ years of relevant experience and track record in Data Science: Machine Learning, Deep Learning and Statistical Data Analysis. MSc or PhD degree in CS or Mathematics, Bioinformatics, Statistics, Engineering, Physics ...
... skills; Experience as a Data Engineer or in a similar role, with a strong understanding of data engineering concepts and methodologies, including data modeling and database design to support scalable data solutions. Strong knowledge of writing and optimizing SQL queries to retrieve, manipulate, and analyze data efficiently. Experience ...
... Have :- SAP Business Data Cloud data products / BTP administration.- SAP Databricks / Databricks (PySpark, Delta Lake, Unity Catalog).- Git-based workflows, CI/CD for data artefacts (abapGit, gCTS).- Microsoft Fabric / Power BI exposure.- SAP certification (Datasphere, HANA Cloud Modeling, or BW/4HANA). (ref:hirist.tech)
... with Apache Airflow — DAG design, operator customisation, dependency management, and operational best practices - Solid understanding of dimensional modelling, data vault, and semantic layer design Technical Skills — Preferred - Experience designing data pipelines and feature engineering workflows for ML model training and ...
... Technically familiar with the stack to recommend infrastructure improvements Desirable Skills Experience with Plotly library structure (For pictorial presentation of the data) Experience with Databricks/Database and schema knowledge Soft Skills Solid communications skills Negotiation skills with the highly technical SMEs
... Glue, Lambda, Step Functions, and Apache Spark . Develop and optimize data lake and data warehouse architectures using AWS S3, Redshift, and Athena . Implement data ingestion, transformation, and storage solutions using AWS Glue, Kinesis, and Kafka . Work with structured and unstructured data to develop efficient data pipelines ...