... data assets by consistently innovating, eliminating friction in how users access data from its Big Data repositories and enforce standards and principles in the Big Data space. The candidate will be part of an exciting, fast paced environment developing Data Engineering solutions in the data and analytics domain. Develop ...
... Experience with cloud platforms (AWS / Azure / GCP).- Familiarity with MLOps tooling and practices (e.g., model versioning, CI/CD for ML, monitoring).- Experience with big data technologies (Spark, Databricks, etc.).- Bachelor's / Master's / PhD in Computer Science, Statistics, Mathematics, or a related quantitative https://jobeax.com/link/sEOksQVqU173JbV8 ...
... Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering experience in Banking Risk, Retail Products, Cards, Mortgage, Deposits, or Wealth Management. Assumed Requirements / Certifications: Databricks Certified Data Engineer, AWS Certified Data Analytics, or Azure Data Engineer Associate.
... Engineer Professional Additional Certifications (Preferred) - Databricks Certified Associate Developer for Apache Spark - Cloud platform certifications (Azure Data Engineer Associate, AWS Certified Data Analytics, or Google Cloud Professional Data Engineer) - Relevant data engineering or big data certifications Soft Skills ...
... efficiently: spatial indexing (R-tree / GiST), optimised spatial joins, partitioning, query-plan diagnosis, and geometry simplification - to control runtime and cost.- Big-data and pipeline fluency: Advanced SQL plus distributed processing for large spatial workloads (Spark or Dask), and building reliable, repeatable data pipelines.- ...
... composable platform, Aera empowers organizations to optimize and automate all types of decisions, across every business area. We are looking for an Associate Data Scientist who is exceptional at understanding data, extracting insights, and turning those insights into robust machine learning or statistical models. You will ...
... AI-first data strategy at scale across 120M+ customer interactions. Databricks Apps & Self-Serve AI Develop and deploy internal AI-powered applications using Databricks Apps — enabling business users to interact with ML models, RAG systems, and analytics agents through governed, self-serve interfaces. Integrate Databricks ...
... batch jobs, database replication & Kafka- Set up and maintain monitoring/observability for data pipelines- Support ML pipelines, Airflow workflows and production data science infrastructureIDEAL PROFILE:Looking for candidates with strong hands-on experience in:- Data Operations- Production Data Engineering- Data Platform Support- ...
... years of experience • Minimum 3 years of experience in design, build and deployment of Ab Initio-based applications • Expertise in handling complex large-scale Data Lake and Warehouse environments • Hands-on experience writing complex SQL queries, exporting and importing large amounts of data using utilities • Excellent verbal ...
... to: - Define and maintain enterprise data models and domain boundaries using traditional and AI-enabled methods - Design and govern reusable data assets and data products - Establish standards and patterns for robust, scalable data pipelines - Ensure data is discoverable, trustworthy, and fit-for-purpose for analytics, ...
... testing, and maintaining architectures such as databases and large-scale processing systems, ensuring that architectures support data analytics, and preparing data for prescriptive and predictive modeling. Data engineers also develop data set processes for data modeling, mining, and production, integrate new data management ...
... affordable healthcare. Build and manage digital analytics datasets and reporting solutions Develop data pipelines and analytical reporting frameworks Support data validation, UAT, process improvements, and analytics initiatives. Experience & Qualification: Python Power BI & Dashboard Development Data Modeling & Data Validation ...
... application development to join our team. The role involves managing development of FastAPI based applications to run large data modeling, and leveraging the Databricks platform. The ideal candidate will have a robust background in backend development, Data science model basics, and cloud-based solutions to drive data-driven ...
... and development Data quality validation and monitoring Schema design for analytics and reporting Performance optimization and scalability Preferred Experience Databricks Delta Lake experience Azure Synapse Analytics Python for data engineering Spark SQL optimization Real-time data streaming Data governance and metadata management ...
... build robust data pipelines, data lakes, and marts to support business analysts and data scientists. Key Responsibilities Modern Data Platform Development: Build data lake components on cloud-based platforms Design and develop data marts for business analysts and data scientists Data Engineering & Pipelines: Design data pipelines ...
... data infrastructure remains a competitive asset in a fast-paced https://jobeax.com/link/CO851KAPNv79RH57 Responsibilities : - Architect and maintain complex data models that serve as the foundation for enterprise-wide reporting and advanced analytics, ensuring high data integrity and performance.- Lead the end-to-end design ...
... optimized SQL queries for data extraction, transformation, and analysis.- Build and manage ETL pipelines using AWS Glue.- Work with AWS Glue Data Catalog for metadata management and data discovery.- Design and implement serverless data processing solutions using AWS Lambda.- Use Amazon S3 for scalable data storage and data ...
We are seeking a Data Engineer to join our team. The role will involve building and maintaining data pipelines and collaborating closely with stakeholders. Key responsibilities Design and develop scalable data pipelines using AWS services. Implement ETL processes with AWS Glue, Lambda, and DBT. Work with PySpark and SQL ...
... inquiries related to Addepar's portfolio data feeds and on general data product functionalities within established SLAs. Manage and complete requests from internal Data teams that require client outreach and/or action to resolve data verification issues. Investigate client reported bugs and data processing issues, and triage ...
... connectors or processing frameworks. - Deep understanding of Apache Airflow and ability to design and manage complex DAGs. - Solid SQL skills and familiarity with data warehouse platforms (e.g., Snowflake, Redshift, BigQuery). - Familiarity with version control tools (Git), CI/CD pipelines, and Agile methodologies. - Exposure ...