Data Engineer SQL - Remote (m/w/d) in India, India - Jobeax
Vacancy description
Data Engineer SQL - Remote (m/w/d) in India, India
artech l.l.c.
RemoteWork from anywhere
India, India
Data Engineer SQL - Remote (m/w/d) in India, India is listed on Jobeax. Browse 30,000+ vacancies available.
Databricks Engineer Location - Remote
Data & AI Platform Engineering (Databricks-Centric): Design, implement, and optimize end-to-end data pipelines on Databricks, following the Medallion Architecture principles. Build robust and scalable ETL/ELT pipelines using Apache Spark and Delta Lake to transform raw (bronze) data into trusted curated (silver) and analytics-ready (gold) data layers. Apply schema evolution and data versioning to support agile data development. Platform Integration & Data Ingestion: Connect and ingest data from enterprise systems such as PeopleSoft, D2L, and Salesforce using APIs, JDBC, or other integration frameworks. Implement connectors and ingestion frameworks that accommodate structured, semi-structured, and unstructured data. Design standardized data ingestion processes with automated error handling, retries, and alerting. Data Quality, Monitoring, and Governance: Develop data quality checks, validation rules, and anomaly detection mechanisms to ensure data integrity across all layers. Integrate monitoring and observability tools (e.g., Databricks metrics, Grafana) to track ETL performance, latency, and failures. Implement Unity Catalog or equivalent tools for centralized metadata management, data lineage, and governance policy enforcement. Enforce data security best practices including row-level security, encryption at rest/in transit, and fine-grained access control via Unity Catalog. Design and implement data masking, tokenization, and anonymization for compliance with privacy regulations (e.g., AI/ML-Ready Data Foundation: Enable data scientists by delivering high-quality, feature-rich data sets for model training and inference. Support AIOps/MLOps lifecycle workflows using MLflow for experiment tracking, model registry, and deployment within Databricks. Collaborate with AI/ML teams to create reusable feature stores and training pipelines. Cloud Data Architecture and Storage: Architect and manage data lakes on Azure Data Lake Storage (ADLS) or Amazon S3, and design ingestion pipelines to feed the bronze layer. Build data marts and warehousing solutions using platforms like Databricks. Optimize data storage and access patterns for performance and cost-efficiency. Maintain technical documentation, architecture diagrams, data dictionaries, and runbooks for all pipelines and components. Provide training and enablement sessions to internal stakeholders on the Databricks platform, Medallion Architecture, and data governance practices. Track deliverables against roadmap milestones and communicate risks or dependencies.
Hands-on experience with Databricks, Delta Lake, and Apache Spark for large-scale data engineering. Deep understanding of ELT pipeline development, orchestration, and monitoring in cloud-native environments. Experience implementing Medallion Architecture (Bronze/Silver/Gold) and working with data versioning and schema enforcement in enterprise grade environments. Strong proficiency in SQL, Python, or Scala for data transformations and workflow logic. Proven experience integrating enterprise platforms (e.g., PeopleSoft, Salesforce, D2L) into centralized data platforms. Familiarity with data governance, lineage tracking, and metadata management tools.
Experience with Databricks Unity Catalog for metadata management and access control. Experience deploying ML models at scale using MLFlow or similar MLOps tools. Familiarity with cloud platforms like Azure or AWS, including storage, security, and networking aspects. Knowledge of data warehouse design and star/snowflake schema modeling.
... troubleshoot Azure resources, pipelines, and deployments for reliability, scalability, and cost efficiency. Stay current with Azure DevOps, IaC, and cloud engineering best practices to drive continuous improvement. What You'll Bring: Strong 5+ years of experience as a DevOps Engineer or Cloud Infrastructure Engineer in ...
... backup/restore, disaster recovery strategies, capacity planning, and performance testing for critical services. 6+ years of experience in DevOps, SRE, Platform Engineering, Systems Engineering, or Software Engineering with significant cloud operations/automation ownership. - Strong hands-on experience designing and operating ...
... right in to help our clients navigate their next in their digital transformation journey this is the place for you Technology AI Generative AI Generative AI for Data Analytics Technology Full stack Net Full stack Technology Reactive Programming react JS A b i l i t y t o w r i t e t e s t c a s e s a n d s c e n a r i o s ...
Nivid Construction Pvt Ltd (NCPL), formerly known as Pioneer Infraprojects, established its headquarters in Ahmedabad, Gujarat, in 2022. Building on over a decade of experience, the company stands as a trusted partner for industrial construction, delivering high-quality, complex projects across Gujarat.
We are seeking a Senior Software Engineer who is ready to design, develop, and deliver robust, scalable, and innovative solutions across the full technology stack. The ideal candidate possesses deep expertise in both front end and back-end development, coupled with proficiency in database design and administration, and ...
... innovation, precision targeting, and impactful advertising that elevates brands to new heights. Team members join a forward-thinking environment centered on data-driven decision-making and measurable outcomes. Role Description The Agency Sales Manager is a full-time, remote role responsible for driving revenue growth through ...
... In specific locations, the pay range may vary from the range posted. JOB FAMILY: Sales Force DIVISION: EPD Established Pharma LOCATION: INDIA BIHAR PATNA : Remote ADDITIONAL LOCATIONS: WORK SHIFT: Standard TRAVEL: Yes, 100 % of the Time MEDICAL SURVEILLANCE: Not Applicable SIGNIFICANT WORK ACTIVITIES: Continuous standing ...
... In specific locations, the pay range may vary from the range posted. JOB FAMILY: Sales Force DIVISION: EPD Established Pharma LOCATION: INDIA BIHAR PATNA : Remote ADDITIONAL LOCATIONS: WORK SHIFT: Standard TRAVEL: Yes, 100 % of the Time MEDICAL SURVEILLANCE: Not Applicable SIGNIFICANT WORK ACTIVITIES: Continuous standing ...
Responsibilities Should have worked with AWS, Dockers and Kubernetes. Should have worked with a scripting language. Should know how to monitor system performance, CPU, Memory. Should be able to do troubleshooting. Should have knowledge of automated deployment Proficient in one programming knowledge - python preferred.
... and development Data quality validation and monitoring Schema design for analytics and reporting Performance optimization and scalability Preferred Experience Databricks Delta Lake experience Azure Synapse Analytics Python for data engineering Spark SQL optimization Real-time data streaming Data governance and metadata management ...
... real-time decisioning. Solid engineering foundation with 5+ years of relevant experience in data engineering and/or distributed systems. Hands-on work building data pipelines (batch and/or streaming) using at least one distributed framework (e.g., Comfortable with SQL and a transformation framework (e.g., Familiarity with ...
... Bachelor's degree (preferably in Computer Science, Mathematics, Data Science, or a related field). - 3-6 years of overall experience in software development and data analysis. - Hands-on experience in Python for application development. - Strong proficiency in SQL for data analysis and query generation. - Experience with hosted ...
... AWS Glue, Lambda, and DBT. Work with PySpark and SQL to transform and cleanse data. Configure and manage Airflow workflows and Amazon EMR clusters. Integrate data services with API Gateway and REST/Soap APIs. Required qualifications - 3 to 5 years of experience in data engineering. - Proficient in AWS S3, Glue, Lambda, ...
... Posted 2 weeks ago Job Be an early applicant About the job Job Purpose: Design and implement scalable data engineering and data warehouse solutions while managing data schemas, SQL query tuning, and code reviews. Who You Are: - 5+ years of experience in Data Engineering, with strong knowledge of Data Platforms and Data Warehousing ...
... teams including data analysts and engineers. Write reusable, efficient, and well-documented code. Integrate APIs and third-party services. Troubleshoot and debug data-related issues. Required Skills Strong proficiency in Python. Experience with data libraries like Pandas, NumPy. Hands-on experience with SQL and relational databases. ...
... data quality, reliability, and performance across data systems Required Skills: - Scala (must-have), python - Apache Spark - Strong coding fundamentals and engineering principles - Azure or on-premises Hadoop ecosystem experience - 5+ years of experience in Data Engineering - Python * SQL * AWS / Azure / GCP Preferred (Bonus) ...
... a related field (Master's degree preferred). Core Technical Skills: Python, PySpark, Databricks, Google ADK, LLMs, FastAPI, Spring Boot, Microservices, Kafka, SQL, Data Mesh, Starburst. Infrastructure and Cloud: Kubernetes, OpenShift, Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering ...
We are looking for a Data Engineer to join our team and play a key role in building, optimizing, and managing data pipelines. You will work closely with Data Analysts and Business Functions to ensure seamless data processing, reporting, and dashboarding. Our tech stack includes BigQuery, Snowflake, Airflow, Stitch/Fivetran, ...