Data Engineer SQL - Remote (m/w/d) in India, India - Jobeax
Vacancy description
Data Engineer SQL - Remote (m/w/d) in India, India
artech l.l.c.
RemoteWork from anywhere
India, India
Data Engineer SQL - Remote (m/w/d) in India, India is listed on Jobeax. Browse 30,000+ vacancies available.
Databricks Engineer Location - Remote
Data & AI Platform Engineering (Databricks-Centric): Design, implement, and optimize end-to-end data pipelines on Databricks, following the Medallion Architecture principles. Build robust and scalable ETL/ELT pipelines using Apache Spark and Delta Lake to transform raw (bronze) data into trusted curated (silver) and analytics-ready (gold) data layers. Apply schema evolution and data versioning to support agile data development. Platform Integration & Data Ingestion: Connect and ingest data from enterprise systems such as PeopleSoft, D2L, and Salesforce using APIs, JDBC, or other integration frameworks. Implement connectors and ingestion frameworks that accommodate structured, semi-structured, and unstructured data. Design standardized data ingestion processes with automated error handling, retries, and alerting. Data Quality, Monitoring, and Governance: Develop data quality checks, validation rules, and anomaly detection mechanisms to ensure data integrity across all layers. Integrate monitoring and observability tools (e.g., Databricks metrics, Grafana) to track ETL performance, latency, and failures. Implement Unity Catalog or equivalent tools for centralized metadata management, data lineage, and governance policy enforcement. Enforce data security best practices including row-level security, encryption at rest/in transit, and fine-grained access control via Unity Catalog. Design and implement data masking, tokenization, and anonymization for compliance with privacy regulations (e.g., AI/ML-Ready Data Foundation: Enable data scientists by delivering high-quality, feature-rich data sets for model training and inference. Support AIOps/MLOps lifecycle workflows using MLflow for experiment tracking, model registry, and deployment within Databricks. Collaborate with AI/ML teams to create reusable feature stores and training pipelines. Cloud Data Architecture and Storage: Architect and manage data lakes on Azure Data Lake Storage (ADLS) or Amazon S3, and design ingestion pipelines to feed the bronze layer. Build data marts and warehousing solutions using platforms like Databricks. Optimize data storage and access patterns for performance and cost-efficiency. Maintain technical documentation, architecture diagrams, data dictionaries, and runbooks for all pipelines and components. Provide training and enablement sessions to internal stakeholders on the Databricks platform, Medallion Architecture, and data governance practices. Track deliverables against roadmap milestones and communicate risks or dependencies.
Hands-on experience with Databricks, Delta Lake, and Apache Spark for large-scale data engineering. Deep understanding of ELT pipeline development, orchestration, and monitoring in cloud-native environments. Experience implementing Medallion Architecture (Bronze/Silver/Gold) and working with data versioning and schema enforcement in enterprise grade environments. Strong proficiency in SQL, Python, or Scala for data transformations and workflow logic. Proven experience integrating enterprise platforms (e.g., PeopleSoft, Salesforce, D2L) into centralized data platforms. Familiarity with data governance, lineage tracking, and metadata management tools.
Experience with Databricks Unity Catalog for metadata management and access control. Experience deploying ML models at scale using MLFlow or similar MLOps tools. Familiarity with cloud platforms like Azure or AWS, including storage, security, and networking aspects. Knowledge of data warehouse design and star/snowflake schema modeling.
... innovation, precision targeting, and impactful advertising that elevates brands to new heights. Team members join a forward-thinking environment centered on data-driven decision-making and measurable outcomes. Role Description The Agency Sales Manager is a full-time, remote role responsible for driving revenue growth through ...
... In specific locations, the pay range may vary from the range posted. JOB FAMILY: Sales Force DIVISION: EPD Established Pharma LOCATION: INDIA BIHAR PATNA : Remote ADDITIONAL LOCATIONS: WORK SHIFT: Standard TRAVEL: Yes, 100 % of the Time MEDICAL SURVEILLANCE: Not Applicable SIGNIFICANT WORK ACTIVITIES: Continuous standing ...
Responsibilities Should have worked with AWS, Dockers and Kubernetes. Should have worked with a scripting language. Should know how to monitor system performance, CPU, Memory. Should be able to do troubleshooting. Should have knowledge of automated deployment Proficient in one programming knowledge - python preferred.
... Bachelor's degree (preferably in Computer Science, Mathematics, Data Science, or a related field). - 3-6 years of overall experience in software development and data analysis. - Hands-on experience in Python for application development. - Strong proficiency in SQL for data analysis and query generation. - Experience with hosted ...
... and detail-oriented Associate Data Engineer to join our growing data team. If you are a recent graduate passionate about data and eager to build a career in data engineering, this is the perfect opportunity for you! In this role, you will work closely with senior engineers to build, maintain, and optimize data pipelines. ...
... tools to streamline data engineering workflows. Collaborate across teams to facilitate smooth data flow and integration. Enforce best practices in observability, data governance, security, and regulatory compliance Minimum 7 years as a Data Engineer or similar role. Hands-on experience with Databricks, Delta Lake, Spark, and ...
... and development Data quality validation and monitoring Schema design for analytics and reporting Performance optimization and scalability Preferred Experience Databricks Delta Lake experience Azure Synapse Analytics Python for data engineering Spark SQL optimization Real-time data streaming Data governance and metadata management ...
... real-time decisioning. Solid engineering foundation with 5+ years of relevant experience in data engineering and/or distributed systems. Hands-on work building data pipelines (batch and/or streaming) using at least one distributed framework (e.g., Comfortable with SQL and a transformation framework (e.g., Familiarity with ...
... AWS Glue, Lambda, and DBT. Work with PySpark and SQL to transform and cleanse data. Configure and manage Airflow workflows and Amazon EMR clusters. Integrate data services with API Gateway and REST/Soap APIs. Required qualifications - 3 to 5 years of experience in data engineering. - Proficient in AWS S3, Glue, Lambda, ...
... Posted 2 weeks ago Job Be an early applicant About the job Job Purpose: Design and implement scalable data engineering and data warehouse solutions while managing data schemas, SQL query tuning, and code reviews. Who You Are: - 5+ years of experience in Data Engineering, with strong knowledge of Data Platforms and Data Warehousing ...
... teams including data analysts and engineers. Write reusable, efficient, and well-documented code. Integrate APIs and third-party services. Troubleshoot and debug data-related issues. Required Skills Strong proficiency in Python. Experience with data libraries like Pandas, NumPy. Hands-on experience with SQL and relational databases. ...
... data quality, reliability, and performance across data systems Required Skills: - Scala (must-have), python - Apache Spark - Strong coding fundamentals and engineering principles - Azure or on-premises Hadoop ecosystem experience - 5+ years of experience in Data Engineering - Python * SQL * AWS / Azure / GCP Preferred (Bonus) ...
... a related field (Master's degree preferred). Core Technical Skills: Python, PySpark, Databricks, Google ADK, LLMs, FastAPI, Spring Boot, Microservices, Kafka, SQL, Data Mesh, Starburst. Infrastructure and Cloud: Kubernetes, OpenShift, Docker, Cloud-Native Infrastructure, CI/CD pipelines. Industry Context: Data engineering ...
... relocation Remote Type: This position is fully remote Commercial Acumen, Communication, Data Analysis, Data cleansing and transformation, Data domain knowledge, Data Integration, Data Management, Data Manipulation, Data Sourcing, Data strategy and governance, Data Structures and Algorithms (Inactive), Data visualization and ...
... pipelines with automated data quality checks to support business needs and demonstrate ownership of live data pipelines. Proficiency understanding and implementing data life cycles, data lineage, custom metadata, and data governance. Proficiency implementing BigQuery SQL procedures, functions, and similar, with actionable logging ...
Treebo Hospitality Ventures (THV) is a hotel group operating with a mission of 'democratising the joy of travel'. We focus on the affordable segments of the market — economy to mid-market — to cater to the next 100 million travellers who will experience the joy of travel over the next 10 years.
At GenY Medium, we don’t just market brands — we make them impossible to ignore. We’re a national digital marketing agency working with enterprise brands and fast-growing startups.
... designers and account managers, occasionally with other leaders Attending client meetings for new projects along with the team, asking the right questions to collect data to work on the project successfully Creating content maps, sitemaps, content inventory and wireframes and writing content primarily for websites, brochures, landing ...
Key Responsibilities Execute 2 brand campaigns/month, collaborating with design, video, and ORM teams Identify 5–10 trending posts/month for witty, brand-safe comment interventions Craft content across formats—social media, scripts, influencer assets, digital copy Collaborate on IP and content integration ideas; own