AWS Data Architect - Python/Apache Airflow in Hyderabad,… - Jobeax
Vacancy description
AWS Data Architect - Python/Apache Airflow in Hyderabad, India
Vision Excel Career Solutions
AWS Data Architect - Python/Apache Airflow in Hyderabad, India is listed on Jobeax. Browse 30,000+ vacancies available.
Role : AWS Data Architect Job Description :Candidate should possess a deep understanding of big data technologies, cloud services, and data architecture, with a proven track record of leading data-driven projects to successful https://jobeax.com/link/CEKJnNMNtDL19sum Responsibilities :- Lead a team of data engineers, providing technical guidance and mentorship. - Develop and execute a strategic roadmap for data processing, storage, and analytics in alignment with organizational goals.- Design, implement, and maintain robust data pipelines using Python and Airflow, ensuring efficient data flow and transformation for analytical and operational purposes.- Utilize AWS services, including S3 for data storage, Glue and EMR for data processing, and orchestrate data workflows that are scalable, reliable, and secure.- Implement real-time data processing solutions using Kafka, SQS, and Event Bridge, addressing high-volume data ingestion and streaming needs.- Oversee the integration of diverse systems and data sources through AppFlow, APIs, and other integration tools, ensuring seamless data exchange and connectivity.- Lead the development of data warehousing solutions, applying best practices in data modelling to support efficient data storage, retrieval, and analysis.- Continuously monitor, optimize, and troubleshoot data pipelines and infrastructure, ensuring optimal performance and scalability.- Ensure adherence to data governance, privacy, and security policies, implementing measures to protect sensitive data and comply with regulatory https://jobeax.com/link/COQih4IcZKKIeDR1 :- Bachelor's or master's degree in computer science, Engineering, or a related field.- 8-10 years of experience in data engineering, with at least 3 years in a leadership role.- Proficient in Python programming and experience with Airflow for workflow management.- Strong expertise in AWS cloud services, particularly in data storage, processing, and analytics (S3, Glue, EMR, etc.).- Experience with real-time streaming technologies like Kafka, SQS, and Event Bridge.- Solid understanding of API based integrations and familiarity with integration tools such as AppFlow.- Deep knowledge of data warehousing concepts, data modelling techniques, and experience in implementing largescale data warehousing solutions.- Demonstrated ability to lead and mentor technical teams, with excellent communication and project management skills. (ref:hirist.tech)
... Python libraries, SDKs, and automation components.- Automate data-contract-driven workflows and platform guardrails.- Build and maintain Terraform modules for AWS, EKS, Airflow, and Databricks.- Manage Harness CI/CD pipelines for infrastructure and application workloads.- Support Kubernetes/EKS, Airflow, Databricks/Unity ...
... and data-driven solutions that optimize operations and drive profitability. As Google Cloud's largest customer, the organization manages over 170 petabytes of data in BigQuery and operates within a centralized data architecture. Design and develop full stack applications using Java 17+ , Spring Boot 3+ , Python 3+ , and ...
... Requirements : - 4 - 6 years of experience in data engineering or a closely related field, with strong hands-on production experience.- Expert-level SQL and strong Python skills, with experience writing production-grade, maintainable, and well-tested code.- Strong hands-on experience with Apache Airflow or a comparable workflow ...
... optimization, contributing to the advancement of our AI capabilities. Key Responsibilities: Design, develop, and implement generative AI applications using Python and GCP's AI/ML services (e.g., Vertex AI, GenAI APIs, Google Kubernetes Engine, Cloud Functions, BigQuery). Collaborate with product managers, data scientists, ...
... distributed data systems. Our technology stack includes AWS services, Airflow, Snowflake, MS SQL Server, and .NET. We are looking for someone who is passionate about data and infrastructure, comfortable making and communicating technical decisions, and interested in continuously improving our architecture, engineering practices, ...
... years of hands-on data engineering experience, ideally within financial services or data intensive environments. - Strong experience building and orchestrating data pipelines with Apache Airflow. • Solid working knowledge of Apache Spark for large-scale data processing. - Strong Python and SQL skills, with a focus on clean, ...
... at least 2 years in a technical leadership role overseeing data engineering or data platform teams. - Strong proficiency in big data technologies, including Apache Spark/PySpark, Apache Airflow and Apache Kafka. - Experience with programming languages such as Scala and Python for developing robust data pipelines. - In-depth ...
... Engineering: Architect Medallion (Bronze/Silver/Gold) or equivalent layered data platform designs using Delta Lake; own Unity Catalog governance, access control, and data lineage. Data Integration: Oversee ETL/ELT pipeline development and orchestration (Apache Airflow, ADF), including legacy migration support (e.g., SSIS, Informatica, ...
... datasets - Good understanding of data modeling, data quality, validation, and reconciliation - Strong troubleshooting and problem-solving skills Good to Have: Apache Spark / PySpark Apache Airflow Talend / SSIS / other ETL tools MongoDB Azure / AWS / GCP Azure Databricks / Delta Lake / Data Lake Kafka or other streaming technologies ...
... featuresfunctionalities. Experience in building or enhancing Automation Framework Hybrid, Data Driven, Keyword Driven etc. Hands on experience in CICD: MAVEN, JenkinsAWS Pipelines EC2 based testing Experience with AgileDevOps methodologies Domain Skills: Experience in Retail Domain preferably with Merchandise DataCommerce retail. ...
... JavaScript. - Knowledge of database systems (e.g., PostgreSQL, MySQL, MongoDB). - Familiarity with version control tools like Git. - Understanding of RESTful APIs and web services integration. - Experience with cloud services (e.g., AWS, Azure) is a plus. - Strong analytical and problem-solving skills, with attention to detail.
... position sizing, max daily loss limits, and 'kill-switch' functionality. Backtesting & Logging: Create a framework to test strategies against historical NSE data and maintain detailed logs for audit and debugging. Required Skills & Qualifications Language Proficiency: Strong expertise in Python (preferred for its libraries ...
... maintain Python applications for data extraction, transformation, automation, and reporting. Build and optimize data pipelines and ETL processes to handle large datasets from multiple sources. Develop reusable Python modules to streamline analytics and reporting workflows. Integrate APIs, databases, and external data sources ...
... using Terraform, Kubernetes, and AWS services such as RDS, MSK, and EC2. Deploy applications into AWS environments. Develop and maintain scripts and tools using Python to automate common tasks. Deploy and manage PostgreSQL and other databases to ensure high availability and data consistency. Deploy and manage Kafka clusters ...
... (PySpark/Scala) development, unit testing, and performance optimization. Strong Python programming skills using libraries such as pandas, requests, json, and awswrangler. Experience on Apache Kafka and Confluent Kafka. Experience designing and optimizing data lakes using Apache Iceberg, including compaction and Iceberg ...
... or equivalent with a strong software architecture mindset and adept in documentation of work Strong understanding and experience in real-time process control, data flow and multi-threaded process control architecture over a heterogeneous compute hardware architecture involving CPU and GPU and networked computers. Software ...
... GCP data engineer- Pyspark- Dataflow- Dataproc- BigQuery- AirflowKey Responsibilities :- Design, develop, and optimize endtoend data pipelines using PySpark, Dataflow, and Dataproc- Implement data ingestion, transformation, and loading processes into BigQuery- Orchestrate workflows and schedule jobs with Apache Airflow- ...
... with agentic frameworks such as AWS Strands, Google ADK, LangChain. Experience designing MCP-based integrations to connect AI agents with enterprise tools and data sources. Familiarity with prompt engineering principles, RAG architectures, or multi-agent orchestration patterns. What started as a humble, little aerial crop-dusting ...