DevOps - Site Reliability Engineer - AWS in Bengaluru, India
KONG
India, Bengaluru
DevOps - Site Reliability Engineer - AWS in Bengaluru, India is listed on Jobeax. Browse 30,000+ vacancies available.
Are you ready to unlock intelligence
As an SRE 2 for Managed Gateways, you will be pivotal in ensuring the rock-solid reliability, scalability, and performance of Kong's critical managed services. Implement and maintain robust automation for deploying and operating Kong's Managed Gateways across various cloud environments.
Monitor system health, performance, and uptime, striving for 99.99% availability for our core infrastructure.
Resolve complex production incidents efficiently, participating actively in on-call rotations to maintain service continuity.
Contribute proactively to the prevention of technical debt, ensuring sustainable and scalable operations as Kong grows.
Collaborate closely with engineering teams to design, review, and implement resilient and highly scalable services.
2+ years of experience applying Site Reliability Engineering (SRE) principles and practices in a production environment.
Proficiency in at least one of Golang or Python for automation, tooling, and infrastructure as code.
Hands-on experience with Kubernetes and major cloud platforms such as AWS, GCP, or Azure.
Familiarity with monitoring, logging, and alerting tools (e.g., Prometheus, Grafana, Datadog).
Solid understanding of networking concepts, distributed systems, and API gateways.
Own the reliability and performance of critical production systems with a strong sense of accountability.
Drive urgent resolution of issues, demonstrating a bias for action and minimizing customer impact.
Bonus Points
Experience with Kong Gateway or other API management platforms.
Active contributions to open-source projects or developer communities.
the AI Connectivity Company, is building the connectivity layer of AI. Trusted by the Fortune 500® and AI-native startups alike, Kong's unified API and AI platform enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI traffic — on any model, any cloud.
... requirements for new telematics API capabilities and enhancements. Develop business cases and requirements to address recurring API issues and improve overall API reliability and customer experience. Lead and coordinate Field Follow Programs for new API endpoints, features, and functionality launches, ensuring successful adoption ...
Role Overview We are hiring backend engineers who specialize in reliability . SRE at our company is a software engineering role — focused on designing, building, and improving highly available distributed systems through code. This is not a DevOps / CI-CD / Terraform-heavy role. What You'll Do Design and build reliable, ...
Full time Type Of Hire : Experienced (relevant combo of work and education) Education Desired : We are hiring a Senior Lead Site Reliability Engineer to define, build, and operate always-on, low-latency , and highly secure payment platforms that power large-scale financial transactions. This is a senior technical role, ...
About the Role :Site Reliability Engineers are responsible for ensuring the availability, reliability, scalability, and performance of the firms most critical, customer-facing microservices that power all eCommerce channels. This role applies Google-inspired SRE principles to balance feature velocity and system reliability ...
Title: Senior Site Reliability Engineer - I, Product Area Focus Noida (Hybrid) Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo's planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your ...
... resolve and/or escalate to service teams Implement changes to enable or improve infrastructure resilience, monitoring, and alerting Experience - 3+ years as a Site Reliability Engineer or in a Cloud Operations/DevOps role - 2+ years using golang, shell scripting and terraform - 2+ years as software developer in a SaaS environment ...
... system reliability and performance. Document processes, configurations, and incident reports with clear and effective communication. 8+ years of experience in Site Reliability Engineering or DevOps role. - Strong experience with Cloud platforms, preferably Microsoft Azure or Google Cloud Platform (GCP). - Minimum 2 years ...
... DFMEA, Encapsulation techniques, Manufacturing testing, and high-volume manufacturing processes and resulting design considerations. Knowledge in manufacturing engineering, quality engineering and reliability engineering. Seasoned in electronic product development processes & life cycle management. A role model in effective ...
... integrity of our corporate network by leveraging network security best practices, innovative products, and rigorous security validation. Reporting to the Network Engineering Manager, this operations-focused role is distinct from core Network Engineering and Network Security, centering primarily on operational execution—including ...
Role: Application Reliability Engineer Location: Gurgaon Who we are Graviton Research Capital is a privately funded quantitative trading firm striving for excellence in financial markets research. We trade across a multitude of asset classes and trading venues using a diverse range of concepts, from time series analysis ...
A leading engineering and infrastructure solutions provider in India, we specialize in designing, deploying, and maintaining integrated building systems for commercial, industrial, and smart city projects. Our team delivers mission-critical ELV (Extra Low Voltage) systems including security, access control, fire alarms, ...
... expertise with cloud infrastructure (AWS, GCP or Azure), containers and orchestration, CI/CD, monitoring stacks, automation - Strong experience in Site Reliability Engineering, DevOps, or Production Operations—preferably supporting large-scale systems - Solid understanding of incident management, reliability engineering, and microservice ...
... investigate incidents, and troubleshoot deployment or production issues. - Maintain containerized services for availability, scalability, and operational reliability. - Partner with software engineers to improve release processes and resolve infrastructure challenges. - Apply DevOps practices covering automation, security, ...
... Ensure availability and use of special tools for Warranty & Customer support Ensure availability of Spares for 0km and Warranty service and efficiently manage site inventory required for maintenance and overhaul Ensure feedback of field experience at all stages to improve reliability of the products Challenge Customer’s ...
... testing of the platform takes place automatically. The individual is expected to collaborate and operate in parallel with the DevOps Engineers and Application Engineers. Mentoring and training other team members will be an expectation from this role. Responsibilities: The Senior Cloud Services Engineer utilizes software that ...
... accuracy, and quality across training data. - Apply your production engineering experience to help improve AI systems' understanding of real-world cloud and DevOps environments. What You Bring - 4+ years of professional experience in cloud infrastructure, DevOps, site reliability engineering, platform engineering, or a ...
... while enjoying the journey together. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. Team The Devops Engineer would be an active member within the Voice Application Services team, responsible for providing automation and test support for the SW releases of Five9. ...
We are seeking a DevOps Engineer to join our team. The role focuses on building and maintaining scalable, reliable, and secure cloud-based infrastructure. Key responsibilities Design and manage AWS infrastructure. Containerize applications using Docker and orchestrate with Kubernetes. Implement Infrastructure as Code using ...