Senior Site Reliability Engineer - Kubernetes in Gurugram, India
GoStravvy
HybridMix of office and remote
India, Gurugram
Senior Site Reliability Engineer - Kubernetes in Gurugram, India is listed on Jobeax. Browse 30,000+ vacancies available.
Job Description : - Own production platform reliability end-to-end.- Kubernetes native multi-region platform.- IC role.- SLOs, observability & automation.- On-call & incident program.- Reliability of our EKS-based deployment platform - GitOps delivery.- Hands-on https://jobeax.com/link/mFvxkowVIupLYK84 Skillset : - Demonstrated expertise in managing large-scale AWS cloud infrastructure, with a deep understanding of EKS, networking, and security best practices.- Proven ability to author and maintain complex Terraform or OpenTofu modules, ensuring infrastructure is scalable, modular, and version-controlled.- Strong proficiency in Kubernetes orchestration, including troubleshooting complex cluster issues and optimizing resource utilization for high-traffic applications.- Advanced experience in observability and incident management, specifically using Datadog to monitor distributed systems and maintain strict Service Level Objectives (SLOs).- Exceptional communication skills, with the ability to articulate technical risks and architectural decisions to both engineering peers and non-technical stakeholders.- A proactive, problem-solving mindset with the ability to thrive in a hybrid work environment in Gurugram, balancing independent deep work with collaborative team https://jobeax.com/link/KJq9ExMBIu6MR34i Experience : 5 - 8 years (ref:hirist.tech)
... and infrastructure while always thinking about reliability, scalability, resilience, security, and https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to ...
... and infrastructure while always thinking about reliability, scalability, resilience, security, and https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to ...
... and infrastructure while always thinking about reliability, scalability, resilience, security, and https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to ...
... requirements for new telematics API capabilities and enhancements. Develop business cases and requirements to address recurring API issues and improve overall API reliability and customer experience. Lead and coordinate Field Follow Programs for new API endpoints, features, and functionality launches, ensuring successful adoption ...
Role Overview We are hiring backend engineers who specialize in reliability . SRE at our company is a software engineering role — focused on designing, building, and improving highly available distributed systems through code. This is not a DevOps / CI-CD / Terraform-heavy role. What You'll Do Design and build reliable, ...
Full time Type Of Hire : Experienced (relevant combo of work and education) Site Reliability Engineer – 4 - 6 Yrs – Pune Location FIS empowers the financial world with payment processing and banking solutions, including software, services and technology outsourcing. FIS’ more than 55,000 worldwide employees are passionate ...
... innovation to customers worldwide. We celebrate diversity and foster an inclusive environment, empowering our employees to be their authentic selves. The Devops Engineer would be an active member within the Voice Application Services team, responsible for providing automation and test support for the SW releases of Five9. Build ...
As a Senior DevOps Engineer , you will play a key role in designing, building, and operating cloud infrastructure and CI/CD platforms that support HERE's products and services. You will contribute to improving system reliability, automation, observability, and security while collaborating with cross‑functional engineering ...
We're looking for a seasoned Senior DevOps Engineer to lead cloud-native infrastructure initiatives across AWS, Azure and GCP. You'll architect scalable CI/CD pipelines using GitLab, manage containerized workloads with Kubernetes, Docker and Helm, and drive automation, security, and governance across multi-cloud environments. ...
Description Job Title: Senior Site Reliability Engineer Job Summary We are seeking a highly motivated and experienced Site Reliability Engineer to join our team. As a Site Reliability Engineer, you will be responsible for ensuring the reliability, scalability, and availability of our systems by leveraging your expertise ...
... neutral infrastructure that enables organizations to safely embrace this new era. This is an opportunity to do career-defining work. As a Senior Site Reliability Engineer you will champion all things pertaining to reliability at Okta for Auth0. Working closely with the Product Engineers, Quality Engineers, Platform Engineers ...
Title: Senior Site Reliability Engineer - I, Product Area Focus Noida (Hybrid) Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo's planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your ...
... DevSecOps and Kubernetes operations. 5+ years of experience in DevOps, Site Reliability Engineering, or Platform Engineering roles. - 2+ years of hands-on Kubernetes experience, including cluster provisioning, scaling, and troubleshooting. - Experience in containerization and orchestration: Docker and Kubernetes. - Solid ...
... enjoy solving complex infrastructure challenges, we'd love to hear from you. Required Skills & Experience SRE & Production Operations (mandatory) - 5+ years in a Site Reliability Engineer or production operations engineering role, operating a SaaS or always-on service at scale. - Proven hands-on experience defining and operating ...
... Performs peer reviews of specifications, designs, and code Identifies the technical debt & scaling issues in the basecode and drives the improvement Works alongside Site Reliability Engineers and cross functional teams to diagnose/troubleshoot any production performance related issues We work in Java, Golang, and Python. Our systems ...
We are seeking a Site Reliability Engineer (SRE) to support and maintain a 24×7 Azure cloud environment, ensuring high availability, reliability, and performance of infrastructure and hosted services. This role requires the engineer to operate across L1 and L2 support responsibilities, combining proactive monitoring with ...
... container technologies including Docker, Kubernetes Proficiency in Python and Shell scripting to enable automation, platform operations, and continuous improvement Experience with monitoring, observability, Site Reliability Engineering (SRE), networking, governance, and DevSecOps principles within enterprise cloud environments
Full time Type Of Hire : Experienced (relevant combo of work and education) Education Desired : We are hiring a Senior Lead Site Reliability Engineer to define, build, and operate always-on, low-latency , and highly secure payment platforms that power large-scale financial transactions. This is a senior technical role, ...
... Markets - Wealth Management - Investment BankingLocation : Gurugram (Hybrid - 3 Days Office)Experience : 8+ YearsAbout the Role :We are seeking a highly skilled Site Reliability Engineer (SRE) to join a leading Asset Management and Investment Operations team. The ideal candidate will be responsible for ensuring application ...