... https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to any tasks or parts of the system which are performed manually.- Able to troubleshoot complicated, cross-platform ...
... https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to any tasks or parts of the system which are performed manually.- Able to troubleshoot complicated, cross-platform ...
... continuously improve operational excellence. Bachelor's degree in Computer Science, Engineering, or a related technical discipline. - 6+ years of experience in Site Reliability Engineering, infrastructure engineering, DevOps, systems engineering, or a closely related field. - Strong Linux system administration expertise and ...
... utilization for high-traffic applications.- Advanced experience in observability and incident management, specifically using Datadog to monitor distributed systems and maintain strict Service Level Objectives (SLOs).- Exceptional communication skills, with the ability to articulate technical risks and architectural decisions ...
... https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to any tasks or parts of the system which are performed manually.- Able to troubleshoot complicated, cross-platform ...
... Certificate Manager, ELB, EBS, ECS, CloudFront/WAF, SQS, SNS, SES. Expertise in tools like Prometheus, Grafana, AppDynamics, CloudWatch, and Thousand Eyes for system health monitoring and alerting. Hands-on experience with Docker and at least one Docker Container orchestration system like ECS or Kubernetes. Expertise with ...
Role Overview We are hiring backend engineers who specialize in reliability . SRE at our company is a software engineering role — focused on designing, building, and improving highly available distributed systems through code. This is not a DevOps / CI-CD / Terraform-heavy role. What You'll Do Design and build reliable, ...
Description Job Title: Senior Site Reliability Engineer Job Summary We are seeking a highly motivated and experienced Site Reliability Engineer to join our team. As a Site Reliability Engineer, you will be responsible for ensuring the reliability, scalability, and availability of our systems by leveraging your expertise ...
... infrastructure. Ensure that applications and websites run smoothly and efficiently. Work with software developers, engineers, and operations teams to improve system performance. What you will need to be successful (experience and qualifications) A bachelor's degree in computer science, engineering, or a related field or ...
... releases of Five9. Build and manage scalable cloud infrastructure on AWS and GCP Automate deployment, monitoring, and operational workflows across distributed systems Collaborate with Development, QA, and Operations teams to ensure seamless delivery and high system reliability Monitor system performance, troubleshoot issues, ...
As a Senior DevOps Engineer , you will play a key role in designing, building, and operating cloud infrastructure and CI/CD platforms that support HERE's products and services. You will contribute to improving system reliability, automation, observability, and security while collaborating with cross‑functional engineering ...
We're looking for a seasoned Senior DevOps Engineer to lead cloud-native infrastructure initiatives across AWS, Azure and GCP. You'll architect scalable CI/CD pipelines using GitLab, manage containerized workloads with Kubernetes, Docker and Helm, and drive automation, security, and governance across multi-cloud environments. ...
Title: Senior Site Reliability Engineer - I, Product Area Focus Noida (Hybrid) Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo's planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your ...
... career-defining work. As a Senior Site Reliability Engineer you will champion all things pertaining to reliability at Okta for Auth0. Working closely with the Product Engineers, Quality Engineers, Platform Engineers and Architecture teams, your primary focus will be on ensuring production systems remain operational at all times, while ...
MetLife is seeking a Site Reliability Engineer (SRE) to ensure the reliability, availability, and performance of critical applications and platforms. The SRE Engineer will monitor production systems, respond to incidents, improve observability, maintain runbooks, and automate operational tasks. Working closely with engineering, ...
... standards. Conduct site audits, prepare as-built documentation, and validate system functionality against project specifications and client requirements. Perform system diagnostics, fault isolation, and preventive maintenance for uninterrupted system performance across multiple project sites. Collaborate with design teams to ...
... lifecycle (Shift-Left Security). Research, recommend, and implement best practices for DevSecOps and Kubernetes operations. 5+ years of experience in DevOps, Site Reliability Engineering, or Platform Engineering roles. - 2+ years of hands-on Kubernetes experience, including cluster provisioning, scaling, and troubleshooting. ...
... maintenance, overhauling. Deliver warranty services from delivery of the product to the end of warranty period Support during the commissioning of the Product/System on train and Locomotive Execute all Field scheduled operations (retrofit, overhaul, maintenance etc.) Provide support to Engineering & Quality team for effective ...
Site Reliability Engineer (Private Cloud / Virtualization) Cisco is transforming its platforms to run the next generation of cloud-native and multi-cloud services. This role offers a superb opportunity to transform how infrastructure platforms are developed and managed with full software automation. This team is responsible ...