Sre - site reliability engineer - remote in Kochi, India - Jobeax
Vacancy description
Sre - site reliability engineer - remote in Kochi, India
Velodata Global Pvt Ltd
RemoteWork from anywhere
India, Kochi
Sre - site reliability engineer - remote in Kochi, India is listed on Jobeax. Browse 30,000+ vacancies available.
We are seeking a Site Reliability Engineer (SRE) to support and maintain a 24×7 Azure cloud environment, ensuring high availability, reliability, and performance of infrastructure and hosted services. This role requires the engineer to operate across L1 and L2 support responsibilities, combining proactive monitoring with advanced troubleshooting and root cause analysis. The ideal candidate is an IT Generalist with strong networking, system administration, and customer service skills, capable of owning to customer issues end-to-end in dynamic and evolving environments. Azure Cloud Infrastructure Support (L1 & L2)
Provide 24X7 monitoring, support, and maintenance of Azure cloud infrastructure to ensure high availability, performance, security, and reliability.
Perform real-time monitoring and alert response using Azure Monitor, Log Analytics, Application Insights, and third-party monitoring tools.
Manage and support Azure Virtual Machines (Windows and Linux) including provisioning, scaling, start/stop, patching, backup, restore, and performance troubleshooting.
Support Azure networking components including Virtual Networks (VNets), Subnets, Network Security Groups (NSGs), User Defined Routes (UDRs), Load Balancers, Application Gateways, Azure Firewall, VPN Gateways, and ExpressRoute connectivity.
Support Azure Storage services (Blob, File, Disk, Queue, Table) including access control, performance tuning, capacity management, and issue resolution.
Provide L1/L2 support for Azure PaaS services such as App Services, Azure SQL, Managed Instances, and Azure Kubernetes Service (AKS), focusing on availability, connectivity, and configuration-related issues.
Perform capacity planning and performance analysis, proactively identifying resource constraints and recommending scaling or optimization actions.
Manage cost monitoring and optimization activities by identifying underutilized resources, supporting right-sizing efforts, and providing usage insights. System Administration & Remote Support
Manage users in Windows Terminal Server / Remote Desktop Services (RDS) environments, both onpremises and hosted.
Act as a Remote Administrator for Customer Windows Servers, including user account management, shared file access, print services, and print queue configuration.
Support Windows print services, server backup, restore, and recovery operations.
Perform routine OS patching, system maintenance, and health checks across Windows and Linux environments. Networking & Firewall Administration
Support and administer Fortinet / FortiGate firewalls, including policy management, site-to-site and dial-up VPN configuration, VLAN setup, and basic network troubleshooting.
Troubleshoot customer WAN, LAN, and VPN connectivity issues, including routing and firewall policy problems.
Diagnose and resolve local LAN and network-related issues impacting application and infrastructure availability.
Collaborate with internal and customer network teams to ensure secure and reliable connectivity. Remotely troubleshoot and resolve technical issues over phone and remote tools, taking full ownership from initial contact to resolution.
Communicate complex technical issues and solutions clearly to non-technical or less tech-savvy users.
Handle after-hours and emergency support requests on a rotational on-call basis.
Perform incident, problem, and change management activities in line with ITIL processes, including triage, escalation, RCA, and post-incident reviews.
Document incidents, resolutions, and troubleshooting steps in the call tracking / ticketing system in accordance with defined documentation standards.
Build strong working relationships with customers and internal teams to improve service quality and operational efficiency
... billions of transactions annually that move over $9 trillion around the globe. At FIS we’re passionate about building software that solves problems. We count on our site reliability engineers (SREs) to empower users with a rich feature set, high availability, and stellar performance level to pursue their missions. As we expand ...
... and infrastructure while always thinking about reliability, scalability, resilience, security, and https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to ...
... and infrastructure while always thinking about reliability, scalability, resilience, security, and https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to ...
... and infrastructure while always thinking about reliability, scalability, resilience, security, and https://jobeax.com/link/Im0ECCXXkw7iwTBY : - Help build a Site Reliability Engineering culture by sharing best practices, approaches, documentation, and code with other engineering teams.- Apply automation and software to ...
... an added advantage, along with knowledge and process in ITSM, ITIL. Precision in monitoring, alerting, and writing reliable code and continuous improvement. An SRE mindset geared toward reducing toil, automating repetitive tasks, and improving system reliability. Preferred Skills: Knowledge of Azure Cloud is an added advantage. ...
Role Overview We are hiring backend engineers who specialize in reliability . SRE at our company is a software engineering role — focused on designing, building, and improving highly available distributed systems through code. This is not a DevOps / CI-CD / Terraform-heavy role. What You'll Do Design and build reliable, ...
... container technologies including Docker, Kubernetes Proficiency in Python and Shell scripting to enable automation, platform operations, and continuous improvement Experience with monitoring, observability, Site Reliability Engineering (SRE), networking, governance, and DevSecOps principles within enterprise cloud environments
Full time Type Of Hire : Experienced (relevant combo of work and education) Education Desired : We are hiring a Senior Lead Site Reliability Engineer to define, build, and operate always-on, low-latency , and highly secure payment platforms that power large-scale financial transactions. This is a senior technical role, ...
... Markets - Wealth Management - Investment BankingLocation : Gurugram (Hybrid - 3 Days Office)Experience : 8+ YearsAbout the Role :We are seeking a highly skilled Site Reliability Engineer (SRE) to join a leading Asset Management and Investment Operations team. The ideal candidate will be responsible for ensuring application ...
... neutral infrastructure that enables organizations to safely embrace this new era. This is an opportunity to do career-defining work. As a Senior Site Reliability Engineer you will champion all things pertaining to reliability at Okta for Auth0. Working closely with the Product Engineers, Quality Engineers, Platform Engineers ...
Title: Senior Site Reliability Engineer - I, Product Area Focus Noida (Hybrid) Own availability, the most important product feature, by continually striving for sustained operational excellence of Sumo's planet-scale observability and security products. Work alongside your global SRE team, executing on projects in your ...
... scalability Leverage AI-assisted coding tools such as Claude Code or GitHub Copilot to improve productivity and code quality 5+ years of experience in DevOps, SRE, or infrastructure engineering roles (preferably in telecom, SaaS, or service provider environments) - Hands-on experience with Ansible and Terraform, including ...
... comfortable working in distributed teams and enjoy improving systems through thoughtful https://jobeax.com/link/cUAbOSDobkfnp7aS working as a DevOps, Platform, or Site Reliability Engineer Hands‑on experience with public cloud platforms such as AWS, Azure, or GCP Experience with containerization and orchestration technologies ...
... workflows for secure and compliant releases Champion cloud-native principles: immutability, declarative configuration, and microservices Collaborate with developers, SREs, and security teams to ensure seamless delivery Monitor and optimize pipeline performance, cost, and reliability across clouds Lead incident response and root ...
Description Job Title: Senior Site Reliability Engineer Job Summary We are seeking a highly motivated and experienced Site Reliability Engineer to join our team. As a Site Reliability Engineer, you will be responsible for ensuring the reliability, scalability, and availability of our systems by leveraging your expertise ...
... lifecycle (Shift-Left Security). Research, recommend, and implement best practices for DevSecOps and Kubernetes operations. 5+ years of experience in DevOps, Site Reliability Engineering, or Platform Engineering roles. - 2+ years of hands-on Kubernetes experience, including cluster provisioning, scaling, and troubleshooting. ...
... systems. If you thrive in a fast-paced environment and enjoy solving complex infrastructure challenges, we'd love to hear from you. Required Skills & Experience SRE & Production Operations (mandatory) - 5+ years in a Site Reliability Engineer or production operations engineering role, operating a SaaS or always-on service ...
... Performs peer reviews of specifications, designs, and code Identifies the technical debt & scaling issues in the basecode and drives the improvement Works alongside Site Reliability Engineers and cross functional teams to diagnose/troubleshoot any production performance related issues We work in Java, Golang, and Python. Our systems ...