Sre - site reliability engineer - remote in Kochi, India - Jobeax
Vacancy description
Sre - site reliability engineer - remote in Kochi, India
Velodata Global Pvt Ltd
RemoteWork from anywhere
India, Kochi
Sre - site reliability engineer - remote in Kochi, India is listed on Jobeax. Browse 30,000+ vacancies available.
We are seeking a Site Reliability Engineer (SRE) to support and maintain a 24×7 Azure cloud environment, ensuring high availability, reliability, and performance of infrastructure and hosted services. This role requires the engineer to operate across L1 and L2 support responsibilities, combining proactive monitoring with advanced troubleshooting and root cause analysis. The ideal candidate is an IT Generalist with strong networking, system administration, and customer service skills, capable of owning to customer issues end-to-end in dynamic and evolving environments. Azure Cloud Infrastructure Support (L1 & L2)
Provide 24X7 monitoring, support, and maintenance of Azure cloud infrastructure to ensure high availability, performance, security, and reliability.
Perform real-time monitoring and alert response using Azure Monitor, Log Analytics, Application Insights, and third-party monitoring tools.
Manage and support Azure Virtual Machines (Windows and Linux) including provisioning, scaling, start/stop, patching, backup, restore, and performance troubleshooting.
Support Azure networking components including Virtual Networks (VNets), Subnets, Network Security Groups (NSGs), User Defined Routes (UDRs), Load Balancers, Application Gateways, Azure Firewall, VPN Gateways, and ExpressRoute connectivity.
Support Azure Storage services (Blob, File, Disk, Queue, Table) including access control, performance tuning, capacity management, and issue resolution.
Provide L1/L2 support for Azure PaaS services such as App Services, Azure SQL, Managed Instances, and Azure Kubernetes Service (AKS), focusing on availability, connectivity, and configuration-related issues.
Perform capacity planning and performance analysis, proactively identifying resource constraints and recommending scaling or optimization actions.
Manage cost monitoring and optimization activities by identifying underutilized resources, supporting right-sizing efforts, and providing usage insights. System Administration & Remote Support
Manage users in Windows Terminal Server / Remote Desktop Services (RDS) environments, both onpremises and hosted.
Act as a Remote Administrator for Customer Windows Servers, including user account management, shared file access, print services, and print queue configuration.
Support Windows print services, server backup, restore, and recovery operations.
Perform routine OS patching, system maintenance, and health checks across Windows and Linux environments. Networking & Firewall Administration
Support and administer Fortinet / FortiGate firewalls, including policy management, site-to-site and dial-up VPN configuration, VLAN setup, and basic network troubleshooting.
Troubleshoot customer WAN, LAN, and VPN connectivity issues, including routing and firewall policy problems.
Diagnose and resolve local LAN and network-related issues impacting application and infrastructure availability.
Collaborate with internal and customer network teams to ensure secure and reliable connectivity. Remotely troubleshoot and resolve technical issues over phone and remote tools, taking full ownership from initial contact to resolution.
Communicate complex technical issues and solutions clearly to non-technical or less tech-savvy users.
Handle after-hours and emergency support requests on a rotational on-call basis.
Perform incident, problem, and change management activities in line with ITIL processes, including triage, escalation, RCA, and post-incident reviews.
Document incidents, resolutions, and troubleshooting steps in the call tracking / ticketing system in accordance with defined documentation standards.
Build strong working relationships with customers and internal teams to improve service quality and operational efficiency
Description Job Title: Senior Site Reliability Engineer Job Summary We are seeking a highly motivated and experienced Site Reliability Engineer to join our team. As a Site Reliability Engineer, you will be responsible for ensuring the reliability, scalability, and availability of our systems by leveraging your expertise ...
... lifecycle (Shift-Left Security). Research, recommend, and implement best practices for DevSecOps and Kubernetes operations. 5+ years of experience in DevOps, Site Reliability Engineering, or Platform Engineering roles. - 2+ years of hands-on Kubernetes experience, including cluster provisioning, scaling, and troubleshooting. ...
... monitoring of ML models in production Ensure infrastructure and services align with security guidelines and relevant regulatory standards Partner with data engineers and data scientists to make AI systems production-ready Foster a reliability-first culture across engineering teams Contribute to on-call rotations and continuously ...
Role: SRE & Deployments Engineer (SDE2) Function: Site Reliability Engineering / DevOps / Platform Engineering Type: Full-time Industry: Information Technology & Services, Computer Software, Fintech A Bengaluru-based enterprise tech startup founded in 2016. The company powers India's digital transformation through paperless, ...
... we're looking for lifelong learners and people who can make us better with their unique experiences. We're building a world where Identity belongs to you. The Engineering Opportunity We are seeking a Principal Site Reliability Engineer to serve as a technical leader for reliability engineering within Okta's Emerging Products ...
JOB TITLE: SENIOR DEVSECOPS ENGINEER JOB LOCATION: NAVI MUMBAI (Reliance Corporate Park) About the Role: As a Senior DevOps/SRE Engineer within the DevSecOps Engineering team at JioSaavn, you will be the driving force behind our application deployment strategies, platform reliability, and CI/CD pipelines across all product ...
... Engineer at Harness, you will play a pivotal role in designing, building, and maintaining our cloud infrastructure. You will be responsible for ensuring the reliability, scalability, and performance of our systems, incorporating a blend of Cloud Engineering and Site Reliability Engineering (SRE) practices. This role requires ...
... residential properties. It consists of over 30 well-known brands and nearly 8,900 properties situated in 141 countries and territories. Role Title: Senior Network Engineer I The Senior Network Engineer for DDI will be part of Network Site Reliability Engineering (SRE) project engineering team. The right candidate will be subject ...
... DFMEA, Encapsulation techniques, Manufacturing testing, and high-volume manufacturing processes and resulting design considerations. Knowledge in manufacturing engineering, quality engineering and reliability engineering. Seasoned in electronic product development processes & life cycle management. A role model in effective ...
... integrity of our corporate network by leveraging network security best practices, innovative products, and rigorous security validation. Reporting to the Network Engineering Manager, this operations-focused role is distinct from core Network Engineering and Network Security, centering primarily on operational execution—including ...
Role: Application Reliability Engineer Location: Gurgaon Who we are Graviton Research Capital is a privately funded quantitative trading firm striving for excellence in financial markets research. We trade across a multitude of asset classes and trading venues using a diverse range of concepts, from time series analysis ...
... alone. - Partner with ML and platform teams to keep large runs alive and serving latency predictable. What We're Looking For - 5+ years in infrastructure or site reliability engineering, including 2+ years operating GPU clusters at scale.* - Demonstrated on-call ownership of infrastructure that mattered, with a track record ...
... expertise with cloud infrastructure (AWS, GCP or Azure), containers and orchestration, CI/CD, monitoring stacks, automation - Strong experience in Site Reliability Engineering, DevOps, or Production Operations—preferably supporting large-scale systems - Solid understanding of incident management, reliability engineering, and microservice ...
A leading engineering and infrastructure solutions provider in India, we specialize in designing, deploying, and maintaining integrated building systems for commercial, industrial, and smart city projects. Our team delivers mission-critical ELV (Extra Low Voltage) systems including security, access control, fire alarms, ...
... Cisco experience, Reinvent applications, and Showcase the power of Cisco: our people, products, processes, systems, and data. You will be part of the Database Engineering and Site Reliability Engineering (SRE) team responsible for operating and maintaining Oracle Database platforms. Your Impact In this role, you will combine ...
... ensuring system reliability, security, and performance. This role is critical in providing a standardized, observability platform that gives both internal engineering teams and external, customer-facing services deep, reliable visibility into system health, performance, and reliability 3+ years of academic or work experience ...