Infosys
Senior Site Reliability Engineer
Hyderabad, India
Apply through hirly
Upload your resume and get a version tailored to this job, plus a cover letter, in about thirty seconds — before you create an account.
Apply with hirlyhirly's read of this role
- Role family
- Engineering
- Seniority
- Senior
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 27 Sept 2026
Derived automatically from the posting. Sign up to see how the role scores against your own resume.
the posting
We are looking for an experienced Site Reliability Engineer (SRE) / Production Engineer to build and operate highly available, scalable, and resilient production systems. The ideal candidate will have strong expertise in cloud infrastructure, automation, observability, incident management, performance engineering, and DevOps practices.
The role requires close collaboration with software engineering, infrastructure, security, and platform teams to ensure operational excellence and improve system reliability at scale.
Responsibilities
Reliability Engineering
Design, build, and maintain highly available and fault-tolerant production systems.
Define and monitor SLIs, SLOs, and SLAs for critical services.
Drive reliability improvements through automation and proactive engineering.
Conduct capacity planning and performance optimization activities.
Production Support & Operations
Manage production environments and ensure service uptime.
Lead incident response, troubleshooting, and root cause analysis (RCA).
Develop runbooks, operational playbooks, and disaster recovery procedures.
Technical requirements
Cloud & Infrastructure
Deploy and manage cloud-native infrastructure across AWS, Azure, or GCP.
Automate infrastructure provisioning using Infrastructure as Code (IaC).
Implement scalable and secure infrastructure solutions.
Support Kubernetes-based platforms and containerized workloads.
Observability & Monitoring
Build monitoring, logging, tracing, and alerting solutions.
Implement observability frameworks using industry-standard tools.
Monitor application health, performance metrics, and infrastructure utilization.
Drive continuous improvements in platform visibility and diagnostics.
Additional responsibilities
Automation & DevOps
Automate deployments, infrastructure management, and operational workflows.
Improve CI/CD pipelines and release processes.
Implement self-healing, auto-scaling, and operational automation solutions.
Promote DevOps and SRE best practices across engineering teams.
Security & Compliance
Ensure production environments meet security and compliance requirements.
Manage secrets, access controls, and vulnerability remediation.
Partner with security teams to implement security best practices.
Education
MCA,MSc,Bachelor of Engineering,BBA,BCom,BCS
Similar jobs
- Site Reliability EngineerNTT America, Inc. · Hyderabad, Telangana, IndiaFirst seen today
- Senior Site Reliability EngineerAkamai · IndiaFirst seen todayremote
- Site Reliability Engineer IIIAmerican Express · Bengaluru, KA, IndiaFirst seen today
- Lead Site Reliability Engineer - ObservabilitySimcorp · HyderabadFirst seen today
- Senior DevSecOps/Site Reliability Engineer (AWS)Stryker · Bengaluru, IndiaFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job