Infosys
Prometheus Engineer
Bangalore, India
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Mid level
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 29 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Responsibilities
Stack Management: Deploy, configure, and maintain the core observability stack using Prometheus, Grafana, Alertmanager, and Loki.
Dashboarding & Visualization: Collaborate with engineering teams to design and build comprehensive Grafana dashboards for application and infrastructure health monitoring.
Alerting Strategy: Configure and fine-tune Alertmanager rules to ensure accurate, actionable alerts while minimizing alert fatigue.
Log Management: Architect and manage centralized logging solutions using Loki to ensure efficient log aggregation and querying.
System Optimization: Monitor the performance of the observability stack itself, optimizing resource usage and scaling infrastructure as needed.
Continuous Improvement: Evaluate and integrate modern observability tools (such as VictoriaMetrics and VictoriaLogs) to enhance system correlation and analysis capabilities.
Experience: 9+ years of hands-on experience in DevOps, SRE, or Observability roles.
Core Stack: Deep technical expertise in Prometheus, Grafana, Alertmanager, and Loki (PLG Stack).
Automation: Proficiency in scripting (Python, Bash) and Infrastructure as Code (e.g., Terraform, Ansible).
Infrastructure & OS: Strong working knowledge of Linux/Unix administration.
Containerization: Experience monitoring containerized environments (Docker, Kubernetes).
Technical requirements
Hands-on knowledge or prior exposure to VictoriaMetrics (for scalable time-series data) and VictoriaLogs.
Experience with distributed tracing tools (e.g., Jaeger, Tempo, or VictoriaTraces).
Understanding of service level indicators (SLIs) and service level objectives (SLOs).
Additional responsibilities
Location of posting: Chennai, Bangalore, Hyderabad, Pune
Education
Bachelor of Engineering
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job