hirly

Encora10

Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)

Colombia

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Encora10 first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

hirly's read of this role

Role family
Engineering
Seniority
Senior
Country
CO
Work mode
Remote-friendly
First seen by hirly
30 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

Job Title: Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure)

Key Skills: Azure IaaS, Azure Monitor, Application Insights, New Relic, Databricks, DBT, SQL, Azure DevOps, GitHub Actions, Site Reliability Engineering (SRE), Application Performance Monitoring (APM), Log Analytics (KQL), Observability, Distributed Tracing, Structured Logging

Experience: 5+ years of experience in Site Reliability Engineering, Cloud Operations, or related roles. Mandatory minimum of 1 year of hands-on experience with DBT, Databricks, and SQL.

Location: Legal residents of Peru, Colombia, Bolivia, Costa Rica, Mexico, and Brazil.

Work Mode: Remote

At Coforge, we are looking for a Senior Site Reliability Engineer (SRE) – Application Observability & Readiness (Azure) (#21013-18-1) with the following profile.

Main Responsibilities

Collaborate with development teams to design and implement monitoring, alerting, dashboards, and APM instrumentation across applications and services.

Lead the implementation, configuration, and optimization of Application Performance Monitoring (APM) solutions.

Apply observability best practices using tools such as Azure Monitor, Application Insights, New Relic, and Log Analytics (KQL).

Enable code-level instrumentation, distributed tracing, and structured logging to improve application visibility and reliability.

Design and maintain application-level monitoring dashboards and operational health metrics.

Define and implement Service Level Indicators (SLIs), Service Level Objectives (SLOs), and effective alerting strategies based on latency, error rates, traffic, and resource saturation.

Continuously improve monitoring and alerting mechanisms through production insights and incident learnings.

Participate in production readiness reviews, identifying operational risks, observability gaps, and potential failure scenarios before deployment.

Support incident analysis and post-incident improvements through enhanced telemetry and monitoring practices.

Partner with engineering teams to ensure applications are reliable, scalable, and production-ready.

Mandatory Requirements

Strong experience supporting and operating applications in Microsoft Azure IaaS environments.

Hands-on experience with application observability, monitoring, and reliability engineering practices.

Mandatory experience with DBT, Databricks, and SQL (minimum 1 year of experience).

Experience implementing and managing APM solutions such as Application Insights, New Relic, or similar platforms.

Experience designing dashboards and monitoring solutions using Azure Monitor, Application Insights, and Log Analytics (KQL).

Familiarity with CI/CD environments including Azure DevOps and GitHub Actions.

Solid understanding of cloud-native architectures and distributed application systems.

Practical SRE mindset with experience in incident analysis, root cause investigation, and proactive problem prevention.

Strong verbal and written English communication skills, with the ability to collaborate effectively with global teams.

Preferred Requirements

Experience with scripting and automation using PowerShell and/or Bash.

Knowledge of scalability, availability, and resilience patterns in modern cloud environments.

Experience driving production readiness and operational excellence initiatives.

Exposure to reliability engineering best practices in enterprise-scale environments.

Posted on: 28-09-2026

At Coforge, we hire professionals solely based on their skills and qualifications and do not discriminate on the basis of age, disability, religion, gender, sexual orientation, socioeconomic status, or nationality.

Original posting on Encora10's site ↗

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job