Clearwater Analytics
Site Reliability Engineer II
Office - Chicago
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Engineering
- Seniority
- Mid level
- Stated salary
- $95,098 per year
- Country
- US
- Work mode
- On-site / unstated
- First seen by hirly
- 27 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
- About Us:
- At Clearwater, we are dedicated to provide world-class enterprise applications, ensuring their performance and availability to support our clients in the ever-evolving fintech landscape. Our Enterprise Application Support team plays a vital role in maintaining the smooth operation of critical business systems, driving innovation and excellence in our technology offerings.
- The Role:
- We are seeking an experienced and motivated individual to join our Enterprise Application Support team. In this role, you will be responsible for monitoring and maintaining the performance, availability, and stability of our enterprise-level applications. You will work collaboratively with cross-functional teams to ensure our systems run efficiently and effectively.
What You’ll Do:
Administer and manage monitoring tools such as Datadog, ELK, Grafana, and Prometheus.
Perform scheduled job monitoring and incident management, addressing issues as they arise.
Facilitate outage communication and coordination among stakeholders.
Utilize ITIL principles for effective problem management.
Monitor applications, servers, networks, databases, and storage systems.
Troubleshoot and identify issues within the IT infrastructure environment.
Demonstrate a solid understanding of Linux/Unix fundamentals, focusing on subsystems such as memory, storage, networks, CPU, disk, filesystems, and services.
Scripting knowledge in Linux/Unix is a plus.
Exhibit strong logical and analytical skills with a good understanding of IT infrastructure.
Be comfortable working in a 24/7 shift environment.
Manage infrastructure monitoring tools effectively, responding to alerts, performing initial triage, and escalating as necessary.
Maintain and update dashboards for server and network monitoring, including monitoring network traffic, bandwidth, hardware, uptime, and performance.
Prepare and maintain documentation and reports, providing follow-up status on identified tasks.
Create daily and weekly reports based on specified formats for designated recipients.
Implement and maintain standard escalation procedures in alignment with the communication plan, ensuring timely escalation and reporting of alerts according to defined roles and responsibilities.
What You’ll Need:
2 to 5 years of experience in application and infrastructure monitoring within a global enterprise environment.
Strong experience in the creation and fine-tuning of advanced alerts for monitored nodes.
Excellent documentation and communication skills.
A Bachelor’s degree in Computer Science or a related field, or an equivalent combination of education and experience.
Salary Range
$95,098.80 - $122,550.00 This is the pay range the Company believes it will pay for this position at the time of this posting. Consistent with applicable law, compensation will be determined based on relevant experience, other job-related qualifications/skills, and geographic location (to account for comparative cost of living). The Company reserves the right to modify this pay range at any time. For this role, benefits include: health/vision/dental insurance, 401(k), PTO, parental leave, and medical leave, STD/LTD insurance benefits. Clearwater Analytics is An Equal Opportunity/Affirmative Action Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability or veteran status, age or any other federally protected class.
Similar jobs
- Site Reliability Engineer IIRelx · 3 LocationsFirst seen today
- Site Reliability Engineer IIRelx · 3 LocationsFirst seen today
- Senior Site Reliability Engineer - US Federal (VDI & Infrastructure)Workday · USA.VA.RestonFirst seen today
- Senior Site Reliability Engineer (US Federal)Workday · USA.VA.RestonFirst seen today
- Principal Site Reliability Engineer, Infrastructure ObservabilityTroweprice · Owings Mills, MDFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job