RTA FLeet
Manager of Infrastructure, Cloud Operations & IT (Chaos Conductor)
Remote Worker
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.6M live jobs from 190,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Lead / management
- Work mode
- Remote-friendly
- First seen by hirly
- 26 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Description
Cloud infrastructure. DevOps. Security. Databases. IT. Production reliability. AI and automation.
If looking at that list makes you want to organize it rather than run from it, keep reading.
RTA is looking for a Manager of Infrastructure, Cloud Operations & IT to lead the technology foundation that keeps our people productive and our Fleet360 platform reliable, scalable, secure, and ready for what’s next.
This isn’t a narrow infrastructure management role. You’ll lead across Cloud Engineering, Infrastructure, DevOps, Database Engineering, Security Engineering, and IT operations—bringing these disciplines together around clear priorities, strong operating practices, and shared accountability.
We need someone who can operate at 30,000 feet without being afraid to come down to 30 feet. You can build the roadmap, communicate technical priorities and risk to senior leadership, develop a high-performing distributed team—and still jump into a production incident when your team needs you.
And yes, AI matters here.
We want a leader who is genuinely curious about where AI and automation can eliminate toil, improve observability, accelerate incident response, strengthen decision-making, and make a great technical team even better.
If you’re a technical operations leader who loves building reliable systems, developing people, solving hard problems, and making things better than you found them, you might be our kind of Chaos Conductor.
Your Top 3 Objectives
If you join RTA, these are the three things we’ll look back on and say, “Yep, that’s why we hired you.”
1. Build a Reliable, Scalable & Secure Technology Foundation
Own the operational health of RTA’s cloud infrastructure and the teams and systems that support it.
Continuously improve reliability, availability, scalability, observability, capacity planning, incident response, database resiliency, security posture, and operational readiness.
As RTA grows, our technology foundation needs to grow with us—without reliability, security, or performance becoming an afterthought.
2. Build & Lead a High-Performing Infrastructure and Operations Team
Create clarity, accountability, strong operating rhythms, and professional growth across Cloud Engineering, Infrastructure, DevOps, Database Engineering, Security Engineering, and IT.
You won’t just coordinate technical specialists. You’ll develop people, establish priorities, remove roadblocks, drive accountability, and help multiple technical disciplines operate as one team.
Because this is a remote role, you’ll know how to create connection, visibility, accountability, and strong communication across a distributed team.
3. Turn AI & Automation Into Measurable Operational Improvement
Move AI and automation from experimentation to actual business and operational value.
Reduce manual toil, improve monitoring and observability, accelerate incident response, strengthen root-cause analysis, automate repeatable work, and increase team productivity.
Develop and maintain an AI and automation roadmap for Infrastructure and IT Operations, prioritizing opportunities based on business impact, reliability, efficiency, risk, and measurable ROI.
You don’t need to be an ML engineer. You do need to experiment, learn quickly, challenge old ways of working, and turn emerging technology into meaningful operational improvement.
What You’ll Actually Own
- Cloud Infrastructure & Reliability
- Guide the architecture, performance, scalability, availability, monitoring, alerting, and capacity planning of RTA’s AWS environment. Use strong observability practices, predictive analytics, and intelligent capacity planning to identify potential issues before they become incidents.
- DevOps & Engineering Operations
- Strengthen CI/CD, infrastructure-as-code, deployment reliability, automation, observability, and the operational practices that help Engineering move quickly without sacrificing stability.
- Production Incident Management
- Provide calm, structured leadership when things go sideways. Improve incident response and resolution times, and ensure post-incident reviews turn lessons learned into lasting improvements.
- Infrastructure & Platform Operations
- Maintain a forward-looking view of RTA’s technical infrastructure, identify risks and capacity needs, and build roadmaps that support company and product growth.
- Database & Data Platform Reliability
- Guide database performance, resiliency, capacity planning, backups, recovery, tuning, and high availability.
- Security Engineering
- Guide vulnerability management, cloud security hardening, security engineering priorities, compliance initiatives, and incident-response preparedness. Ensure security risks are identified, prioritized, communicated, and addressed as RTA scales.
- IT Strategy & Oversight
- Set strategic direction for endpoint management, identity and access management, networking, internal tooling, and the employee technology experience while empowering the IT Manager to lead day-to-day execution.
- AI & Intelligent Automation
- Evaluate AI-powered tools, AIOps, intelligent monitoring, predictive analytics and scaling, automated remediation, AI-assisted development tools, automated runbooks, and other technologies that can improve reliability and team effectiveness.
- Hands-On Technical Leadership
- Stay close enough to the technology to jump in when needed—whether that’s troubleshooting an AWS configuration, reviewing a deployment issue, investigating a security alert, evaluating observability data, refining alerting thresholds, or helping the team through a difficult production incident.
- People Leadership
- Lead, coach, develop, and create accountability across a multidisciplinary technical team. Establish clear goals and operating rhythms and make sure people understand not only what matters, but why.
- Cross-Functional Partnership
- Work closely with Engineering, Product, Support, Security, and other teams so reliability, scalability, security, and operational considerations are built into decisions early—not discovered after something breaks.
- Executive Communication
- Translate complex technical topics, risks, investments, and tradeoffs into clear business language. Senior leaders shouldn’t need a decoder ring to understand what matters and what you recommend.
What We’re Looking For
Our ideal candidate is a hands-on technical operations leader who has grown beyond simply being the strongest technical person in the room.
You know the technology, but you also know how to lead people, establish operational discipline, prioritize competing demands, communicate with executives, and build systems that scale.
You’ll likely bring:
- 7+ years of progressive experience in cloud infrastructure, DevOps, SRE/platform engineering, infrastructure operations, or a closely related technical discipline.
- Meaningful hands-on experience supporting AWS production environments.
- Demonstrated experience leading and developing technical teams, including setting expectations, creating accountability, coaching performance, and helping strong technical people grow.
- Experience owning or significantly influencing production reliability, availability, monitoring, and incident management.
- Strong knowledge of modern DevOps, CI/CD, observability, infrastructure-as-code, and automation practices.
- Strong understanding of modern cloud architecture and software design principles, including microservices and how infrastructure decisions affect application reliability and scalability.
- Experience with technologies such as AWS, Docker, Terraform/CloudFormation, Grafana, Prometheus, Datadog, New Relic, Splunk, or comparable platforms.
- The ability to translate complex technical problems into clear priorities, risks, tradeoffs, and recommendations for business and executive stakeholders.
- Experience using AI and/or automation to improve technical, engineering, infrastructure, or operational workflows.
- Strong organizational and prioritization skil
Listed on hirly, a job board. hirly is not the employer: RTA FLeet is hiring for this role.
Similar jobs
- Category Manager InfrastructureLseg · London, United KingdomFirst seen today
- Associate Manager Infrastructure Services (Platform Engineering Lead )Dxctechnology · IND - KA - BANGALOREFirst seen today
- Associate Manager Infrastructure Services (Container/Linux)Dxctechnology · EGY - C - CAIROFirst seen today
- Manager/Senior Manager Infrastructure Planning & Management (Cruise)Sggovterp · STB - TOURISM COURT BUILDINGFirst seen 3d ago
- Senior Manager Infrastructure Support ManagerGlobalhr · US-CA-REMOTEFirst seen 4d agoremote
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job