Amazon
Systems Development Engineer, AWS Incident Response (AIR)
Dublin, IRL
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Mid level
- Country
- IE
- Work mode
- On-site / unstated
- First seen by hirly
- 27 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
AWS Incident Response (AIR) ensures the high availability of Amazon Web Services by making customer-impacting events shorter and less frequent through incident detection, management, and automated mitigation. Our systems monitor AWS infrastructure in real-time, automatically detect impairments, and orchestrate responses to minimize customer impact across regions and services.
As a Systems Development Engineer on the AIR team, you will lead the response to critical customer-impacting events — triaging impact, identifying root causes, coordinating mitigation actions with service teams, and driving resolution in real-time. Not every event is solved by automation; you will use your technical judgment to assess situations, engage the right teams, and direct mitigation strategies when manual intervention is required. Insights from these events directly inform the automation and tooling you build — creating a continuous improvement loop where each event makes the next one shorter or prevents it entirely.
This role offers a unique combination of systems development and real-time operational leadership, with direct impact on the availability of AWS services used by millions of customers.
- Key job responsibilities
- Drive the resolution of large-scale customer-impacting incidents as part of an on-call rotation (including weekends and holidays), leading incident calls and coordinating resolver teams across AWS service organizations
- Design, build, and enhance incident detection, triage, and mitigation automation tools
- Author COEs and event deep-dive documents to identify improvement opportunities; create and lead action items that improve processes, tooling, and automation
- Identify recurring platform issues and own projects that eliminate entire classes of operational problems
- Collaborate with teams globally to expand incident response capabilities across AWS regions and services
- A day in the life
- A Systems Development Engineer on the AWS Incident Response (AIR) team has full visibility on all AWS services! There are limitless opportunities to learn as you will work with all AWS internal teams and have exposure to AWS products and services.
When on-call, your day may start with large scale event — you join the conference bridge, assess the scope of impact using real-time dashboards, identify impaired services, engage the right teams, and drive mitigation until the event is resolved. After the event, you lead the deep-dive, document findings, and create action items to prevent recurrence.
When off-call, you spend your time building and improving the tools that make incident response faster and more automated. You might be writing code to improve event detection logic, building dashboards that surface the right signals during triage, or working on automation that reduces manual steps during mitigation. You participate in design / code reviews and collaborate with engineers across AIR to drive operational improvements. You also invest time in learning AWS service architectures — understanding how services fail helps you respond faster when they do.
- About the team
- AWS Incident Response (AIR) is a globally distributed team responsible for leading the large-scale customer-impacting events across AWS. We operate 24/7, providing incident leadership and coordination for events that span multiple services and regions. Our engineers combine hands-on incident leadership with systems development — we build the automation and tooling we use, and every event teaches us how to make the next one shorter or prevent it entirely. The team values operational excellence, continuous learning, and a bias for action. We work closely with service teams, networking, and infrastructure organizations across AWS, giving our engineers broad exposure to how AWS operates under the hood.
Basic qualifications
- - Knowledge of systems engineering fundamentals (networking, storage, operating systems)
- - Experience designing or architecting (design patterns, reliability and scaling) of new and existing systems
- - Experience in networking, storage systems, operating systems and hands-on systems engineering
- - Experience programming with at least one modern language such as C++, C#, Java, Python, Golang, PowerShell, Ruby
Preferred qualifications
- - Experience handling large enterprise technical customer escalations
- - Experience delivering results for large, cross-functional initiatives/projects, or experience in identifying incomplete or inaccurate data, identifying the root cause and creating/implementing an escalation plan
- - Experience in automating, deploying, and supporting large-scale infrastructure
Amazon is an equal opportunities employer. We believe passionately that employing a diverse workforce is central to our success. We make recruiting decisions based on your experience and skills. We value your passion to discover, invent, simplify and build. Protecting your privacy and the security of your data is a longstanding top priority for Amazon. Please consult our Privacy Notice ( https://www.amazon.jobs/en/privacy_page ) to know more about how we collect, use and transfer the personal data of our candidates.
Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.
Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.
Similar jobs
- Fluidics Development EngineerAnalogdevices · Ireland, LimerickFirst seen 2d ago
- Software Development EngineerCvshealth · IRL - GalwayFirst seen 5d ago
- Research & Development EngineerVertiv · Burnfoot, IrelandFirst seen 5d ago
- Product Development Engineer (R&D)Vertiv · Burnfoot, IrelandFirst seen 5d ago
- Senior Process Development EngineerZoetis · RathdrumFirst seen 4d ago
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job