AirAsia
Senior Site Reliability Engineer
Wisma Capital A
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.6M live jobs from 190,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Engineering
- Seniority
- Senior
- Country
- MY
- Work mode
- On-site / unstated
- First seen by hirly
- 27 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Job Description
We are looking for a highly motivated Senior Site Reliability Engineer (SRE) to join our Platform Engineering team. In this role, you will design, build, and operate highly available, secure, and scalable cloud platforms while driving automation across the software delivery lifecycle.
You will partner closely with Engineering, Security, DevOps, and Product teams to improve platform reliability, developer productivity, operational excellence, and cloud governance. This is a hands-on engineering role requiring strong expertise in cloud infrastructure, Infrastructure as Code (IaC), GitOps, observability, and incident management.
Key Responsibilities
- Design, build, and operate highly available production platforms on Google Cloud Platform (GCP).
- Develop Infrastructure as Code (IaC) using Terraform to provision and manage cloud infrastructure.
- Implement and maintain GitOps workflows using Argo CD and GitLab.
- Build and enhance CI/CD pipelines using GitLab to enable secure, reliable, and automated software delivery.
- Develop automation solutions to eliminate repetitive operational tasks using scripting and APIs.
- Manage and optimize Cloudflare services including DNS, WAF, CDN, Load Balancing, Zero Trust, and security controls.
- Build and maintain observability platforms including monitoring, logging, alerting, tracing, dashboards, and SLO/SLI reporting.
- Drive platform reliability through proactive monitoring, capacity planning, performance tuning, resilience testing, and automation.
- Participate in an on-call rotation, troubleshoot production incidents, lead incident response, perform root cause analysis (RCA), and implement permanent corrective actions.
- Improve operational excellence by reducing toil through automation and self-service capabilities.
- Collaborate with development teams to improve application reliability, deployment strategies, and operational readiness.
- Ensure platform security by implementing infrastructure best practices, policy enforcement, secrets management, and least-privilege access.
- Create and maintain technical documentation, operational runbooks, and standard operating procedures.
- Mentor junior engineers and promote SRE best practices across engineering teams.
Required Qualifications
- Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent practical experience.
- 5+ years of experience in Site Reliability Engineering, Platform Engineering, Cloud Engineering, or DevOps.
- Strong hands-on experience with Google Cloud Platform (GCP).
- Strong experience with Terraform and Infrastructure as Code.
- Hands-on experience with GitLab CI/CD.
- Experience implementing GitOps using Argo CD.
- Experience managing Cloudflare services including DNS, WAF, CDN, and Load Balancing.
- Strong Linux administration and troubleshooting skills.
- Experience with container technologies including Docker and Kubernetes.
- Strong scripting skills using Bash, Python, or Go.
- Experience with monitoring, logging, and observability platforms.
- Experience with incident management, production support, and on-call operations.
- Excellent troubleshooting and root cause analysis skills.
- Strong communication and stakeholder management skills.
Preferred Qualifications
- Experience operating Kubernetes platforms such as GKE.
- Experience with service mesh technologies (Istio, Linkerd, or Envoy).
- Knowledge of SRE principles including SLIs, SLOs, Error Budgets, and Toil Reduction.
- Experience implementing platform security and DevSecOps practices.
- Experience with FinOps and cloud cost optimization.
- Experience with policy-as-code and infrastructure governance.
- Google Cloud Professional certifications are an advantage.
- Knowledge in API’s and gateways is added advantages
What Success Looks Like
Within your first 12 months, you will:
- Improve platform reliability and availability through automation and engineering improvements.
- Reduce operational toil by automating manual processes.
- Improve deployment reliability using GitOps and CI/CD best practices.
- Enhance observability with actionable monitoring and alerting.
- Strengthen platform security and operational governance.
- Enable engineering teams to deliver software faster and more reliably.
Listed on hirly, a job board. hirly is not the employer: AirAsia is hiring for this role.
Similar jobs
- Senior Site Reliability Engineer (Data Platform)Guidewire · Malaysia - Kuala LumpurFirst seen today
- Site Reliability Engineer III (Platform)Guidewire · Malaysia - Kuala LumpurFirst seen 5d ago
- Senior Site Reliability EngineerBpinternational · Malaysia - Kuala LumpurFirst seen 6d ago
- Senior Site Reliability EngineerGen Digital Inc. · MYS - Kuala LumpurFirst seen 6d ago
- Security Site Reliability EngineerUobgroup · The Gardens North TowerFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job