hirly

Bpinternational

Senior Site Reliability Engineer

Malaysia - Kuala Lumpur

Apply through hirly

Upload your resume and get a version tailored to this job, plus a cover letter, in about thirty seconds — before you create an account.

Apply with hirly

hirly's read of this role

Role family
Engineering
Seniority
Senior
Country
MY
Work mode
On-site / unstated
First seen by hirly
28 Sept 2026

Derived automatically from the posting. Sign up to see how the role scores against your own resume.

the posting

Entity:

  • Technology
  • Job Family Group:
  • IT&S Group
  • Job Description:

Role Summary

As a Site Reliability Engineer, you will be responsible for improving the reliability, resilience and operational effectiveness of our technology platforms and services.

You will work closely with engineering and product teams to ensure systems are highly available, scalable, secure and supportable in production. You will use software engineering, automation and modern cloud practices to reduce manual effort, improve performance and strengthen production reliability.

Key Responsibilities

  • Improve the reliability, availability, performance and scalability of cloud-based applications and services.
  • Design and implement automation to reduce manual operational activities and improve engineering efficiency.
  • Build and improve monitoring, logging, alerting and observability across production systems.
  • Investigate complex production issues and drive improvements to prevent recurring failures.
  • Improve system resilience, recovery and operational readiness.
  • Build and improve CI/CD pipelines to enable reliable and repeatable software delivery.
  • Develop and maintain infrastructure using Infrastructure as Code and automation.
  • Identify reliability risks, operational gaps and technical debt and drive appropriate improvements.
  • Improve cloud infrastructure security and operational practices.
  • Develop reusable engineering patterns and mentor engineers across teams.

Required Experience and Qualifications

  • 7+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Engineering, Software Engineering or related technical disciplines, with strong experience operating production systems.
  • Strong understanding of cloud infrastructure security, including identity and access management, least privilege, network security, secrets management and secure configuration.
  • Understanding of security practices within CI/CD pipelines and Infrastructure as Code.
  • Able to independently investigate and resolve complex technical and production problems.
  • Strong communication and collaboration skills across engineering, product, security and operational teams.
  • Able to influence engineering practices, drive technical improvements and mentor other engineers.
  • Degree in Computer Science, Engineering or a related discipline, or equivalent professional experience.
  • Relevant cloud or engineering certifications are beneficial but not essential.

Technical Skills

  • Strong experience operating and improving production systems in cloud-based environments.
  • Strong troubleshooting skills across applications, infrastructure, networking and cloud services.
  • Experience managing system reliability, scalability, availability and performance.
  • Strong knowledge of monitoring, logging, alerting and production diagnostics.
  • Experience with incident investigation, root cause analysis and operational improvement.
  • Good understanding of distributed systems, resilience and recovery practices.

Software Engineering

  • Strong programming and scripting skills using Python, Ruby, Go or equivalent technologies.
  • Strong understanding of software engineering practices including source control, code review, automated testing and software delivery.
  • Strong experience designing, building and maintaining CI/CD pipelines.
  • Strong experience with deployment automation, release management and rollback or recovery practices.
  • Experience building automation, tooling and reusable engineering solutions.

Cloud Infrastructure

  • Strong hands-on experience with AWS, Microsoft Azure or equivalent cloud platforms.
  • Strong experience with Infrastructure as Code, using technologies such as Terraform, CloudFormation or equivalent.
  • Strong knowledge of Linux/Unix systems, networking and infrastructure troubleshooting.
  • Experience with containers and modern cloud application infrastructure.
  • Experience with observability technologies such as Prometheus, Grafana, OpenTelemetry or cloud-native equivalents.
  • Good understanding of cloud services including compute, networking, storage, databases, identity and messaging.

Skills That Set You Apart

  • Experience improving reliability and operational practices across multiple services or engineering teams.
  • Experience with automated recovery, resilience engineering or self-service platform capabilities.
  • Experience operating large-scale or highly available distributed systems.
  • Strong understanding of cloud-native engineering and modern operational practices.

About bp

At bp, we provide the following environment and benefits to you:

  • A company culture where we respect our diverse and unified teams, where we are proud of our achievements and where fun and the attitude of giving back to our environment are highly valued.
  • Possibility to join our social communities and networks
  • Learning opportunities and other development opportunities to craft your career path
  • Life and health insurance, medical care package

And many other benefits. We are an equal opportunity employer and value diversity at our company. We do not discriminate based on race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform crucial job functions, and receive other benefits and privileges of employment.

Travel Requirement

  • No travel is expected with this role
  • Relocation Assistance:
  • This role is not eligible for relocation
  • Remote Type:
  • This position is a hybrid of office/remote working
  • Skills:

Agility core practices, Agility core practices, Analytics, API and platform design, Business Analysis, Cloud Platforms, Coaching, Communication, Configuration management and release, Continuous deployment and release, Data Structures and Algorithms (Inactive), Digital Project Management, Documentation and knowledge sharing, Facilitation, Information Security, iOS and Android development, Mentoring, Metrics definition and instrumentation, NoSql data modelling, Relational Data Modeling, Risk Management, Scripting, Service operations and resiliency, Software Design and Development, Source control and code management {+ 4 more} .

  • 

  • Legal Disclaimer:

We are an equal opportunity employer. We do not discriminate on the basis of protected characteristics like race, religion, color, sex, national origin, sexual orientation, veteran status or disability status. Individuals with an accessibility need may request an adjustment/accommodation related to bp’s recruiting process (e.g., accessing the job application, completing required assessments, participating in telephone screenings or interviews, etc.). If you would like to request an adjustment/accommodation related to the recruitment process, please contact us .

If you are selected for a position and depending upon your role, your employment may be contingent upon adherence to local policy. This may include pre-placement drug screening, medical review of physical fitness for the role, and background checks.

Original posting on Bpinternational's site ↗

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job