Senior Site Reliability Engineer

bpinternational· MY0B BP Business Service Centre Asia Sdn Bhd
Apply Now ↗
🌍 Remote📍 Malaysia - Kuala LumpurFull time

About this role

Entity:

Technology


Job Family Group:

IT&S Group


Job Description:

Role Summary

As a Site Reliability Engineer, you will be responsible for improving the reliability, resilience and operational effectiveness of our technology platforms and services.


You will work closely with engineering and product teams to ensure systems are highly available, scalable, secure and supportable in production. You will use software engineering, automation and modern cloud practices to reduce manual effort, improve performance and strengthen production reliability.


Key Responsibilities

  • Improve the reliability, availability, performance and scalability of cloud-based applications and services.
  • Design and implement automation to reduce manual operational activities and improve engineering efficiency.
  • Build and improve monitoring, logging, alerting and observability across production systems.
  • Investigate complex production issues and drive improvements to prevent recurring failures.
  • Improve system resilience, recovery and operational readiness.
  • Build and improve CI/CD pipelines to enable reliable and repeatable software delivery.
  • Develop and maintain infrastructure using Infrastructure as Code and automation.
  • Identify reliability risks, operational gaps and technical debt and drive appropriate improvements.
  • Improve cloud infrastructure security and operational practices.
  • Develop reusable engineering patterns and mentor engineers across teams.

Required Experience and Qualifications

  • 7+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Engineering, Software Engineering or related technical disciplines, with strong experience operating production systems.
  • Strong understanding of cloud infrastructure security, including identity and access management, least privilege, network security, secrets management and secure configuration.
  • Understanding of security practices within CI/CD pipelines and Infrastructure as Code.
  • Able to independently investigate and resolve complex technical and production problems.
  • Strong communication and collaboration skills across engineering, product, security and operational teams.
  • Able to influence engineering practices, drive technical improvements and mentor other engineers.
  • Degree in Computer Science, Engineering or a related discipline, or equivalent professional experience.
  • Relevant cloud or engineering certifications are beneficial but not essential.

Technical Skills

  • Strong experience operating and improving production systems in cloud-based environments.
  • Strong troubleshooting skills across applications, infrastructure, networking and cloud services.
  • Experience managing system reliability, scalability, availability and performance.
  • Strong knowledge of monitoring, logging, alerting and production diagnostics.
  • Experience with incident investigation, root cause analysis and operational improvement.
  • Good understanding of distributed systems, resilience and recovery practices.

Software Engineering

  • Strong programming and scripting skills using Python, Ruby, Go or equivalent technologies.
  • Strong understanding of software engineering practices including source control, code review, automated testing and software delivery.
  • Strong experience designing, building and maintaining CI/CD pipelines.
  • Strong experience with deployment automation, release management and rollback or recovery practices.
  • Experience building automation, tooling and reusable engineering solutions.

Cloud Infrastructure

  • Strong hands-on experience with AWS, Microsoft Azure or equivalent cloud platforms.
  • Strong experience with Infrastructure as Code, using technologies such as Terraform, CloudFormation or equivalent.
  • Strong knowledge of Linux/Unix systems, networking and infrastructure troubleshooting.
  • Experience with containers and modern cloud application infrastructure.
  • Experience with observability technologies such as Prometheus, Grafana, OpenTelemetry or cloud-native equivalents.
  • Good understanding of cloud services including compute, networking, storage, databases, identity and messaging.

Skills That Set You Apart

  • Experience improving reliability and operational practices across multiple services or engineering teams.
  • Experience with automated recovery, resilience engineering or self-service platform capabilities.
  • Experience operating large-scale or highly available distributed systems.
  • Strong understanding of cloud-native engineering and modern operational practices.


About bp

At bp, we provide the following environment and benefits to you:

  • A company culture where we respect our diverse and unified teams, where we are proud of our achievements and where fun and the attitude of giving back to our environment are highly valued.
  • Possibility to join our social communities and networks
  • Learning opportunities and other development opportunities to craft your career path
  • Life and health insurance, medical care package


And many other benefits. We are an equal opportunity employer and value diversity at our company. We do not discriminate based on race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.


We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform crucial job functions, and receive other benefits and privileges of employment.


Travel Requirement

No travel is expected with this role


Relocation Assistance:

This role is not eligible for relocation


Remote Type:

This position is a hybrid of office/remote working


Skills:

Agility core practices, Agility core practices, Analytics, API and platform design, Business Analysis, Cloud Platforms, Coaching, Communication, Configuration management and release, Continuous deployment and release, Data Structures and Algorithms (Inactive), Digital Project Management, Documentation and knowledge sharing, Facilitation, Information Security, iOS and Android development, Mentoring, Metrics definition and instrumentation, NoSql data modelling, Relational Data Modeling, Risk Management, Scripting, Service operations and resiliency, Software Design and Development, Source control and code management {+ 4 more}

.


Legal Disclaimer:

We are an equal opportunity employer. We do not discriminate on the basis of protected characteristics like race, religion, color, sex, national origin, sexual orientation, veteran status or disability status. Individuals with an accessibility need may request an adjustment/accommodation related to bp’s recruiting process (e.g., accessing the job application, completing required assessments, participating in telephone screenings or interviews, etc.). If you would like to request an adjustment/accommodation related to the recruitment process, please contact us.

If you are selected for a position and depending upon your role, your employment may be contingent upon adherence to local policy. This may include pre-placement drug screening, medical review of physical fitness for the role, and background checks.

Frequently Asked Questions

Is the salary disclosed for the Senior Site Reliability Engineer position at bpinternational?
The salary for this Senior Site Reliability Engineer role at bpinternational is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Is the Senior Site Reliability Engineer job at bpinternational remote?
Yes, this Senior Site Reliability Engineer position at bpinternational is remote, with team members based in Malaysia - Kuala Lumpur. You can work from home or anywhere in the supported regions.
Is the Senior Site Reliability Engineer role at bpinternational full-time or part-time?
This is listed as a Full time position. It is posted as a Senior Site Reliability Engineer role in the MY0B BP Business Service Centre Asia Sdn Bhd department at bpinternational.
Which team or department does the Senior Site Reliability Engineer at bpinternational belong to?
This Senior Site Reliability Engineer position is part of the MY0B BP Business Service Centre Asia Sdn Bhd department at bpinternational. See the full job description for more information about the team structure and responsibilities.
How do I apply for the Senior Site Reliability Engineer position at bpinternational?
Click the "Apply Now" button on this page. You will be redirected to bpinternational's official application portal hosted on workday where you can submit your application directly.
When was the Senior Site Reliability Engineer job at bpinternational posted?
This Senior Site Reliability Engineer position at bpinternational was posted on Sep 28, 2026. Apply as soon as possible — early applications are often reviewed first.
Senior Site Reliability Engineer
bpinternational
Apply for this role ↗

You'll be redirected to bpinternational's official application page on Workday.