Hiring.Camp

Senior Site Reliability Engineer

Bpinternational

·

Today

Location
MY: Kuala Lumpur - Bangsar South Campus, Malaysia
Workplace
Remote
Type
Full-time
Department
Engineering
Seniority
Senior
Experience
7+ years
Closing date
Today
Source
Workday

Description

Entity:

Technology


Job Family Group:

IT&S Group


Job Description:

Role Summary

As a Site Reliability Engineer, you will be responsible for improving the reliability, resilience and operational effectiveness of our technology platforms and services.


You will work closely with engineering and product teams to ensure systems are highly available, scalable, secure and supportable in production. You will use software engineering, automation and modern cloud practices to reduce manual effort, improve performance and strengthen production reliability.


Key Responsibilities

  • Improve the reliability, availability, performance and scalability of cloud-based applications and services.
  • Design and implement automation to reduce manual operational activities and improve engineering efficiency.
  • Build and improve monitoring, logging, alerting and observability across production systems.
  • Investigate complex production issues and drive improvements to prevent recurring failures.
  • Improve system resilience, recovery and operational readiness.
  • Build and improve CI/CD pipelines to enable reliable and repeatable software delivery.
  • Develop and maintain infrastructure using Infrastructure as Code and automation.
  • Identify reliability risks, operational gaps and technical debt and drive appropriate improvements.
  • Improve cloud infrastructure security and operational practices.
  • Develop reusable engineering patterns and mentor engineers across teams.

Required Experience and Qualifications

  • 7+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Engineering, Software Engineering or related technical disciplines, with strong experience operating production systems.
  • Strong understanding of cloud infrastructure security, including identity and access management, least privilege, network security, secrets management and secure configuration.
  • Understanding of security practices within CI/CD pipelines and Infrastructure as Code.
  • Able to independently investigate and resolve complex technical and production problems.
  • Strong communication and collaboration skills across engineering, product, security and operational teams.
  • Able to influence engineering practices, drive technical improvements and mentor other engineers.
  • Degree in Computer Science, Engineering or a related discipline, or equivalent professional experience.
  • Relevant cloud or engineering certifications are beneficial but not essential.

Technical Skills

  • Strong experience operating and improving production systems in cloud-based environments.
  • Strong troubleshooting skills across applications, infrastructure, networking and cloud services.
  • Experience managing system reliability, scalability, availability and performance.
  • Strong knowledge of monitoring, logging, alerting and production diagnostics.
  • Experience with incident investigation, root cause analysis and operational improvement.
  • Good understanding of distributed systems, resilience and recovery practices.

Software Engineering

  • Strong programming and scripting skills using Python, Ruby, Go or equivalent technologies.
  • Strong understanding of software engineering practices including source control, code review, automated testing and software delivery.
  • Strong experience designing, building and maintaining CI/CD pipelines.
  • Strong experience with deployment automation, release management and rollback or recovery practices.
  • Experience building automation, tooling and reusable engineering solutions.

Cloud Infrastructure

  • Strong hands-on experience with AWS, Microsoft Azure or equivalent cloud platforms.
  • Strong experience with Infrastructure as Code, using technologies such as Terraform, CloudFormation or equivalent.
  • Strong knowledge of Linux/Unix systems, networking and infrastructure troubleshooting.
  • Experience with containers and modern cloud application infrastructure.
  • Experience with observability technologies such as Prometheus, Grafana, OpenTelemetry or cloud-native equivalents.
  • Good understanding of cloud services including compute, networking, storage, databases, identity and messaging.

Skills That Set You Apart

  • Experience improving reliability and operational practices across multiple services or engineering teams.
  • Experience with automated recovery, resilience engineering or self-service platform capabilities.
  • Experience operating large-scale or highly available distributed systems.
  • Strong understanding of cloud-native engineering and modern operational practices.


About bp

At bp, we provide the following environment and benefits to you:

  • A company culture where we respect our diverse and unified teams, where we are proud of our achievements and where fun and the attitude of giving back to our environment are highly valued.
  • Possibility to join our social communities and networks
  • Learning opportunities and other development opportunities to craft your career path
  • Life and health insurance, medical care package


And many other benefits. We are an equal opportunity employer and value diversity at our company. We do not discriminate based on race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.


We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform crucial job functions, and receive other benefits and privileges of employment.


Travel Requirement

No travel is expected with this role


Relocation Assistance:

This role is not eligible for relocation


Remote Type:

This position is a hybrid of office/remote working


Skills:

Agility core practices, Agility core practices, Analytics, API and platform design, Business Analysis, Cloud Platforms, Coaching, Communication, Configuration management and release, Continuous deployment and release, Data Structures and Algorithms (Inactive), Digital Project Management, Documentation and knowledge sharing, Facilitation, Information Security, iOS and Android development, Mentoring, Metrics definition and instrumentation, NoSql data modelling, Relational Data Modeling, Risk Management, Scripting, Service operations and resiliency, Software Design and Development, Source control and code management {+ 4 more}

.


Legal Disclaimer:

We are an equal opportunity employer. We do not discriminate on the basis of protected characteristics like race, religion, color, sex, national origin, sexual orientation, veteran status or disability status. Individuals with an accessibility need may request an adjustment/accommodation related to bp’s recruiting process (e.g., accessing the job application, completing required assessments, participating in telephone screenings or interviews, etc.). If you would like to request an adjustment/accommodation related to the recruitment process, please contact us.

If you are selected for a position and depending upon your role, your employment may be contingent upon adherence to local policy. This may include pre-placement drug screening, medical review of physical fitness for the role, and background checks.

Skills

PythonRubyAWSAzureTerraformCI/CDLinuxiOSAndroidDevOpsRisk ManagementProject Management

Similar Jobs

30

Senior Platform Engineer

Two Six Technologies·Arlington, Virginia +1·Onsite

2d ago

Senior Platform Engineer

ACI Worldwide·Co. Limerick, IE·Hybrid

2d ago

Senior Platform Engineer

Accenturefederalservices·Hill AFB, UT +1·Onsite

3d ago

Senior Platform Engineer

Encora·Peru

3d ago

Senior Platform Engineer

Damia Group·Porto·Remote, Hybrid, Onsite

4d ago

Senior Platform Engineer

Drawbridge Partners·Remote

5d ago

Senior Platform Engineer

Defenseunicorns·United States - Remote·Remote

5d ago

Senior SRE

Banyansoftware·Canada, US +2·Remote

6d ago

Senior Platform Engineer

Unum·Dorking, Surrey +1

6d ago

Senior Platform Engineer

Recordedfuture·Remote - USA·Remote

6d ago

Senior Platform Engineer

Vultr·Remote - United States·Remote

1w ago

Senior Platform Engineer

Mintlify·San Francisco

1w ago

Senior Platform Engineer

Draftkings·Remote - US, US·Remote

1w ago

Senior Platform Engineer

Sydecar·San Francisco Office - Hybrid·Hybrid

1w ago

Senior Platform Engineer

H&M Group·Stockholm, Stockholms län

1w ago

Senior Platform Engineer

Avaloq·Fort Lauderdale, Florida·Hybrid

1w ago

Senior Platform Engineer

Mastercard·Pune, India

1w ago

Senior Platform Engineer

Westpacnz·Westpac on Takutai Square, New Zealand·Hybrid

1w ago

Senior Platform Engineer

Takeaway·Bristol Office, UK

1w ago

Senior Platform Engineer

Earnin·Mountain View, US +1·Hybrid, Onsite

1w ago

Senior Engineer (Platform)

Later·Boston, MA +3·Remote

1w ago

Senior Platform Engineer

ITV·London, GB

1w ago

Senior Platform Engineer

Nikohealth·New York, NY

1w ago

Senior DevOps

Damia Group·Lisbon·Remote, Hybrid, Onsite

1w ago

Senior Platform Engineer

LSPedia·Farmington Hills, MI

2w ago

Senior Platform Engineer

DoiT·Remote EMEA·Remote

2w ago

Senior Platform Engineer

DoiT·Remote EMEA·Remote

2w ago

Senior Platform Engineer

DoiT·Remote Ireland·Remote

2w ago

Senior Platform Engineer

DoiT·Remote EMEA·Remote

2w ago

Senior Platform Engineer

DoiT·Remote Estonia·Remote

2w ago
Remote Senior Site Reliability Engineer at Bpinternational | Hiring.Camp