Hiring.Camp

Site Reliability Engineer (SRE)

Acquireai

·

May 5, 2026

Location
Taguig City, Philippines
Type
Full-time
Department
Engineering
Source
Workday

Description

We’re an award-winning global outsourcer providing contact center and back office services on behalf of our global clients. Come work at a place where innovation and teamwork come together to support the most exciting missions in the world!

Role objective
The Site Reliability Engineer serves as the guardian of our production systems, ensuring the reliability, scalability, and performance of our IoT telemetry platform. You will define and enforce Service Level Objectives (SLOs), automate operational processes, and build the infrastructure and tooling that enables our engineering teams to deploy with confidence. By implementing comprehensive monitoring, incident response procedures, and reliability practices, you will play a pivotal role in maintaining the uptime and data freshness that our customers depend on for their critical fleet operations.

The role will focus on the following key areas:
SLO Management
Infrastructure Automation
Incident Response
Security & compliance


Key Responsibilities
Responsibilities of the Site Reliability Engineer will include but are not limited to:
Service Level Management & Reliability
• Define, monitor, and enforce Service Level Objectives (SLOs) and error budgets across
all production systems
• Track error budget burn rates and make data-driven decisions to halt risky
deployments when thresholds are exceeded
• Implement comprehensive monitoring and alerting strategies using Prometheus,
Grafana, and PagerDuty
• Establish and maintain reliability standards that support business-critical uptime
requirements
Infrastructure Automation & Management
• Design and implement Infrastructure as Code (IaC) solutions using Pulumi with
TypeScript
• Manage and optimize AWS services including EKS (Elastic Kubernetes Service), MSK
(Managed Streaming for Kafka), SingleStore, MongoDB S3
• Automate operational processes to eliminate toil, targeting any task that consumes
more than 2 engineer-days per quarter
Incident Response & Post-Mortem Leadership
• Serve as incident commander during production outages and service degradations
• Lead comprehensive post-mortem processes within 48 hours of incidents
• Drive "never-again" corrective actions to completion, ensuring systemic improvements
• Maintain and improve incident response procedures and runbooks
Security & Compliance
• Implement and enforce least-privilege IAM policies across all AWS resources
• Manage security patch pipelines and vulnerability remediation processes
• Support compliance initiatives including SOC2 and ISO 27001 certification requirements
• Ensure security best practices are embedded in all infrastructure and operational
procedures
On-Call & Operational Excellence
• Participate in follow-the-sun on-call rotation with one week primary/secondary
commitment every five weeks
• Provide 24×7 support coverage across AU/NZ, EU/ZA, and MX time zones
• Maintain operational runbooks and knowledge transfer documentation
• Continuously improve on-call experience and reduce alert fatigu

Join the A-Team and experience the A-Life!

Skills

TypeScriptAWSKubernetesMongoDBComplianceISO 27001

Similar Jobs

30

Site Reliability Engineer

Experian · Cyberjaya, Selangor, Malaysia · Hybrid

Yesterday

Site Reliability Engineer

Cisco · USA-SAN FRANCISCO, United States of America +11 · Remote

Yesterday

Site Reliability Engineer

LSEG · ROU-Bucharest-Iuliu Maniu Boulevard, Romania

Yesterday

Site Reliability Engineer

Rbs · Bengaluru, India

Yesterday

Site Reliability Engineer

Huntington · Easton Ops Cols C Oh, United States of America +10 · Onsite

4 days ago

Site Reliability Engineer

Westpac Group · Sydney, NSW, Australia

5 days ago

Site Reliability Engineer

Professional Kyndryl · PELML Lima (PELML) La Molina, Peru · Remote

5 days ago

Site Reliability Engineer

"Arena Intelligence, Inc." · Bay Area · Hybrid

5 days ago

Site Reliability Engineer

Damia Group · Lisbon, Coimbra, Braga · Remote, Hybrid, Onsite

5 days ago

Site Reliability Engineer

Accenture · Taguig, Uptown Bonifacio Tower 3, Philippines

5 days ago

Site Reliability Engineer

NielsenIQ · Mexico City, MEX, Mexico

5 days ago

Site Reliability Engineer

Tinybird · Spain · Remote

5 days ago

Site Reliability Engineer

Cisco · USA-RESEARCH TRIANGLE PARK, United States of America +4 · Hybrid

6 days ago

Site Reliability Engineer

Miqdigital · Bengaluru, India

6 days ago

Site Reliability Engineer

Forwardnetworks · Santa Clara, CA +1

6 days ago

Site Reliability Engineer

Bosonai · Toronto · Remote

6 days ago

Site Reliability Engineer

Experian · Cyberjaya, Selangor, Malaysia · Hybrid

6 days ago

Site Reliability Engineer

CRC Careers · Charlotte NC - 600 S Tryon St., United States of America

1 week ago

Site Reliability Engineer

Citi Bank · 5900 HURONTARIO STREET MISSISSAUGA, Canada · Hybrid

1 week ago

Site Reliability Engineer

HostPapa · Remote

1 week ago

Site Reliability Engineer

citibank · Mississauga, ON,CA, CA · Hybrid

1 week ago

Site Reliability Engineer

CRC Careers · CRC - Charlotte, NC 600 S. Tryon St., United States of America

1 week ago

Site Reliability Engineer

LSEG · Taipei - Nan Shan Plaza, Taiwan

1 week ago

Site Reliability Engineer

Careers Home · Zagreb (Croatia) +4 · Hybrid

1 week ago

Site Reliability Engineer

Braiins · Braiins · Hybrid

1 week ago

Site Reliability Engineer

Acronis · Serbia +2 · Remote

1 week ago

Site Reliability Engineer

Earnin · Mountain View, US +1 · Hybrid, Onsite

1 week ago

Site Reliability Engineer

Quberesearchandtechnologies · Hong kong +1

1 week ago

Site Reliability Engineer

Runloop · San Francisco, CA · Onsite

1 week ago

Site Reliability Engineer

Triomics · India Office · Hybrid

1 week ago
Site Reliability Engineer (SRE) at Acquireai | Hiring.Camp