Hiring.Camp

Site Reliability Engineer (Managed Patching & Platform Automation)

Swift

·

Today

Location
Kuala Lumpur, Malaysia
Department
Engineering
Education
Bachelor
Source
Workday

Description

ABOUT US

We’re the world’s leading provider of secure financial messaging services, headquartered in Belgium. We are the way the world moves value – across borders, through cities and overseas. No other organisation can address the scale, precision, pace and trust that this demands, and we’re proud to support the global economy. 

We’re unique too. We were established to find a better way for the global financial community to move value – a reliable, safe and secure approach that the community can trust, completely. We’re always striving to be better and are constantly evolving in an ever-changing landscape, without undermining that trust. Five decades on, our vibrant community reflects the complexity and diversity of the financial ecosystem. We innovate diligently, test exhaustively, then implement fast. In a connected and exciting era, our mission has never been more relevant. Swift now has a presence in 200+ countries and legal territories to serve a community of more than 12,000 banks and financial institutions.   

Experience and Qualifications

  • 3 plus years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Infrastructure Engineering, or related disciplines
  • Proven experience building and operating enterprise-scale automation solutions
  • Strong hands-on experience in infrastructure automation, Linux administration, and system reliability
  • Experience working within large-scale enterprise environments
  • Bachelor's Degree in Computer Science, Engineering, Information Technology, or equivalent practical experience

Key Responsibilities

Support the Managed Patching Service (MPS)

  • Contribute to the enhancement and continuous improvement of the Managed Patching Service (MPS)
  • Take ownership of assigned technical deliverables and service improvements
  • Support the scalability, reliability, performance, and maintainability of the service
  • Help implement engineering solutions that meet enterprise and regulatory requirements
  • Participate in the evolution of MPS toward a platform-driven and self-service operating model

Design and Build Enterprise Automation Solutions

  • Design, develop, and maintain automation workflows using Ansible Automation Platform
  • Develop reusable automation components, scripts, and operational tooling
  • Integrate automation solutions with ServiceNow, inventory systems, CI/CD platforms, and related enterprise services
  • Apply engineering best practices including testing, version control, peer review, and release management
  • Develop automation and integrations using Python and related technologies where required

Platform Integration and Service Engineering

  • Support onboarding, subscription, scheduling, and maintenance window capabilities within MPS
  • Improve service delivery through automation and standardization
  • Contribute to platform enhancements that improve user experience and operational efficiency
  • Assist in building scalable integration patterns across enterprise platforms
  • Support initiatives that reduce manual effort and improve service adoption

Reliability, Observability and Reporting

  • Implement and maintain operational monitoring, logging, and reporting capabilities
  • Contribute to the definition and measurement of service reliability objectives and operational metrics
  • Improve visibility of patching outcomes, compliance status, and service health
  • Support the development of dashboards and reporting solutions for operational and regulatory requirements
  • Identify opportunities to improve service reliability and reduce operational complexity

Incident, Problem and Operational Management

  • Investigate and resolve complex technical issues affecting service availability or performance
  • Participate in incident response, troubleshooting, and service recovery activities
  • Contribute to root cause analysis and corrective actions following incidents
  • Develop and maintain operational documentation, runbooks, and troubleshooting guides
  • Support continuous improvement initiatives to improve service stability and resilience

Compliance and Governance

  • Support compliance with enterprise security, risk, and regulatory requirements
  • Ensure automation workflows maintain appropriate traceability and auditability
  • Contribute to the implementation of governance controls and operational standards
  • Support evidence collection and reporting requirements for audits and compliance reviews
  • Assist in maintaining service documentation and operational records

Platform Operations and Automation Engineering

  • Contribute to infrastructure automation and platform engineering initiatives beyond Managed Patching Service
  • Apply automation and SRE practices to improve operational efficiency and reliability across related platform services
  • Support service transition, operational readiness, and continuous improvement activities
  • Collaborate with engineering teams to identify automation opportunities and operational improvements
  • Participate in shared engineering responsibilities aligned with evolving business priorities

Collaboration and Technical Contribution

  • Collaborate closely with Infrastructure, Security, Architecture, Service Management, and Engineering teams
  • Work with engineers across teams to identify and resolve technical challenges
  • Participate in design discussions, solution reviews, and technical workshops
  • Share knowledge, best practices, and lessons learned with team members
  • Provide guidance and mentoring to less experienced engineers when required
  • Support service adoption by collaborating with stakeholders and platform consumers

Required Skills

  • Strong expertise in Ansible Automation Platform
  • Strong Linux administration and troubleshooting experience (RHEL preferred)
  • Experience developing automation solutions using scripting languages such as Python
  • Experience integrating enterprise platforms such as ServiceNow, CMDBs, monitoring solutions, and CI/CD tools
  • Strong understanding of Site Reliability Engineering principles and operational practices
  • Experience managing and supporting large-scale infrastructure environments
  • Good understanding of automation governance, change management, and operational controls
  • Strong analytical and problem-solving skills
  • Strong communication and collaboration skills


Preferred Skills

  • Experience with CloudBees, Jenkins, GitHub Actions, or similar CI/CD platforms
  • Familiarity with infrastructure as code practices
  • Experience with observability, monitoring, and enterprise reporting solutions
  • Experience with Power BI or similar reporting tools
  • Experience working in regulated or financial services environments
  • Exposure to platform engineering or self-service operational models


What Success Looks Like (6–12 Months)

  • Managed Patching Service operates reliably and efficiently within its defined scope
  • Automation capabilities are enhanced with reduced manual intervention
  • Service onboarding and operational processes become increasingly standardized
  • Operational and compliance reporting is available and trusted by stakeholders
  • Service reliability and operational performance improve through continuous enhancement
  • Strong collaboration is established across engineering and support teams
  • Contributions made to broader infrastructure automation and platform engineering initiatives
  • Knowledge is actively shared within the team, helping raise overall engineering capability

What we offer

We give you the freedom to be yourself. We are creating an environment of unique individuals – like you – with different perspectives on the financial industry and the world. A diverse and inclusive environment in which everyone’s voice counts and where you can reach your full potential.

We are committed to an inclusive and accessible recruitment process. If you require a reasonable accommodation related to accessibility during your application or interview, please contact [email protected] or indicate this in your application.

Please note that this mailbox is not monitored for general recruitment enquiries and should only be used for accessibility or accommodation-related requests (for example related to vision, hearing or neurodiversity).

All requests are confidential and will not affect your candidacy.

Don’t meet every single requirement? At Swift, we are dedicated to building a workplace where people can bring their full selves and ideas to the team, so if you are excited about this role, we encourage you to apply even if you do not meet every single qualification.

Skills

PythonSwiftAnsibleJenkinsCI/CDLinuxGitHubPower BIServiceNowDevOpsSREComplianceChange Management

Similar Jobs

30

Site Reliability Engineer

Tennr · New York City Office · Onsite

Yesterday

DevOps Engineer

SmartRent · Phoenix, Arizona · Remote

2 days ago

Staff Platform Engineer

Shieldai · San Diego, California +4 · Onsite

2 days ago

DevOps Engineer

NS2 Mission · Chantilly, VA

2 days ago

ServiceNow Platform Engineer

Lightfeatheriollc · Washington, DC +1 · Hybrid

2 days ago

DevOps Engineer

Miratech · Brasília, DF, Brazil · Remote

2 days ago

DevOps Engineer

E-INFOSOL · Washington, DC +1

2 days ago

DevOps Engineer

Netcompany · London, England, United Kingdom · Hybrid

2 days ago

Senior Platform Engineer

LSPedia · Farmington Hills, MI

3 days ago

Staff Platform Engineer

InPost · Warszawa, Województwo mazowieckie, Poland · Remote

3 days ago

Release Platform Engineer

bet365 · Manchester, England, United Kingdom · Hybrid

3 days ago

Release Platform Engineer

bet365 · Stoke-on-Trent, England, United Kingdom · Hybrid

3 days ago

DevOps Engineer

Netcompany · Leeds, England, United Kingdom · Hybrid

3 days ago

AI Platform Engineer

Sysdig · Raleigh · Remote

3 days ago

DevOps Engineer

Intertec · Skopje, Macedonia, the former Yugoslav Republic of · Hybrid

3 days ago

DevOps / Platform Engineer

SIA · Paris, IDF, France

3 days ago

DevOps Engineer

Talan · Buenos Aires, Buenos Aires, Argentina · Remote

3 days ago

DevOps Engineer

SIA · Mumbai, Maharashtra, India

3 days ago

Senior Platform Engineer

DoiT · Remote EMEA · Remote

3 days ago

Senior Platform Engineer

DoiT · Remote EMEA · Remote

3 days ago

Senior Platform Engineer

DoiT · Remote Ireland · Remote

3 days ago

Senior Platform Engineer

DoiT · Remote EMEA · Remote

3 days ago

Senior Platform Engineer

DoiT · Remote Estonia · Remote

3 days ago

Senior Platform Engineer

DoiT · Remote EMEA · Remote

3 days ago

Senior Platform Engineer

DoiT · Israel · Remote

3 days ago

Senior Platform Engineer

DoiT · Remote Sweden · Remote

3 days ago

Senior Platform Engineer

DoiT · Remote Netherlands · Remote

3 days ago

Senior Platform Engineer

DoiT · Remote UK · Remote

3 days ago

Palantir Platform Engineer

Accenture Federal Services · Herndon, VA +1 · Onsite

3 days ago

Senior Platform Engineer

Blinq · Melbourne, Victoria +1 · Hybrid

3 days ago
Site Reliability Engineer (Managed Patching & Platform Automation) at Swift | Hiring.Camp