- Location
- Manila, Philippines
- Workplace
- Hybrid
- Type
- Full-time
- Department
- Engineering
- Source
- Workday
Description
About this job opportunity
Our Vision
To be the world's most trusted global payroll partner, simplifying pay for all employees.
Our Mission
Empowering global workforces with seamless, compliant, and innovative payroll and payment solutions, enabling businesses to thrive in a connected world.
Our People
Our fundamental beliefs at CloudPay are built on core values of professionalism, passion, empowerment, innovation, and teamwork. We value our employees and strive to create a great workplace where everyone is valued, heard, inspired, and encouraged to bring their authentic selves to work. We're committed to providing an excellent employee experience through fulfilling projects, empowerment to make a difference, and an environment that inspires innovation.
What makes this role exciting
The Platform Operations Engineer is responsible for ensuring the reliability, performance, security, and operational efficiency of enterprise platforms and services. The role combines infrastructure management, automation, monitoring, incident management, and continuous improvement practices to support business-critical systems, working collaboratively with other platform engineers, software developers and SecOps to maintain a stable and efficient platform environment.
Main responsibilities
Platform Monitoring & Maintenance:
- Monitor platform health, performance metrics, and proactively identify and resolve issues.
- Perform routine maintenance tasks, such as patching, updates, backups, and configuration changes.
- Ensure platform availability and uptime by managing incidents, troubleshooting problems, and implementing corrective actions.
- Manage and optimise the platform's infrastructure resources (servers, networks, storage) for optimal performance and cost-effectiveness.
- Create and maintain documentation for platform configurations, procedures, and best practices.
Incident Management & Response:
- Investigate and resolve platform incidents, collaborating with other teams as needed.
- Document incident details, root causes, and resolution steps for future reference and continuous improvement.
- Assist with security incident investigation, containment, eradication, recovery, and post-incident reviews. Support root cause analysis, evidence collection, documentation, and the implementation of corrective actions to reduce the likelihood and impact of future security incidents.
- Participate in on-call rotation to provide 24/7 support for critical platform issues.
Automation & Tooling:
- Create and manage scripts using Python, Bash, or similar technologies to automate routine administration and support activities.
- Implement Infrastructure as Code (IaC) using tools such as Terraform.
- Utilise automation and AI tools to enhance operational efficiency and reduce manual effort.
- Develop and maintain CI/CD pipelines to support platform deployments and operational changes.
Collaboration & Support:
- Provide support to wider business teams to ensure smooth deployment and operation of applications on the platform.
- Collaborate with engineering and security teams to embed automation into platform lifecycle management and compliance processes.
- Provide technical support to the engineering, QA and development teams
- Work closely with the SecOps team to ensure platform security controls are effectively implemented, monitored, and maintained. Support vulnerability management activities including vulnerability assessment, patching, remediation, and validation of security findings within agreed SLAs, helping to strengthen the overall security posture and compliance of the platform.
Experience needed for this role
Experience:
- Proven experience in platform operations and DevOps
- Skill in Linux OS or other Unix-based operating systems.
- Good knowledge of the AWS platform and its technologies. (EC2, ECS, Lambda, etc)
- Strong experience with DevOps tools such as Terraform and GitHub
- Knowledge of monitoring and observability tools such as Prometheus and Grafana.
Skills:
- Excellent troubleshooting and problem-solving skills.
- Strong scripting or programming skills (e.g., Python, Bash).
- CI/CD Pipelines and Infrastructure as Code (IaC).
- Terraform and GitHub
Additional Considerations (Optional):
- Certifications in relevant technologies such as AWS Solutions Architect and HashiCorp Terraform Associate are beneficial.
- Experience with specific programming languages or frameworks used on the platform can be an asset.
- ITIL Foundation / Managing Professional
- Identity and Access Management (IAM), Zero Trust principles and Security Compliance Frameworks (ISO 27001, SOC 2, GDPR)
About you and Our core values
- Taking ownership, working with integrity and respect
- Being a team player is key to our culture
- Solution and customer focused
- Great initiative with the goal for excellence in achieving results
- Dedicated to developing and always looking for continuous improvements
- Be creative, be committed, be engaged and enjoy what you do
Philippines Package and Benefits:
- Competitive Salary
- Competitive vacation allowance
- Calm app
- Sick Leave
- EAP
- Group Life Insurance, HMO
- Employee Referral Program
- De Minimis Benefit
- WFH Allowance
- Mid-Year Bonus
- 13th Month Pay
- Regularization Bonus, 1st Year Anniversary Bonus
- Bereavement Leave
- Paid Volunteering days
- Study Leave
- Marriage Leave
CloudPay is committed to being an equal opportunities employer
#LI-Hybrid #LI-GH
The CloudPay culture is built upon on five core values, from which we develop our service, our technology and our business strategies. Our fundamental beliefs are a promise to our employees, customers and partners, built on the core values of professionalism, passion, empowerment, innovation, and teamwork.
Glassdoor