- Location
- IND - Mumbai - Corporate Office, India
- Type
- Full-time
- Department
- Engineering
- Source
- Workday
Description
Cloud DevOps Engineer
Role Purpose
The Cloud DevOps Engineer will be responsible for designing, building, automating and managing cloud-based infrastructure using DevOps and Infrastructure as Code (IaC) practices.
Reporting to the Infrastructure Services Technical Manager / Cloud DevOps Manager, the engineer will work closely with application, security and infrastructure teams to design, implement and operate secure, scalable, highly available and cost-effective AWS cloud solutions.
The role will contribute to cloud migration, infrastructure automation, operational excellence, continuous improvement and the adoption of Cloud DevOps best practices across Travelex.
What You Will Be Doing
Cloud Infrastructure & Engineering
Design, build and manage AWS cloud infrastructure for different product lines across Travelex.
Implement secure, scalable, highly available and resilient cloud infrastructure solutions.
Work on cloud migration and modernisation initiatives.
Build reusable infrastructure patterns and automation to improve engineering efficiency.
Implement and maintain infrastructure using Infrastructure as Code (IaC), primarily Terraform and/or CloudFormation.
Manage and support AWS services including EC2, VPC, IAM, S3, RDS, Lambda, ECS/EKS, Elastic Load Balancing and related services.
Implement and maintain AWS networking components including VPCs, subnets, route tables, security groups, NACLs and DNS.
Support containerised workloads using Docker and AWS ECS/EKS/Fargate.
Automation & CI/CD
Develop automation to reduce manual operational activities and prevent recurring support incidents.
Build and maintain CI/CD pipelines for infrastructure and application deployments.
Implement automated infrastructure provisioning, configuration and deployment processes.
Develop scripts using Python, Bash, PowerShell, AWS CLI or equivalent automation tools.
Promote reusable automation frameworks and self-service capabilities across engineering teams.
Monitoring, Reliability & Operations
Implement monitoring, logging, alerting and observability for cloud infrastructure and applications.
Work with tools such as Amazon CloudWatch, Grafana, Datadog and other observability platforms.
Proactively identify performance, availability and reliability issues and implement preventative solutions.
Participate in production releases, infrastructure changes and resolution of production incidents.
Perform troubleshooting, root cause analysis and implement permanent corrective actions.
Support patching, vulnerability remediation and End-of-Support/End-of-Life remediation activities.
Contribute to Disaster Recovery, Business Continuity and operational resilience initiatives.
Security & Compliance
Implement AWS security best practices, including IAM, least privilege, encryption and secure configuration.
Support cloud security and compliance requirements, including PCI and other organisational controls.
Identify and remediate infrastructure vulnerabilities and security findings.
Ensure infrastructure changes follow appropriate governance, change management and security standards.
Cost Optimisation
Monitor AWS resource utilisation and identify opportunities for cost optimisation.
Support FinOps initiatives through rightsizing, resource optimisation, lifecycle management and automation.
Promote cost-aware cloud architecture and engineering practices.
Agile / Team Responsibilities
Participate actively in Agile/Scrum ceremonies including sprint planning, daily stand-ups, sprint reviews and retrospectives.
Work with the Cloud DevOps Manager on backlog management, prioritisation and refinement.
Proactively identify and resolve technical blockers and dependencies.
Contribute to improving the team's delivery velocity, quality and engineering practices.
Conduct demonstrations of completed Cloud DevOps capabilities to consuming teams and stakeholders.
Work collaboratively with team members to promote cross-skilling and breadth of technical knowledge.
Contribute to knowledge management through Confluence, technical documentation, runbooks and knowledge-sharing sessions.
Provide regular status and progress updates to the management team.
Actively contribute to continuous improvement initiatives and ensure agreed retrospective actions are implemented.
What We Are Looking For
Essential Experience & Skills
2–3 years of hands-on experience in AWS Cloud and DevOps.
Good understanding of AWS core services
Hands-on experience with Docker and containerisation.
Operational experience with AWS ECS and Fargate; exposure to Kubernetes/EKS is desirable.
Experience building Docker images using Dockerfiles and Docker Compose.
Good understanding of AWS networking including VPC, subnets, routing, security groups and load balancers.
Hands-on experience with Infrastructure as Code tools such as Terraform and/or CloudFormation.
Experience with configuration management tools such as Ansible.
Working knowledge of CI/CD pipelines and tools such as Jenkins, GitHub Actions, CircleCI or equivalent.
Experience with Git and source-control-based development practices.
Experience with scripting/automation using Python, Bash, PowerShell or AWS CLI.
Working knowledge of monitoring, logging and observability using CloudWatch, Grafana, Datadog or similar tools.
Understanding of cloud security principles, IAM and least-privilege access.
Understanding of production support, incident management and root cause analysis.
Good understanding of Agile/Scrum delivery practices and JIRA.
Desirable Skills
AWS certification such as AWS Solutions Architect, Developer or SysOps Administrator.
Experience with Kubernetes/EKS and Helm.
Experience developing reusable Terraform modules and managing Terraform state.
Experience with AWS Lambda-based automation.
Knowledge of AWS cost optimisation and FinOps practices.
Experience with vulnerability remediation, patching and EoSL/EoL remediation.
Knowledge of Disaster Recovery and Business Continuity practices.
Experience working in a regulated environment with security and compliance requirements.
Experience with application performance monitoring and distributed tracing.
Experience working with multiple AWS accounts / AWS Organisations.
Key Behaviours
Strong problem-solving and analytical skills.
Proactive approach to identifying and resolving technical issues.
Strong focus on automation and continuous improvement.
Good communication and stakeholder-management skills.
Ability to work independently while collaborating effectively with cross-functional teams.
Willingness to learn new AWS technologies and share knowledge with other team members.
Strong ownership of production infrastructure and service reliability.