- Location
- Bengaluru, India
- Type
- Full-time
- Department
- Engineering
- Seniority
- VP
- Closing date
- Today
- Source
- Workday
Description
Join us as a Site Reliability Engineer
- This is an opportunity for a driven Software Engineer to take on an exciting new career challenge
- Day-to-day, you'll build a wide network of stakeholders of varying levels of seniority
- It’s a chance to hone your existing technical skills and advance your career
- We're offering this role at associate vice president level
What you'll do
In your new role, you’ll engineer and maintain innovative, customer centric, high performance, secure and robust solutions. You’ll be working within a feature team and using your extensive experience to engineer software, scripts and tools that are often complex, as well as liaising with other engineers, architects and business analysts across the platform.
You’ll also be:
- Producing complex and critical software rapidly and of high quality which adds value to the business
- Working in permanent teams who are responsible for the full life cycle, from initial development, through enhancement and maintenance to replacement or decommissioning
- Collaborating to optimise our software engineering capability
- Designing, producing, testing and implementing our working code
- Working across the life cycle, from requirements analysis and design, through coding to testing, deployment and operations
The skills you'll need
You’ll need you to Own the reliability, availability, and operational performance of customer-facing and business-critical applications, support and troubleshoot distributed systems, APIs, microservices, and cloud-native platforms in production environments.
You'll also lead major incident management, service restoration, root cause analysis, and post-incident reviews, partner with Product, Engineering, and Customer Support teams to resolve customer-impacting issues and improve service stability and define and drive Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budget practices.
You’ll also need:
- Strong experience supporting REST APIs, microservices, and distributed systems at scale and deep understanding of API monitoring, performance analysis, troubleshooting, and dependency management.
- Expertise in diagnosing and resolving issues across complex microservice architectures, APIs, and interconnected systems
- Experience managing customer-facing incidents, ensuring service availability, and implementing observability, monitoring, SLI/SLOs, and reliability best practices
- Experience with cloud platforms such as AWS or GCP and hands-on expertise with Kubernetes, Docker, and Infrastructure as Code (Terraform)
- Experience with CI/CD pipelines using Azure DevOps, GitHub Actions, or similar tools
- Strong observability experience using tools such as Grafana, Prometheus, Splunk, or Datadog
- Proficiency in one or more scripting/programming languages such as Python, PowerShell, Bash, or Go
- Experience with performance tuning, scalability testing, and production support operations
Hours
45Job Posting Closing Date:
09/09/2026