Hiring.Camp

Site Reliability Engineer

SolarWinds

Location
Bangalore, India · Bangalore Office
Department
Education
Experience
2+ years

Description

At SolarWinds, we’re a people-first company. Our purpose is to enrich the lives of the people we serve—including our employees, customers, shareholders, partners, and communities. Join us in our mission to help customers accelerate business transformation with simple, powerful, and secure solutions.

The ideal candidate thrives in an innovative, fast-paced environment and is collaborative, accountable, ready, and empathetic. We’re looking for individuals who believe they can accomplish more as a team and create lasting growth for themselves and others. We hire based on attitude, competency, and commitment. Solarians are ready to advance our world-class solutions in a fast-paced environment and accept the challenge to lead with purpose. If you’re looking to build your career with an exceptional team, you’ve come to the right place. Join SolarWinds and grow with us!

At SolarWinds, we’re a people-first company. Our purpose is to enrich the lives of the people we serve-including our employees, customers, shareholders, Partners, and communities. Join us in our mission to help customers accelerate business transformation with simple, powerful, and secure solutions.

The ideal candidate thrives in an innovative, fast-paced environment and is collaborative, accountable, ready, and empathetic. We’re looking for individuals who believe they can accomplish more as a team and create lasting growth for themselves and others. We hire based on attitude, competency, and commitment. Solarians are ready to advance our world-class solutions in a fast-paced environment and accept the challenge to lead with purpose. If you’re looking to build your career with an exceptional team, you’ve come to the right place. Join SolarWinds and grow with us!

 

What you will do
SolarWinds is looking for a Senior Site Reliability Engineer with 2+ years of experience to help build, operate, and improve the systems that support our products and internal platforms. This role is ideal for someone who is hands-on, dependable, and comfortable working across Kubernetes, AWS, Azure, Linux, and Database environments in a production setting.

  • The right candidate has a foundation in site reliability engineering, enjoys solving problems in live environments, and understands the importance of reliability, automation, and operational discipline.
  • Operate, maintain, upgrade and improve production Kubernetes clusters and workloads across AWS and Azure environments.
  • Manage Kubernetes platform components and technologies such as Helm, Kustomize, operators, istio, autoscaling, and cluster/node lifecycle management.
  • Support and maintain production database platforms such as ClickHouse, Aurora and other distributed data systems, including performance troubleshooting and operational health.
  • Build and maintain infrastructure using Terraform.
  • Develop automation and tooling to reduce operational toil and improve reliability using Python, Go, Bash, or similar technologies.
  • Participate in an on-call rotation, respond to production incidents, and lead or contribute to incident resolution and root-cause analysis.
  • Improve observability, monitoring, logging, alerting, and incident response across infrastructure and services.
  • Partner closely with software engineering and platform teams to design and deploy reliable, scalable services.
  • Contribute to capacity planning, performance optimization, patching, upgrades, and infrastructure lifecycle management.
  • Develop and maintain operational documentation, runbooks, and troubleshooting guides.
  • Participate in reliability initiatives, including automation, resilience engineering, disaster recovery, and continuous improvement.

Required Qualifications

  • 2+ years of experience in Site Reliability Engineering, DevOps, Systems Engineering, Platform Engineering, or a related field.
  • Strong hands-on production Kubernetes experience, including operating, upgrading and troubleshooting Kubernetes clusters and workloads.
  • Strong understanding of Kubernetes fundamentals, including:
  • Pods, Deployments, StatefulSets, DaemonSets, Jobs, CronJobs
  • Services, Ingress, DNS, and Kubernetes networking
  • ConfigMaps, Secrets, and persistent storage
  • Resource requests/limits and scheduling
  • Autoscaling across workloads, clusters, and nodes
  • Pod Disruption Budgets and high availability patterns
  • Kubernetes upgrades and cluster lifecycle management
  • Experience troubleshooting Kubernetes at both the application workload and cluster infrastructure levels.
  • Strong hands-on experience with AWS and Azure cloud infrastructure.
  • Strong Linux systems administration and troubleshooting skills.
  • Experience operating customer facing, highly available production systems.
  • Experience participating in a production on-call rotation and responding to high-severity incidents.
  • Strong experience with Terraform and Infrastructure as Code.
  • Experience with scripting and automation using Python, Go, Bash, or similar languages.
  • Strong understanding of infrastructure concepts including compute, networking, storage, DNS, load balancing, and security.
  • Strong troubleshooting skills with the ability to methodically diagnose complex distributed-system failures.
  • Strong communication skills and the ability to collaborate effectively with engineering and cross-functional teams.

What we are looking for

  • Ownership mindset and strong operational discipline
  • Ability to stay calm and effective during incidents
  • Willingness to learn, improve systems, and drive reliability-focused changes
  • Practical approach to solving infrastructure and operational problems
  • Team player who values clear communication and documentation
  • Continuous learner, stays current with Kubernetes, cloud infrastructure, and modern SRE practices.

On-Call Expectations
This position requires participation in a scheduled on-call rotation to support production systems and help ensure service reliability and availability.

SolarWinds is an Equal Employment Opportunity Employer. SolarWinds will consider all qualified applicants for employment without regard to race, color, religion, sex, age, national origin, sexual orientation, gender identity, marital status, disability, veteran status or any other characteristic protected by law.

All applications are treated in accordance with the SolarWinds Privacy Notice: https://www.solarwinds.com/applicant-privacy-notice

Skills

PythonAWSAzureKubernetesTerraformLinuxDevOpsSREGo

Similar Jobs

30

DevOps Engineer

Razorpaysoftwareprivatelimited·Malaysia

Today

Site Reliability Engineer

RELX Jobs·Home Based - Ireland·Remote

Today

Technology Platform Engineer

Accenture·Coimbatore, CODC1A

Today

DevOps Engineer

Accenture·Navi Mumbai, MDC5C

Today

DevOps Engineer

Accenture·Bengaluru, BDC7B

Today

OpenShift Platform Engineer

Westpacnz·Westpac on Takutai Square, New Zealand·Hybrid

Today

Senior Platform Engineer

Mintlify·San Francisco

1d ago

DevOps Engineer

" MAXISIQ, Inc."·Chantilly, VA

3d ago

Site Reliability Engineer

Armor·Pune, Maharashtra

3d ago

Site Reliability Engineer

Point72·Bengaluru, India +1

3d ago

DevOps Engineer

TBCBANK·Tbilisi, Georgia·Hybrid

3d ago

Software Engineer - Platform

Nokia·India·Hybrid

3d ago

Site Reliability Engineer

Binance·Asia·Remote

3d ago

DevOps Engineer

YourCode·Onsite

3d ago

Databricks Platform Engineer

Mckesson·Columbus, OH +1·Hybrid, Remote

3d ago

DevOps Engineer

Roche·Indianapolis, US

3d ago

Databricks Platform Engineer

The Future of Health Starts with You·Columbus, OH +1·Hybrid, Remote

3d ago

Platform Desktop Engineer

Gresearch·London, UK

3d ago

Senior Platform Engineer

Draftkings·Remote - US, US·Remote

3d ago

DevOps Engineer

Nasdaq·Vilnius, Lithuania·Hybrid, Onsite

3d ago

DevOps Engineer

Booz Allen Hamilton·Lexington, MA +6

3d ago

DevOps Engineer

Barclays·Knutsford, Radbroke Hall

3d ago

Site Reliability Engineer

AutoDesk·APAC - India - Pune - ABIL Boulevard

3d ago

DevOps Engineer

Usbank·Warsaw, Poland

3d ago

Senior Platform Engineer

Mastercard·Pune, India

3d ago

Site Reliability Engineer

Chevron Corporation is one of·Makati City Chevron 6750 Office, Philippines

3d ago

Cloud Platform Engineer

Singlestore·India +1

3d ago

DevOps Engineer

SOSi·Remote, US·Remote

3d ago

AI Platform Engineer

Starburst·India·Remote

3d ago

DevOps Engineer

Investorflow·Santiago de los Caballeros·Onsite

4d ago
Site Reliability Engineer at SolarWinds | Hiring.Camp