Hiring.Camp

Site Reliability Engineer

Leidos

·

Yesterday

Salary
$87k – $157k
Location
10585 Chantilly VA, United States of America · 10587 Boulder CO · 10588 San Antonio TX · 10586 Columbus OH
Type
Full-time
Department
Engineering
Source
Workday

Description

Who We Are:

Kudu Dynamics is a 100% employee-owned company, forged out of a decade of experience in computer network operations and staffed with talent who have built, overseen, and enhanced capabilities throughout the entire USG arsenal. Our team of hackers, engineers, makers, and shakers brings deep experience across research, development, deployment, and operations.

Kudu Dynamics is uniquely qualified to anticipate tomorrows threats and build the next generation of capabilities.

When you come in for your interview, youll see that Kudu is an amazing place to work, where youll be surrounded by experts who are ready to teach and learn. Our team has flexible work hours and work-from-home options. When we do work from the offices, we enjoy home-roasted coffee and award-winning workspaces.

Job Description:

Hello! We’re a small team in a fun company looking for someone to come in and help us build systems that stay reliable when things get complicated.

We need a Site Reliability Engineer who has experience building, deploying, automating, and operating complex compute platforms. Youll work across infrastructure, Linux systems, networking, distributed storage, observability, security, and deployment automation, with a particular emphasis on creating systems that are reproducible, maintainable, resilient, and easy to operate.

A major part of this role will involve Nix and NixOS. We want someone who appreciates declarative systems, reproducible environments, infrastructure-as-code, and reducing configuration drift. You may be building NixOS-based servers, improving deployment pipelines, developing reusable Nix modules, debugging distributed systems, or making sure a platform can be rebuilt predictably from scratch.

Critical components of the environment may include high-performance compute, distributed Linux file systems, network design, security-in-depth, ML/AI infrastructure, high-bandwidth data processing, cloud deployment, and fielded systems.

You dont need to be an expert at everything. Whatever part you bite off, though, will be yours to own. This is a great opportunity to apply deep systems expertise while learning adjacent areas.

Our team needs your help making sure our customers are repeatedly provided with operational insights from a multi-domain data environment—and that the infrastructure producing those insights is dependable, observable, reproducible, and recoverable.

These Are the Things Youll Get to Do:

           Own the reliability and operation of critical compute and data platforms.

           Design, deploy, and maintain Linux infrastructure, including NixOS-based systems.

           Build reproducible system configurations and deployment workflows using Nix and infrastructure-as-code.

           Develop reusable NixOS modules, packages, flakes, and system configurations.

           Improve platform resilience, fault tolerance, recoverability, and maintainability.

           Build monitoring, logging, alerting, and observability capabilities that help us understand system behavior before customers notice problems.

           Diagnose difficult failures across Linux, networking, storage, containers, and distributed systems.

           Automate repetitive operational tasks and eliminate configuration drift.

           Design and build security and maintainability components for deployed systems.

           Help define operational standards, deployment practices, upgrade strategies, and disaster-recovery procedures.

           Perform capacity planning and identify performance bottlenecks across compute, storage, and networking.

           Own solutions spanning CNO, analytics, security, infrastructure, and field deployments.

           Learn new skills and teach the rest of us what you know.

Minimum Qualifications:

           Enthusiasm for learning new stuff.

           Bachelors degree in Computer Science, Computer Engineering, a related field, or amazing equivalent skills.

           Strong experience administering and troubleshooting Linux systems.

           Experience with Linux networking, storage, and file systems.

           Experience automating system configuration and deployment.

           Experience with Nix and/or NixOS in production, lab, or significant personal environments.

           Python, Go, Bash, or similar scripting/programming experience of 2+ years.

           Ability to debug complex systems methodically across multiple layers of the stack.

Nice-to-Have Qualifications:

           Deep experience with NixOS, including custom modules, overlays, flakes, packaging, and reproducible deployments.

           Experience operating fleets of Linux systems.

           Infrastructure-as-code experience with Nix, Terraform, Ansible, or similar tooling.

           Experience designing highly available or fault-tolerant systems.

           Monitoring and observability experience with tools such as Prometheus, Grafana, Loki, OpenTelemetry, Elasticsearch, or similar systems.

If you're looking for comfort, keep scrolling. At Leidos, we outthink, outbuild, and outpace the status quo — because the mission demands it. We're not hiring followers. We're recruiting the ones who disrupt, provoke, and refuse to fail. Step 10 is ancient history. We're already at step 30 — and moving faster than anyone else dares.

Original Posting:

August 14, 2026

For U.S. Positions: While subject to change based on business needs, Leidos reasonably anticipates that this job requisition will remain open for at least 3 days with an anticipated close date of no earlier than 3 days after the original posting date as listed above.

Pay Range:

Pay Range $87,100.00 - $157,450.00

The Leidos pay range for this job level is a general guideline only and not a guarantee of compensation or salary. Additional factors considered in extending an offer include (but are not limited to) responsibilities of the job, education, experience, knowledge, skills, and abilities, as well as internal equity, alignment with market data, applicable bargaining agreement (if any), or other law.

Skills

PythonTerraformAnsibleLinuxElasticsearchGo

Similar Jobs

30

DevOps Engineer

Pythian · India

Today

DevOps Engineer

Pythian · Mexico +4

Today

Devops Engineer

Sutherland · Bogotá, Bogota, Colombia · Hybrid

Today

Senior Platform Engineer

H&M Group · Stockholm, Stockholms län, Sweden

Today

Data Platform Engineer

Accenturefederalservices · Suitland, MD +1 · Onsite

Today

Data Platform Engineer

Accenture · Mumbai, MDC2B, India

Yesterday

Data Platform Engineer

Accenture · Pune, PDC2C, India

Yesterday

Cloud Platform Engineer

Accenture · Bhubaneswar, BBDC1A, India

Yesterday

Cloud Platform Engineer

Accenture · Bhubaneswar, BBDC1A, India

Yesterday

Site Reliability Engineer

Lbg · London 1-10 Praed Mews, United Kingdom · Remote, Hybrid

Yesterday

DevOps Engineer

Rockwellautomation · Colombia Medellin · Hybrid

Yesterday

Site Reliability Engineer

Lbg · London 1-10 Praed Mews, United Kingdom · Remote, Hybrid

Yesterday

Celonis Platform Engineer

Mdlz · Business Office (Joy House 2) - Mumbai, India

Yesterday

Site Reliability Engineer

LSEG · IND-BLR-Divyasree Technopolis, India

Yesterday

Staff Platform Engineer

Robotsandpencils · Calgary, Alberta, Canada, Canada - Remote +2 · Remote

Yesterday

Staff Platform Engineer

Robotsandpencils · Austin, TX +1 · Remote

Yesterday

Cloud Platform Engineer

Accenturefederalservices · Chantilly, VA +1 · Onsite

Yesterday

Engineer - AI Platform

Ebury · Madrid · Hybrid

Yesterday

DevOps Engineer

Wrike · Prague · Hybrid

Yesterday

Site Reliability Engineer

NielsenIQ · Mumbai, MH, India

Yesterday

AI Platform Engineer

Vodafone · London, England,GB, GB · Hybrid

Yesterday

DevOps Engineer

Quess · Coimbatore, Tamil Nadu, India · Onsite

Yesterday

DevOps Engineer

GDIT · USA MD Annapolis Junction - 135 National Business Parkway (MDS048), United States of America

2 days ago

Salesforce Platform Engineer

Limble CMMS · Remote · Remote

2 days ago

DevOps Engineer

MMIST · Ottawa, Ontario, Canada

2 days ago

DevOps Engineer

Accenture · Chennai, CDC2A, India

2 days ago

Technology Platform Engineer

Accenture · Kolkata, KDC1A, India

2 days ago

DevOps Engineer

Motorola Solutions · Allen, TX (TX139), United States of America

2 days ago

Platform DevOps Engineer

Booz Allen Hamilton · USA, VA, Chantilly (14151 Park Meadow Dr), United States of America

2 days ago

DevOps Engineer

Booz Allen Hamilton · USA, OH, Wright Patterson AFB (4180 Watson Way), United States of America

2 days ago
Site Reliability Engineer at Leidos • $87k – $157k | Hiring.Camp