Hiring.Camp

Senior Staff Platform Engineer

Nvidia

·

Yesterday

Location
US, CA, Santa Clara, United States of America
Type
Full-time
Department
Engineering
Seniority
Senior
Source
Workday

Description

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

Ready to build the platforms that make AI at scale possible? NVIDIA is seeking a Senior Staff Platform Engineer to architect, build, and scale foundational infrastructure for some of our most demanding compute and AI/ML workloads. The role spans distributed systems, cloud, networking, content delivery, automation, and reliability engineering. Success means turning complex infrastructure challenges into resilient platforms that can grow with NVIDIA’s rapidly evolving needs. This is a hands-on technical leadership role with broad influence across Cloud, Networking, Security, AI/ML, and Developer Infrastructure teams.

What You Will Be Doing:

This role focuses on architecting, building, and scaling highly available platform services for AI/ML, distributed compute, and data-intensive workloads. The work spans cloud, compute, GPU infrastructure, networking, storage, DNS, load balancing, proxies, traffic management, and content delivery, with end-to-end ownership from architecture and design through implementation, production readiness, and large-scale adoption. A key part of the role is advancing CDN and edge infrastructure, including HTTP caching, origin architecture, TLS, WAF, rate limiting, traffic routing, and global load balancing. The role also drives automation, infrastructure-as-code, and self-service capabilities using technologies such as Python, Go, Kubernetes, and Terraform. Observability, capacity analytics, incident learnings, and performance data are used to continuously improve reliability, scalability, efficiency, and operational simplicity. Close collaboration with Cloud, Networking, Security, AI/ML, and infrastructure teams is central to solving complex problems that span multiple technology domains.

What We Need To See:

  • Bachelor’s degree in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field, or equivalent experience

  • 12+ years of relevant industry experience.

  • Proven success architecting, building, and operating large-scale distributed platforms or infrastructure systems in production.

  • Strong technical depth in several areas such as cloud infrastructure, distributed systems, networking, compute, storage, platform engineering, or content delivery.

  • Deep knowledge of Linux/Unix, TCP/IP, DNS, TLS, HTTP/S, proxies, load balancing, availability, scalability, and fault-tolerant system design.

  • Strong programming and automation skills with Python, Go, or similar languages, plus hands-on experience with infrastructure-as-code and orchestration.

  • Experience with AWS, Azure, or Google Cloud Platform and the ability to troubleshoot complex systems across application, operating system, network, and infrastructure layers.

  • Demonstrated ability to independently drive architecture and implementation across multiple teams, communicate effectively in complex situations, and mentor other engineers.

Ways To Stand Out From The Crowd:

  • Experience building platforms for AI/ML training, inference, model serving, GPU-accelerated workloads, distributed compute, or high-performance computing.

  • Deep expertise with CDN and edge platforms such as Akamai, AWS CloudFront, Fastly, or Cloudflare, including caching, origin design, WAF, DNS, TLS, and global traffic management.

  • Experience developing self-service platform capabilities that enable engineering teams to consume infrastructure reliably and at scale.

  • Proven use of SLIs, SLOs, error budgets, capacity analytics, and reliability metrics to deliver measurable improvements.

  • Experience distributing models, datasets, containers, software artifacts, or other large objects across globally distributed environments.

Want to work on infrastructure where scale, AI, networking, and reliability come together? Join us!

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/ 

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 200,000 USD - 322,000 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until August 23, 2026.

This posting is for an existing vacancy. 

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Skills

PythonAWSAzureKubernetesTerraformLinuxTCP/IPGo

Similar Jobs

30

Senior Staff Machine Learning Platform Engineer

Faire · Remote - US; San Francisco, CA +1 · Remote

Today

Senior Staff Data Platform Software Engineer (Persistence & Data Management)

ServiceNow · Hyderabad, India · Hybrid

Today

Senior Staff Platform Engineer

Nvidia · Santa Clara, CA,US, US

Yesterday

Senior Staff Engineer – Core Platform Products

Qualcomm · Hyderabad, TS,IN, IN

Yesterday

Senior Staff Engineer, Kubernetes Platform & Networking

Equinix · Bangalore Office BLS2, India · Hybrid

Yesterday

Senior /Staff Data & Cloud Platform Engineer

Micron Technology · Hyderabad, TS,IN, IN · Hybrid

3 days ago

Senior /Staff Data & Cloud Platform Engineer

Micron · Hyderabad - Phoenix Aquila, India · Hybrid

3 days ago

Sr. Staff Platform/Data Reliability Engineer, Databricks (R5537)

Shieldai · Remote · Remote

1 week ago

Senior/Staff - Agent Platform (Cortex Code)

Snowflake · US-CA-Menlo Park

1 week ago

Senior/Staff System Software Engineer — ADAS Vision Platform

Phantom AI · Mountain View, California, United States

1 week ago

Senior Staff Engineer - Platform & Partner Experience

Spotify · London +1 · Hybrid

1 week ago

Senior/Staff Platform Engineer (m/f/x)

Cortea AI · Berlin · Onsite

1 week ago

Senior/Staff Engineer - Exchange Middle Platform

OKX · Singapore, Singapore

1 week ago

Senior/Staff Software Engineer, Platform Infrastructure

Verkada · San Mateo, CA United States +1 · Onsite

2 weeks ago

Senior Staff Technical Product Manager, Vehicle Transaction Platform

Scoutmotors · Charlotte, North Carolina, United States; Fremont, California, United States +1 · Remote

2 weeks ago

Sr Staff Platform Engineer / Team Lead

Inari · OKS 7-501, United States of America +1

2 weeks ago

Sr. Staff Software Development Engineer - AI Platform

Zscaler · Bangalore, IND; Pune, IND +1 · Remote

2 weeks ago

Senior Staff Software Engineer, Platform

Sandboxaq · United States · Remote

3 weeks ago

Senior/Staff Software Engineer, Data Platform

Sovrn · Boulder, Colorado +1

3 weeks ago

Senior Staff Software Engineer, iOS Platform

"Gusto, Inc." · Denver, CO;San Francisco, CA;New York, NY;Seattle, WA +3 · Remote

3 weeks ago

Senior/Staff Embedded Software Engineer, Robotics Platform

Embedding Vc · Milpitas, CA · Onsite

3 weeks ago

Senior/Staff QA Engineer – Agentic AI Platform

Ontrac Solutions · Chicago, IL, US · Remote

3 weeks ago

Senior Staff Inbound Product Manager, ServiceNow SDK -Platform Agentic Build Lifecycle Team

ServiceNow · San Diego, California, United States · Remote

1 month ago

Senior Staff Client Platform Engineer

Nvidia · Santa Clara, CA,US, US +1 · Remote

1 month ago

Senior Staff Client Platform Engineer

Nvidia · US, CA, Santa Clara, United States of America +1 · Remote

1 month ago

Senior/Staff Platform Engineer

CodeRabbit · San Francisco · Hybrid

1 month ago

Senior/Staff/Principal Backend Engineer, Platform & Data Ingestion

Veranahealth · San Francisco, California, United States, New York, United States +2 · Remote, Hybrid, Onsite

1 month ago

Senior Staff Software Engineer - Cloud Platform

Lytx Careers · Office - San Diego, CA, United States of America +1 · Remote

1 month ago

Senior / Staff AI Platform Engineer

Clearstreet · Remote · Remote

1 month ago

Senior / Staff Platform Engineer

Radar · New York +1 · Onsite

1 month ago