You and the process

With a person
Anna, in-house recruiter
9 years hiring engineers
30 minutes with a real recruiter
They read your CV with you, on a call, and say where the offers are being lost.
Didn’t find what you were looking for? Tell us what to build
Be the first to open itNo views yet
NVIDIA

Senior Manager, Kubernetes Runtime Engineering

  • Office
  • 6+ years

Salary

$272,000 - 431,250/ year

AI summary

For members

The whole posting in a few lines. Sign up to read it here and on every role you open.

Sign Up to Read

Description

The NVIDIA Kubernetes Engine (NKE) team is looking for a technical leader to lead the Runtime Engineering team responsible for the full configuration lifecycle of NKE tenant workload clusters. This team is responsible for software components that keep GPU workloads reliable and secure at scale. Their scope includes cluster bootstrapping, node configuration, and the container execution environment, including NVIDIA's AI Container Runtime (AICR). You will work across networking, storage, GPU resource management, and cluster security to deliver a production-grade, multi-tenant Kubernetes platform. Your team's decisions directly shape the runtime foundation that internal and external customers depend on.

What You'll Be Doing:

Be responsible for the build, implementation, and operational reliability of cluster configurations for NKE tenant workloads across all supported topologies

Manage a team of engineers coordinating the entire container runtime stack: AICR, GPU management operator, DCGM, and related node-level components

Drive architecture decisions for cluster networking (CNI), storage (CSI), cluster HA , and GPU resource partitioning (MIG, MPS, time-slicing)

Define and implement cluster hardening standards, RBAC models, pod security policies, and multi-tenancy isolation boundaries

Partner with NKE platform, infrastructure, and cybersecurity teams to integrate new capabilities and resolve cross-cutting runtime concerns

Build and maintain tooling for AICR lifecycle management — provisioning, upgrades, configuration drift detection, and remediation

Represent the runtime team in architecture reviews, roadmap planning, and customer communications with NVIDIA leadership

Contribute to open source communities anywhere NKE has upstream dependencies or influence

What We Need to See:

BS/MS degree in Computer Science or related field (or equivalent experience)

12+ overall years of relevant experience designing and delivering large-scale distributed software systems, including 5+ years of people-management experience leading, developing, and scaling high-performing software engineering teams responsible for complex, production-critical software.

Experience leading a group of engineers with varying specializations and seniority levels — bridging runtime, networking, and security fields is a core part of this role

Kubernetes internals knowledge — not just usage; you understand how the scheduler, kubelet, API server, and admission controllers interact

Cluster lifecycle management experience — Cluster API, kubeadm, or equivalent; experience leading fleet-scale cluster provisioning and upgrades

Security and compliance posture — CIS Kubernetes Benchmark, pod security admission, image signing, supply chain integrity

Proven ability to design and implement maintainable APIs for consumers

Familiarity with Identity and Access Management approaches

Excel in managing up, down, and across organizations

Demonstrated ability to reach cross-organization consensus without all the details

Ways to Stand Out from the crowd:

Prior experience with NVIDIA GPU Operator, DCGM Exporter, or NVLink-aware scheduling

Experience running Kubernetes at hyperscale with GPU node pools

Track record of upstream open source contributions in the Kubernetes or any open source runtime ecosystem

Experienced, persuasive, and effective interpersonal skills — written, verbal, and in front of engineering leadership

Demonstrated skills in coaching, analysis, problem solving, and short/long-term technical planning

NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing, and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction — from artificial intelligence to autonomous vehicles. NVIDIA is widely considered one of the technology world's most desirable employers. We have some of the most forward-thinking and hard-working people in the world working for us. If you're passionate about building the infrastructure that runs AI at scale, we want to hear from you.

Widely considered to be one of the technology world’s most desirable employers, NVIDIA offers highly competitive salaries and a comprehensive benefits package. As you plan your future, see what we can offer to you and your family www.nvidiabenefits.com/

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 272,000 USD - 431,250 USD.

Responsibilities

  • Applications for this job will be accepted at least until October 3, 2026.
  • This posting is for an existing vacancy.
  • NVIDIA uses AI tools in its recruiting processes.
  • NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Where you’d work

From the office

About the company

NVIDIA

Offices in United States, Seattle, United States

Also hiring in Munich, Germany, France, Bristol, United Kingdom and 12 more places

273 of their 660 open roles are remote

Your chances

Worth a look before you spend an evening tailoring a CV for it.

  • 21 checks run
  • 1 red flag

Still hiring?

14 checks

No red flags

How crowded?

7 checks

1 red flag

Fits Me

How well does this role fit you?

Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.

  • Your field
  • Level
  • Stack
  • Work model
  • Salary floor
  • Must-haves
Details12 facts · Role, Location, Compensation, Employment
Tech stack
  • Kubernetes
Seniority
Senior
Type
Full-time
Equity
Equity offered
Specialty
DevOps
Region
United States
Show 6 more factsShow less

Role

Category
DevOps & Infrastructure
Specialty
DevOps
Seniority
Senior
Experience
5+ years
Tech stack
  • Kubernetes

Location

Work model
Office
Region
United States
Offices
  • United States
  • Seattle, United States

Compensation

Salary
$272,000 - 431,250 / year
Pay period
Annual
Equity
Equity offered

Employment

Type
Full-time

Something wrong with this vacancy?

Similar vacancies

  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Software Engineer (DevOps)

    Salary by agreement

    • Office · Krakow
    • Junior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Principal Platform Engineer

    $255,000 - 280,000 / year

    • Remote
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Senior Staff Cloud Platform Engineer

    $230,000 / year

    • Remote · US, Canada
    • Senior

Share this vacancy

What's wrong with it?

The employer never sees who reported.

Reason