





Strong employer brand plus metro location but specialized Kubernetes and networking requirements reduce applicant density.
Deep Kubernetes, cluster lifecycle, and network infrastructure expertise required, limiting cross-industry transferability.
Explicit 8+ years requirement and mandatory Kubernetes, GitOps, Go/Python, and production on-call experience make filters strict.
Login to See Your Match Score
Create a free account or log in to unlock your CV match score across:
Own end-to-end lifecycle management of Kubernetes platform powering NVIDIA Global Network Infrastructure, including cluster provisioning, upgrades, capacity, availability, and recovery.
Develop and maintain automation and software for Kubernetes cluster operations such as provisioning, validation, upgrades, and multi-cluster delivery using GitOps.
Provide production support and lead incident response for network services running on the Kubernetes platform, ensuring resolution and corrective actions through to completion.
Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent experience.
8+ years experience building or operating production Kubernetes platforms, network infrastructure, or distributed systems.
Deep expertise with Kubernetes at scale including cluster lifecycle, upgrades, networking, storage, and recovery.
Proficiency in at least one general-purpose programming language (e.g., Go, Python); experience with GitOps, infrastructure as code, CI/CD, and production on-call/incident management required.
Senior engineer experienced operating large-scale, multi-region Kubernetes platforms with deep Kubernetes lifecycle and automation expertise.
Hands-on experience with GitOps and infrastructure automation for reliable, scalable multi-cluster Kubernetes deployments.
Background supporting network automation or telemetry services on Kubernetes, including diagnosing complex platform issues and driving production incident response effectively.