At NVIDIA, we are pushing the boundaries of AI, graphics, and computing. The GitHub Actions Runner team manages self-hosted GPU-enabled GitHub Actions runners, using Actions Runner Controller with KubeVirt to deploy ephemeral VM-based runners for NVIDIA’s open source projects on GitHub. The team operates both on-premise and in the cloud (AWS) to support 100+ developers with whom they collaborate regularly to ensure a seamless CI/CD experience. We are looking for a Senior Infrastructure Engineer to help scale, optimize, and expand our platform.
Our team is fully remote and distributed across multiple time zones. If you're passionate about infrastructure, Kubernetes, automation, and observability, this is an opportunity to work with exciting technology at one of the most innovative companies in the world.
Preferred work location: Eastern/Central time zones
What you'll be doing:
Manage and scale self-hosted GitHub Actions runners using Kubernetes
Help expand runner support for various hardware and operating system combinations, including Linux, Windows, single-GPU, multi-GPU, NVLink, and more
Use Infrastructure as Code (Terraform and ArgoCD) to deploy and maintain infrastructure both on-premise and in AWS
Build and maintain runner VM images using HashiCorp Packer
Connect distributed services securely using mTLS, PKI, and HashiCorp Vault
Develop, package, and deploy custom Golang tools to support platform observability, stability, and efficiency
Configure alerting and monitoring to identify and address issues quickly, using tools like Prometheus and Grafana
Contribute upstream to open-source tools and libraries that our team depends on
Periodically update platform dependencies and address CVEs
What we need to see:
B.S. or M.S. in Computer Science, Computer Engineering, or a related field (or equivalent experience)
7+ years of proven experience in infrastructure, DevOps, or platform engineering
Strong Kubernetes expertise (running, debugging, and scaling workloads)
Experience with GitOps tools (ArgoCD or similar)
Proficiency in Linux administration and troubleshooting
Experience with Infrastructure as Code using Terraform/Terragrunt
Proficiency in Golang, Python, and TypeScript
Hands-on experience with monitoring, logging, and tracing (Prometheus, Grafana, OpenTelemetry, etc.)
Solid understanding of CI/CD pipelines, particularly GitHub Actions
Ability to work and collaborate effectively with a fully remote, distributed team
Ways to stand out from the crowd:
Experience instrumenting telemetry for distributed systems
Strong background in GPU workloads on Kubernetes with experience writing custom Kubernetes controllers
Deep understanding of KubeVirt and/or virtualization
Experience with self-hosted GitHub Actions runners
Contributions to open-source Kubernetes-related projects
With competitive salaries and a generous benefits package, NVIDIA is considered one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking individuals in the industry working for us. Due to unprecedented growth, our exclusive engineering teams are expanding rapidly. If you're a creative and autonomous engineer with a genuine passion for technology, we want to hear from you!
The base salary range is 168,000 USD - 333,500 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.
If an employer mentions a salary or salary range on their job, we display it as an "Employer Estimate". If a job has no salary data, Rise displays an estimate if available.
Lead and scale NVIDIA's global go-to-market sales strategies with strategic ISV partners to drive joint success and innovation in AI and enterprise solutions.
NVIDIA is looking for an experienced AI/ML Solutions Architect to drive large scale AI infrastructure and hyperscale cloud customer engagements.
Multiple remote OSP engineering and project management roles available for experienced professionals to support fiber construction nationwide.
Lead the strategic design and scaling of cutting-edge AI platforms at Palo Alto Networks to drive enterprise-wide AI innovation and impact.
Innovate electric vehicle powertrain design as a key engineer at Ford’s EVDD team committed to shaping the future of mobility.
Medtronic is looking for a highly motivated Field Service Engineer III to provide technical support and service for Cardiac Ablation Solutions across the Capitol Region.
Steer NVIDIA's Data Science and Data Engineering products as Vice President of Engineering, leading advanced software initiatives that bridge AI and traditional data processing.
Innovative clean mining startup Phoenix Tailings is searching for a Lead Electrical Engineer to spearhead electrical system design and implementation for sustainable metal production.
AECOM is looking for a skilled Fire Protection Engineer V with a PE license to lead building fire safety projects and contribute to their Tampa-based multidisciplinary team.
Contribute to humanity's next giant leap as an RF Engineer developing innovative communication systems for SpaceX’s Crew Starship lunar missions.
Innovative hardware company seeks a Product Design Engineer to develop and refine mechanical designs for their flagship tablet product in a highly collaborative environment.
Support innovative product and process development as an Associate Product Development Engineer at Porex in Fairburn, GA.
A Staff Engineer role at Samsung Semiconductor focused on architecture and analog mixed signal Serdes/RF circuit design with onsite presence in San Jose, CA.
Technical Engineering Specialist needed at Sonny's Enterprises to oversee car wash equipment performance, lead technical projects, and collaborate across teams to improve industrial automation solutions.
Lead the comprehensive system verification and validation efforts for innovative lunar landers at ispace U.S., advancing the future of lunar exploration.
NVIDIA is a publicly traded, multinational technology company headquartered in Santa Clara, California. NVIDIA's invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, and ignited the era of modern AI.
485 jobsSubscribe to Rise newsletter