NVIDIA Logo

NVIDIA

Senior DevOps Infrastructure Engineer, Open-Source CI and CD

Posted 10 Days Ago
Remote
2 Locations
168K-334K
Senior level
Remote
2 Locations
168K-334K
Senior level
The Senior DevOps Infrastructure Engineer will manage GitHub Actions runners, optimize infrastructure using Kubernetes, Terraform, and promote CI/CD practices at NVIDIA.
The summary above was generated by AI

At NVIDIA, we are pushing the boundaries of AI, graphics, and computing. The GitHub Actions Runner team manages self-hosted GPU-enabled GitHub Actions runners, using Actions Runner Controller with KubeVirt to deploy ephemeral VM-based runners for NVIDIA’s open source projects on GitHub. The team operates both on-premise and in the cloud (AWS) to support 100+ developers with whom they collaborate regularly to ensure a seamless CI/CD experience. We are looking for a Senior Infrastructure Engineer to help scale, optimize, and expand our platform.

Our team is fully remote and distributed across multiple time zones. If you're passionate about infrastructure, Kubernetes, automation, and observability, this is an opportunity to work with exciting technology at one of the most innovative companies in the world.

Preferred work location: Eastern/Central time zones

What you'll be doing:

  • Manage and scale self-hosted GitHub Actions runners using Kubernetes

  • Help expand runner support for various hardware and operating system combinations, including Linux, Windows, single-GPU, multi-GPU, NVLink, and more

  • Use Infrastructure as Code (Terraform and ArgoCD) to deploy and maintain infrastructure both on-premise and in AWS

  • Build and maintain runner VM images using HashiCorp Packer

  • Connect distributed services securely using mTLS, PKI, and HashiCorp Vault

  • Develop, package, and deploy custom Golang tools to support platform observability, stability, and efficiency

  • Configure alerting and monitoring to identify and address issues quickly, using tools like Prometheus and Grafana

  • Contribute upstream to open-source tools and libraries that our team depends on

  • Periodically update platform dependencies and address CVEs

What we need to see:

  • B.S. or M.S. in Computer Science, Computer Engineering, or a related field (or equivalent experience)

  • 7+ years of proven experience in infrastructure, DevOps, or platform engineering

  • Strong Kubernetes expertise (running, debugging, and scaling workloads)

  • Experience with GitOps tools (ArgoCD or similar)

  • Proficiency in Linux administration and troubleshooting

  • Experience with Infrastructure as Code using Terraform/Terragrunt

  • Proficiency in Golang, Python, and TypeScript

  • Hands-on experience with monitoring, logging, and tracing (Prometheus, Grafana, OpenTelemetry, etc.)

  • Solid understanding of CI/CD pipelines, particularly GitHub Actions

  • Ability to work and collaborate effectively with a fully remote, distributed team

Ways to stand out from the crowd:

  • Experience instrumenting telemetry for distributed systems

  • Strong background in GPU workloads on Kubernetes with experience writing custom Kubernetes controllers

  • Deep understanding of KubeVirt and/or virtualization

  • Experience with self-hosted GitHub Actions runners

  • Contributions to open-source Kubernetes-related projects

With competitive salaries and a generous benefits package, NVIDIA is considered one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking individuals in the industry working for us. Due to unprecedented growth, our exclusive engineering teams are expanding rapidly. If you're a creative and autonomous engineer with a genuine passion for technology, we want to hear from you!

The base salary range is 168,000 USD - 333,500 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Top Skills

Argocd
Github Actions
Go
Grafana
Hashicorp Packer
Kubernetes
Opentelemetry
Prometheus
Python
Terraform
Typescript

Similar Jobs

Yesterday
Remote
Hybrid
2 Locations
135K-250K Annually
Expert/Leader
135K-250K Annually
Expert/Leader
Artificial Intelligence • Cloud • Sales • Security • Software • Cybersecurity • Data Privacy
The Senior Staff DevOps Engineer will lead the Infrastructure Platform Team, design and implement SaaS infrastructure, and automate operations for high availability. Responsibilities include improving cloud architecture, managing cross-functional teams, and ensuring security compliance.
Top Skills: AWSKubernetesPythonRubyShell ScriptingTerraform
5 Days Ago
Remote
Hybrid
USA
135K-215K Annually
Senior level
135K-215K Annually
Senior level
Cloud • Computer Vision • Information Technology • Sales • Security • Cybersecurity
The Sr. DevOps Engineer will ensure smooth operations of applications through CI/CD pipelines, optimize infrastructure, automate processes, and enhance reliability and scalability.
Top Skills: BitbucketChrome ExtensionsCi/CdDockerFastapiFlaskGoJenkinsKubernetesPython
5 Days Ago
Remote
Washington, DC, USA
170K-200K Annually
Mid level
170K-200K Annually
Mid level
Artificial Intelligence • Information Technology
The DevOps Engineer at ConductorAI will design system architecture, improve CI/CD processes, and ensure compliance on DoD projects.
Top Skills: Ci/CdDockerGithub ActionsHelmKubernetesPythonTypescript

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account