NVIDIA
Senior Site Reliability Engineer, DGX Cloud
Remote role where the employee must remain based in a particular country.
United States only
Employer listed it yesterday · Added yesterday
First listed yesterday.
Salary
$168,000 to $333,500
Location
United States only
Timezone
US East
Contract
Full-time
Experience
Senior
Category
Software
Published by the employer
Remote flexibility
Work from home
This is a remote role, but the employee must be based in United States. It is work from home rather than work from anywhere.
What the employer says
- Source listing states candidate location: "United States"
What Nomaders makes of it
- Residency required in United States
- Payroll and tax are likely handled in that country only
The quotes above are the employer's own words; the reading is ours. Always check the original listing and employment terms before working from another country.
About the role
NVIDIA DGX Cloud is developing and managing large-scale GPU infrastructure for AI research and production workloads. We are looking for Senior Reliability Engineers to help build the automation, tooling, and operational systems that make GPU clusters reliable, scalable, and safe to run. This role is part of a production engineering team passionate about Kubernetes-based infrastructure, GPU cluster operations, reliability, automation, GitOps, and Day 2 operability across DGX Cloud environments. What you’ll be doing:
Build and operate automation for large-scale Kubernetes clusters across NVIDIA Cloud Partners (NCP) and on-prem environments.
The full brief and the application link are for members.
Requirements, benefits, how to apply and the employer's own careers page, plus every other live role we track. £12.99/month, or £69 for a year.
Unlock this roleSimilar roles
Other open software roles with comparable remote rules.
$168,000 to $333,500 · Free accounts get one application link on us.