Senior Site Reliability Engineer - Datacenter Automation
NVIDIA
About this role
NVIDIA is seeking experienced Site Reliability Engineers to help scale its AI infrastructure, focusing on building reliable, scalable GPU clusters and supporting AI workloads across multiple cloud platforms. The role involves working closely with cross-functional teams to improve system availability, performance, and automation, contributing to NVIDIA's cutting-edge AI computing solutions.
Skills
Qualifications
About NVIDIA
nvidia.comNVIDIA invents the GPU and drives advances in AI, HPC, gaming, creative design, autonomous vehicles, and robotics.
Recent company news
NVIDIA Ignites the Next Industrial Revolution in Knowledge Work With Open Agent Development Platform
2 days ago
Nvidia CEO Says Company Is Firing Up H200 Production for China
1 day ago
Nvidia CEO Huang says company sees more than $1 trillion in sales through 2027
1 day ago
Nvidia's one of the fastest growing companies with one of the lowest valuations, says Jim Cramer
14 hours ago
Nvidia is reskinning games with AI. Gamers are angry about it, and wrong
1 day ago
About NVIDIA
Headquarters
San Francisco, CA
Company Size
201-500 employees
Founded
2018
Industry
Technology
Glassdoor Rating
4.2 / 5
Leadership Team
Sarah Johnson
Chief Executive Officer
Michael Chen
Chief Technology Officer
Emily Williams
VP of Engineering
David Rodriguez
VP of Product
Jessica Thompson
Chief Financial Officer
Andrew Park
VP of Sales
Unlock Company Insights
View leadership team, funding history,
and employee contacts for NVIDIA.
More jobs at NVIDIA
Similar Jobs
Site Reliability Engineer, AI/ML Infrastructure
Boson AI
Staff Cloud Site Reliability Engineer
Wayve
Site Reliability Engineer - AI & ML Infrastructure (Kubernetes & Terraform)
Deepgram
Network Site Reliability Engineer (NetSRE)
Nebius
Site Reliability Engineer, AI/ML Infrastructure
Boson AI
Site Reliability Engineer - AI & ML Infrastructure (Kubernetes, AWS & Terraform)
Deepgram