Compute Architecture Software Engineer
NVIDIA
About this role
An LLM Inference Software Engineer at NVIDIA will work on the TRTLLM project to accelerate large language model inference using GPU technology across environments from single PCs to large GPU clusters. The role is part of a collaborative engineering team focused on advancing AI infrastructure and performance.
Skills
Qualifications
About NVIDIA
nvidia.comNVIDIA invents the GPU and drives advances in AI, HPC, gaming, creative design, autonomous vehicles, and robotics.
Recent company news
NVIDIA Ignites the Next Industrial Revolution in Knowledge Work With Open Agent Development Platform
2 days ago
Nvidia CEO Says Company Is Firing Up H200 Production for China
1 day ago
Nvidia CEO Huang says company sees more than $1 trillion in sales through 2027
1 day ago
Nvidia's one of the fastest growing companies with one of the lowest valuations, says Jim Cramer
14 hours ago
Nvidia is reskinning games with AI. Gamers are angry about it, and wrong
1 day ago
About NVIDIA
Headquarters
San Francisco, CA
Company Size
201-500 employees
Founded
2018
Industry
Technology
Glassdoor Rating
4.2 / 5
Leadership Team
Sarah Johnson
Chief Executive Officer
Michael Chen
Chief Technology Officer
Emily Williams
VP of Engineering
David Rodriguez
VP of Product
Jessica Thompson
Chief Financial Officer
Andrew Park
VP of Sales
Unlock Company Insights
View leadership team, funding history,
and employee contacts for NVIDIA.
Salary
$163k – $219k
per year
More jobs at NVIDIA
Similar Jobs
Software Engineer, Model Performance Tooling
Baseten
Senior Software Engineer - Model Performance
Inference
Distributed Training & Inference Optimization Engineer (LLM) - GPU Optimization Department (GPUOD)
Rakuten Group, Inc.
Member of Technical Staff - Mid-Training Infra
Reflection AI
ML Platform Engineer
eBay
Research Engineering, Inference
Bitdeer