AI Software Engineer, LLM Inference Performance Analysis - New College Grad 2026
NVIDIA
About this role
A Software Engineer, Performance Analysis and Optimization for LLM Inference at NVIDIA focuses on improving the efficiency and scalability of large language model inference on NVIDIA computing platforms. The role centers on advancing compiler and kernel infrastructure to shape runtime behavior and hardware utilization for next-generation LLM deployments across data center and embedded platforms. The position requires close collaboration with compiler, hardware, kernel, and framework teams and influences performance of deployed models.
Skills
Qualifications
About NVIDIA
nvidia.comNVIDIA invents the GPU and drives advances in AI, HPC, gaming, creative design, autonomous vehicles, and robotics.
Recent company news
NVIDIA Ignites the Next Industrial Revolution in Knowledge Work With Open Agent Development Platform
2 days ago
Nvidia CEO Says Company Is Firing Up H200 Production for China
1 day ago
Nvidia CEO Huang says company sees more than $1 trillion in sales through 2027
1 day ago
Nvidia's one of the fastest growing companies with one of the lowest valuations, says Jim Cramer
14 hours ago
Nvidia is reskinning games with AI. Gamers are angry about it, and wrong
1 day ago
About NVIDIA
Headquarters
San Francisco, CA
Company Size
201-500 employees
Founded
2018
Industry
Technology
Glassdoor Rating
4.2 / 5
Leadership Team
Sarah Johnson
Chief Executive Officer
Michael Chen
Chief Technology Officer
Emily Williams
VP of Engineering
David Rodriguez
VP of Product
Jessica Thompson
Chief Financial Officer
Andrew Park
VP of Sales
Unlock Company Insights
View leadership team, funding history,
and employee contacts for NVIDIA.
Salary
$124k – $219k
per year
More jobs at NVIDIA
Similar Jobs
Member of Technical Staff, Inference
Ashby
Senior Hardware/Software ML Inference IP and Compiler Developer
Altera
Performance Engineer - Inference
Cerebras Systems
Software Engineer, ML Inference Performance
SambaNova Systems
Principal Software Engineer China Beijing Beijing
Microsoft
Technical Program Manager, Inference Performance
Anthropic