Senior Deep Learning Software Engineer, LLM Performance jobs in United States
cer-icon
Apply on Employer Site
company-logo

NVIDIA · 21 hours ago

Senior Deep Learning Software Engineer, LLM Performance

NVIDIA is seeking an experienced Deep Learning Engineer passionate about analyzing and improving the performance of LLM inference. The role involves optimizing and deploying deep learning models, collaborating with diverse teams, and contributing to NVIDIA's LLM frameworks and solutions.

Artificial Intelligence (AI)Consumer ElectronicsGPUHardwareSoftwareVirtual Reality
check
Growth Opportunities
check
H1B Sponsor Likelynote
Hiring Manager
Khai N.
linkedin

Responsibilities

Performance optimization, analysis, and tuning of LLM, VLM and GenAI models for DL inference, serving and deployment in NVIDIA/OSS LLM frameworks
Scale performance of LLM models across different architectures and types of NVIDIA accelerators
Scale performance for max throughput, minimum latency and throughput under latency constraints
Contribute features and code to NVIDIA/OSS LLM frameworks, inference benchmarking frameworks, TensorRT, and Triton
Work with cross-collaborative teams across generative AI, automotive, image understanding, and speech understanding to develop innovative solutions

Qualification

Deep LearningPerformance OptimizationPython/C/C++DL FrameworksGPU ProgrammingSoftware DesignSoftware EngineeringPerformance ModelingArchitectural Knowledge

Required

Bachelors, Masters, PhD, or equivalent experience in relevant fields (Computer Engineering, Computer Science, EECS, AI)
At least 8 years of relevant software development experience
Excellent Python/C/C++ programming, software design and software engineering skills
Experience with a DL framework like PyTorch, JAX, TensorFlow

Preferred

Prior experience with a LLM framework or a DL compiler in inference, deployment, algorithms, or implementation
Prior experience with performance modeling, profiling, debug, and code optimization of a DL/HPC/high-performance application
Architectural knowledge of CPU and GPU
GPU programming experience (CUDA or OpenCL)

Benefits

Equity
Benefits

Company

NVIDIA is a computing platform company operating at the intersection of graphics, HPC, and AI.

H1B Sponsorship

NVIDIA has a track record of offering H1B sponsorships. Please note that this does not guarantee sponsorship for this specific role. Below presents additional info for your reference. (Data Powered by US Department of Labor)
Distribution of Different Job Fields Receiving Sponsorship
Represents job field similar to this job
Trends of Total Sponsorships
2025 (1877)
2024 (1355)
2023 (976)
2022 (835)
2021 (601)
2020 (529)

Funding

Current Stage
Public Company
Total Funding
$4.09B
Key Investors
ARPA-EARK Investment ManagementSoftBank Vision Fund
2023-05-09Grant· $5M
2022-08-09Post Ipo Equity· $65M
2021-02-18Post Ipo Equity

Leadership Team

leader-logo
Jensen Huang
Founder and CEO
linkedin
leader-logo
Michael Kagan
Chief Technology Officer
linkedin
Company data provided by crunchbase