NVIDIA · 3 days ago
Senior HPC and AI Networking Performance Research and Analysis Engineer
NVIDIA is a leader in the field of AI and visual computing, seeking a Senior HPC and AI Networking Performance Research and Analysis Engineer to join their Performance group. The role involves profiling and analyzing AI workloads on large GPU and CPU clusters for distributed Deep Learning training, focusing on performance analysis and optimization in high-performance networking.
Artificial Intelligence (AI)Consumer ElectronicsGPUHardwareSoftwareVirtual Reality
Responsibilities
Exploring and researching AI workloads and DL models specifically tailored for large-scale deep learning LLM training on NVIDIA supercomputers and distributed systems focusing on high-performance networking and Nvidia Collective Communications Library (NCCL)
Benchmarking, Profiling, and Analyzing the performance to find bottlenecks and identify areas of improvement and optimizations, with a strong emphasis on networking aspects
Implementing performance analysis tools
Collaborating with many teams from hardware to software to provide performance analysis insights
Defining performance test planning , setting performance expectations for new technologies and solutions, and working to reach the performance targets limits
Qualification
Required
B.Sc in Computer Science or Software Engineering or equivalent experience
5+ years of experience with high-performance Networking (RDMA, MPI, NCCL, Congestion Control Algorithms)
Demonstrated Performance Analysis skills and methodologies
Experience with NVIDIA GPUs, CUDA library, deep learning frameworks like TensorFlow or PyTorch, combined with expertise in networking collective communication libraries (such as NCCL) and protocols (such as RoCE and RDMA)
Fast and self-learning capabilities with strong analytical and problem-solving skills
Programming Languages: Python, Bash and C languages
Experience with Linux OS distros
Great teammate with good communication and interpersonal skills
Preferred
In-depth knowledge and experience with AI workloads and benchmarking for distributed LLM training
Knowledge in CUDA, and NCCL libraries
Knowledge in Congestion Control algorithms
In-depth System knowledge and understanding (Intel / AMD / ARM CPUs, NVIDIA GPUs, HCA, Memory, PCI)
Strong Performance Analysis skills and methodologies using modern tools
Benefits
Equity
Benefits
Company
NVIDIA
NVIDIA is a computing platform company operating at the intersection of graphics, HPC, and AI.
H1B Sponsorship
NVIDIA has a track record of offering H1B sponsorships. Please note that this does not
guarantee sponsorship for this specific role. Below presents additional info for your
reference. (Data Powered by US Department of Labor)
Distribution of Different Job Fields Receiving Sponsorship
Represents job field similar to this job
Trends of Total Sponsorships
2025 (1877)
2024 (1355)
2023 (976)
2022 (835)
2021 (601)
2020 (529)
Funding
Current Stage
Public CompanyTotal Funding
$4.09BKey Investors
ARPA-EARK Investment ManagementSoftBank Vision Fund
2023-05-09Grant· $5M
2022-08-09Post Ipo Equity· $65M
2021-02-18Post Ipo Equity
Recent News
2026-01-07
2026-01-07
Company data provided by crunchbase