NVIDIA · 16 hours ago
Senior Software Engineer - Triton Tools
Maximize your interview chances
Insider Connection @NVIDIA
Get 3x more responses when you reach out via email instead of LinkedIn.
Responsibilities
Develop and enhance functionalities within the GenAI-Perf, Triton Performance Analyzer and Triton Model Analyzer tools.
Collaborate with researchers and engineers to understand their performance analysis needs and translate them into actionable features.
Collaborate closely with cross-functional teams including software engineers, system architects, and product managers to drive performance improvements throughout the development lifecycle.
Responsible for setting up, executing, and analyzing the performance of LLM, Generative AI and deep learning models.
Develop and implement efficient algorithms for measuring deep learning throughput and latency, benchmarking large language models, and deploying models.
Integrate various tools to create a unified and user-friendly experience for deep learning performance analysis.
Automate testing processes to ensure the quality and stability of the tools.
Contribute to technical documentation and user guides. Stay up-to-date on the latest advancements in deep learning performance analysis and LLM optimization techniques.
Qualification
Find out how your skills align with this job's requirements. If anything seems off, you can easily click on the tags to select or unselect skills to reflect your actual expertise.
Required
Bachelor's, Masters or PhD or equivalent experience
8+ years in Computer Science, computer architecture, or related field
Knowledge of distributed systems programming.
Ability to work in a fast-paced, agile team environment
Excellent Python programming and software design skills, including debugging, performance analysis, and test design.
Preferred
Experience with deep learning algorithms and frameworks. Especially experience with Large Language Models and frameworks such as PyTorch, TensorFlow, TensorRT, and ONNX Runtime.
Excellent troubleshooting abilities spanning multiple software (storage systems, kernels and containers).
Experience contributing to a large open source project - use of GitHub, bug tracking, branching and merging code, OSS licensing issues handling patches, etc.
Familiarity with cloud computing platforms (e.g., AWS, Azure, GCP) and Experience building and deploying cloud services using HTTP REST, gRPC, protobuf, JSON and related technologies.
Experience working with NVIDIA GPUs and deep learning inference frameworks is a plus.
Benefits
Equity and benefits
Company
NVIDIA
NVIDIA is a computing platform company operating at the intersection of graphics, HPC, and AI.
H1B Sponsorship
NVIDIA has a track record of offering H1B sponsorships. Please note that this does not
guarantee sponsorship for this specific role. Below presents additional info for your
reference. (Data Powered by US Department of Labor)
Distribution of Different Job Fields Receiving Sponsorship
Represents job field similar to this job
Trends of Total Sponsorships
2023 (735)
2022 (892)
2021 (696)
2020 (534)
Funding
Current Stage
Public CompanyTotal Funding
$4.09BKey Investors
ARPA-EARK Investment ManagementSoftBank Vision Fund
2023-05-09Grant· $5M
2022-08-09Post Ipo Equity· $65M
2021-02-18Post Ipo Equity· undefined
Recent News
2024-11-21
2024-11-21
GlobeNewswire News Room
2024-11-21
Company data provided by crunchbase