Senior Solution Architect, HPC and AI - NVIS @ NVIDIA | Jobright.ai
JOBSarrow
RecommendedLiked
0
Applied
0
External
0
Senior Solution Architect, HPC and AI - NVIS jobs in United States
Be an early applicantLess than 25 applicants
company-logo

NVIDIA · 19 hours ago

Senior Solution Architect, HPC and AI - NVIS

ftfMaximize your interview chances
Artificial Intelligence (AI)GPU
check
Growth Opportunities
check
H1B Sponsor Likelynote
Hiring Manager
Crystal Aggarwal (Crystal Amarante)
linkedin

Insider Connection @NVIDIA

Discover valuable connections within the company who might provide insights and potential referrals.
Get 3x more responses when you reach out via email instead of LinkedIn.

Responsibilities

Primary responsibilities will include building and enabling robust AI/HPC infrastructure for customers
Support operational and reliability aspects of large-scale AI clusters, focusing on performance at scale, training stability, real-time monitoring, logging, and alerting
Engage in and improve services from inception and design through deployment, operation, and optimization
Co-design telemetry of AI workloads to help engineering build solutions for more robust workloads at scale
Communicate across internal teams to support the continuous improvement of NVIDIA's offerings and software designs

Qualification

Find out how your skills align with this job's requirements. If anything seems off, you can easily click on the tags to select or unselect skills to reflect your actual expertise.

PythonC++AI trainingNLP librariesHPC systemsMicroservices architectureCloud platformsParallel filesystemsMachine LearningDeep LearningKubernetesTechnical leadershipTime-management skills

Required

Strong foundational expertise, from a BS, MS, or Ph.D. degree in Engineering, Mathematics, Physics, Computer Science, Data Science, or similar (or equivalent experience).
8+ years of experience and knowledge of neural networks including good understanding of transformer architectures.
Experience designing large scale AI workloads with SLURM and/or Kubernetes.
Proficiency with Python / C++ / Rust or other popular software languages.
Excellent verbal, written communication, and technical presentation skills in English.
You are motivated to work with multiple levels and teams across organizations.
Strong analytical and problem-solving skills.
Strong time-management and organization skills for coordinating multiple initiatives, priorities and implementations of new technology and products into very sophisticated projects.
You are a curious self-starter with a desire for continuous learning and sharing knowledge across the team.

Preferred

Experience orchestrating distributed Deep Learning training with SLURM.
Proficiency in DevOps, including hands-on experience with Ansible, Terraform or similar tools. Equivalent experience will be accepted as well.
8+ years designing solutions with one or more Tier-1 Clouds (AWS, Azure, GCP or OCI) and cloud-native architectures and software.
Technical leadership with a strong understanding of NVIDIA technologies, and success in working with customers.
Expertise with parallel file systems (e.g. Lustre, GPFS, BeeGFS, WekaIO) and high-speed interconnects (InfiniBand, Omni Path, and Gig-E).
Experience with integration and deployment of software products in production enterprise environments, and microservices software architecture.

Benefits

Equity and benefits

Company

NVIDIA is a computing platform company operating at the intersection of graphics, HPC, and AI.

H1B Sponsorship

NVIDIA has a track record of offering H1B sponsorships. Please note that this does not guarantee sponsorship for this specific role. Below presents additional info for your reference. (Data Powered by US Department of Labor)
Distribution of Different Job Fields Receiving Sponsorship
Represents job field similar to this job
Trends of Total Sponsorships
2023 (735)
2022 (892)
2021 (696)
2020 (534)

Funding

Current Stage
Public Company
Total Funding
$4.09B
Key Investors
ARPA-EARK Investment ManagementSoftBank Vision Fund
2023-05-09Grant· $5M
2022-08-09Post Ipo Equity· $65M
2021-02-18Post Ipo Equity· undefined

Leadership Team

leader-logo
Jensen Huang
CEO and Founder
linkedin
leader-logo
Chris Malachowsky
Co-Founder, SVP
linkedin
Company data provided by crunchbase
logo

Orion

Your AI Copilot