Senior Solution Architect, HPC and AI - NVIS (Santa Clara) Job at NVIDIA, Santa Clara, CA

TWJPaVJCVW9RWWFwYzZXSVFUeERKOTFNclE9PQ==
  • NVIDIA
  • Santa Clara, CA

Job Description

Senior Solution Architect, HPC and AI - NVIS

Join to apply for the Senior Solution Architect, HPC and AI - NVIS role at NVIDIA.

Do you want to be part of the team that brings Artificial Intelligence (AI) emerging technology to the field? We are looking for a hardworking Solution Architect (SA) to join the NVIDIA AI Enterprise (NVAIE) SA Segment Team. The mission of the NVAIE Segment team is to guide and enable the successful adoption at scale of NVIDIA AI Enterprise Software in production.

In our Solutions Architecture team, we work with NVIDIA's pioneering hardware and software, driving the latest breakthroughs in artificial intelligence. We need people who enable customer adoption of NVIDIA technology and develop lasting relationships with our technology partners, making NVIDIA a key design choice for enduser solutions. On this team, you will support full stack deployment including architectural designs, workload orchestration and application optimization. At NVIDIA, you will be immersed in a diverse, encouraging environment where everyone is inspired to do their life's work. Come join the team and see how you can make a lasting impact on the world!

What Youll Be Doing

  • Primary responsibilities will include building and enabling robust AI/HPC infrastructure for customers
  • Support operational and reliability aspects of largescale AI clusters, focusing on performance at scale, training stability, realtime monitoring, logging, and alerting
  • Engage in and improve services from inception and design through deployment, operation, and optimization
  • Codesign telemetry of AI workloads to help engineering build solutions for more robust workloads at scale
  • Communicate across internal teams to support the continuous improvement of NVIDIA's offerings and software designs

What We Need To See

  • Strong foundational expertise, from a BS, MS, or Ph.D. degree in Engineering, Mathematics, Physics, Computer Science, Data Science, or similar (or equivalent experience).
  • 8+ years of experience and knowledge of neural networks including good understanding of transformer architectures. Experience designing large scale AI workloads with SLURM and/or Kubernetes
  • Proficiency with Python / C++ / Rust or other popular software languages
  • Excellent verbal, written communication, and technical presentation skills in English
  • You are motivated to work with multiple levels and teams across organizations
  • Strong analytical and problemsolving skills
  • Strong timemanagement and organization skills for coordinating multiple initiatives, priorities and implementations of new technology and products into very sophisticated projects
  • You are a curious selfstarter with a desire for continuous learning and sharing knowledge across the team

Ways To Stand Out From The Crowd

  • Experience orchestrating distributed Deep Learning training with SLURM
  • Proficiency in DevOps, including handson experience with Ansible, Terraform or similar tools. Equivalent experience will be accepted as well.
  • 8+ years designing solutions with one or more Tier1 Clouds (AWS, Azure, GCP or OCI) and cloudnative architectures and software
  • Technical leadership with a strong understanding of NVIDIA technologies, and success in working with customers
  • Expertise with parallel file systems (e.g. Lustre, GPFS, BeeGFS, WekaIO) and highspeed interconnects (InfiniBand, Omni Path, and GigE)
  • Experience with integration and deployment of software products in production enterprise environments, and microservices software architecture

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is $148,000 $235,750 for Level 4, and $176,000 $276,000 for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until July 29, 2025. NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

#J-18808-Ljbffr

Job Tags

Full time,

Similar Jobs

DAVITA

Registered Nurse - Inpatient Dialysis RN Job at DAVITA

 ...meaningful impact in acute care nephrology. DaVita is seeking an RN to provide dialysis care in a hospital setting for patients with...  ...: $37-$51/hr (Final determination will be based on relevant experience) What Youll Do: Deliver inpatient dialysis therapies, including... 

micro1

FT Data Entry Clerk - Work From Home Job at micro1

 ...records of orders, customer correspondences, and inventory movements; Respond efficiently to customer inquiries regarding order status, shipping details, and product information; Monitor and track orders, proactively resolving discrepancies or delays...Hiring Immediately... 

FedEx

Retail Senior Store Manager Job at FedEx

 ...POSITION SUMMARY: The Senior Store Manager and Flagship Store Manager positions are critical to the successful operations of FedEx Offices largest and most impactful retail stores. You will run and grow your business while maintaining Purple Promise service, operational... 

Cavco

Drafter Job at Cavco

 ...Industries, Inc. (NASDAQ CVCO), our 7000 team members are at the heart of everything we do. We design and produce quality, affordable factory-built homes. We are also a leading producer of park model RVs, vacation cabins and factory-built commercial structures. In addition... 

52X Consulting LLC

Junior Recruiter Job at 52X Consulting LLC

 ...Job Description Junior Recruiter (Entry Level Friendly) Architecture Engineering Construction (AEC) Industry Location: Jacksonville, FL Schedule: Hybrid (4 days in-office / 1 day remote) About the Role Are you a natural connector who enjoys talking...