Principal Solution Engineer, Large Scale ML Profiling Services

NVIDIA
Santa Clara CA
13 days ago
NVIDIA
NVIDIA
nvidia.com

Job Description

Our team builds scalable, always-available profiling services for ML applications in distributed environments! As a solutions engineer, you will work with the profiler team to understand performance monitoring services, APIs, and interfaces used by our customers. A strong understanding of large-scale ML performance analysis is essential. In this role, you will build strong customer relationships, have empathy for customer’s user difficulties, and be the customers’ advocate in our mission to drive adoption of our tools and GPUs.

What you’ll be doing:

You will plan and drive scalable profiling initiatives for ML customers. Prototype, design, and develop robust profiling solutions for large-scale ML use cases. Design actionable performance metrics for ML engineers to optimize workloads on CPU and GPU. Develop and implement deployment plans for software delivery and test profiling services in datacenter environments. Identify user friction points and emerging needs. Provide technical leadership to help strategies and status reports to senior leadership. Stay updated on the latest techniques and frameworks for AI training, deployment, and performance optimization.

What we need to see:

  • Ability to work in a multifaceted, fast-paced environment.

  • Experience with customer support, focusing on product and feature delivery

  • 15+ years of proven experience in system design, performance analysis, and shipping production software.

  • BS or higher degree in CS/EE/CE or equivalent experience.

  • Strong interpersonal skills.

  • Strong verbal and written communication skills

  • Experience deploying software in microservices distributed environments.

  • Strong software development experience in Python and C++.

  • Experience collaborating with different teams across the company.

Ways to stand out from the crowd:

  • Experience with performance analysis of AI training/inference applications.

  • Experience in building continuous profiling systems for GPU data centers.

  • Knowledge of GPU architecture and programming (CUDA, OpenCL).

  • Ability to tackle ambiguous situations and make them tractable.

The base salary range is 272,000 USD - 419,750 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Visit Original Source:

http://www.indeed.com/viewjob
why ?Jumpstart your career with our tech sales bootcamp!
Free Guides, Videos and Podcasts
  • The Biggest Red Flags in Sales Interviews: A Complete Guide
    The Biggest Red Flags in Sales Interviews: A Complete Guide
  • Career Change Guide: Breaking Into a Career in Tech Sales
    Career Change Guide: Breaking Into a Career in Tech Sales
  • How to Find a Second Career in Tech Sales
    How to Find a Second Career in Tech Sales
  • SDR Interviews | How to Land the Interview and Stand Out in the Process
    SDR Interviews | How to Land the Interview and Stand Out in the Process
  • See More…

Other Jobs

StackAdapt

Senior Enterprise BI Developer

StackAdapt

We have an exciting opportunity in the newly formed Enterprise Data Office (EDO) with its mandate to serve the business leaders and stakeholders at StackAdapt with trusted data, standard reporting fra

 
CA
Trusscore

Who we are Trusscore is a material science company focused on developing sustainable building materials. We're starting a journey to change the way people build buildings and the environmental foo

 
Calgary AB
Clio

Clio is more than just a tech company–we are a global leader that is transforming the legal experience for all by bettering the lives of legal professionals while increasing access to justice . Summa

 
Toronto ON / Remote