Senior Software Engineer - Distributed Inference

1 Month ago • 8 Years + • Artificial Intelligence • $184,000 PA - $356,500 PA

Job Summary

Job Description

NVIDIA seeks a Senior Software Engineer to enhance user-facing tools for Triton Inference Server, focusing on performance analysis for Generative AI and LLMs. Responsibilities include developing functionalities within GenAI-Perf and Triton analyzers, collaborating with researchers and engineers on performance needs, implementing efficient algorithms for measuring deep learning performance, integrating tools for a user-friendly experience, automating testing, contributing to documentation, and staying current on deep learning advancements. The role involves working with deep learning frameworks like PyTorch, TensorFlow, and TensorRT.
Must have:
  • 8+ years experience in relevant field
  • Distributed systems programming knowledge
  • Excellent Python programming skills
  • Performance analysis and test design expertise
  • Collaboration with cross-functional teams
Good to have:
  • Deep learning algorithm and framework experience (LLMs, PyTorch, TensorFlow, TensorRT, ONNX Runtime)
  • Troubleshooting skills across software systems
  • Open source project contribution experience
  • Cloud computing platform familiarity (AWS, Azure, GCP)
  • Experience with NVIDIA GPUs and inference frameworks
Perks:
  • Equity
  • Benefits

Job Details

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. Come join the team and see how you can make a lasting impact on the world.

We are now looking for a Senior System Software Engineer to work on user facing tools for Triton Inference Server! NVIDIA is hiring software engineers for its GPU-accelerated deep learning software team, and we are a remote friendly work environment. Academic and commercial groups around the world are using GPUs to power a revolution in deep learning, enabling breakthroughs in problems from LLM, image classification to speech recognition to natural language processing. We are a fast-paced team building tools and software to make the design and deployment of new deep learning models easier and accessible to more data scientists.

What you'll be doing:

  • Develop and enhance functionalities within the GenAI-Perf, Triton Performance Analyzer and Triton Model Analyzer tools.

  • Collaborate with researchers and engineers to understand their performance analysis needs and translate them into actionable features.

  • Collaborate closely with cross-functional teams including software engineers, system architects, and product managers to drive performance improvements throughout the development lifecycle.

  • Responsible for setting up, executing, and analyzing the performance of LLM, Generative AI and deep learning models.

  • Develop and implement efficient algorithms for measuring deep learning throughput and latency, benchmarking large language models, and deploying models.

  • Integrate various tools to create a unified and user-friendly experience for deep learning performance analysis.

  • Automate testing processes to ensure the quality and stability of the tools.

  • Contribute to technical documentation and user guides. Stay up-to-date on the latest advancements in deep learning performance analysis and LLM optimization techniques.

What we need to see:

  • Bachelor's, Masters or PhD or equivalent experience

  • 8+ years in Computer Science, computer architecture, or related field

  • Knowledge of distributed systems programming.

  • Ability to work in a fast-paced, agile team environment

  • Excellent Python programming and software design skills, including debugging, performance analysis, and test design.

Ways to stand out from the crowd:

  • Experience with deep learning algorithms and frameworks. Especially experience with Large Language Models and frameworks such as PyTorch, TensorFlow, TensorRT, and ONNX Runtime.

  • Excellent troubleshooting abilities spanning multiple software (storage systems, kernels and containers).

  • Experience contributing to a large open source project - use of GitHub, bug tracking, branching and merging code, OSS licensing issues handling patches, etc.

  • Familiarity with cloud computing platforms (e.g., AWS, Azure, GCP) and Experience building and deploying cloud services using HTTP REST, gRPC, protobuf, JSON and related technologies.

  • Experience working with NVIDIA GPUs and deep learning inference frameworks is a plus.

NVIDIA has continuously reinvented itself over three decades. Our invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined modern computer graphics, and revolutionized parallel computing. More recently, GPU deep learning ignited modern AI — the next era of computing. We are widely considered to be the leader of AI computing, and one of the technology world’s most desirable employers. We have some of the most forward-thinking and committed people in the world working for us. If you're creative and autonomous, we want to hear from you!

The base salary range is 184,000 USD - 356,500 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Similar Jobs

Krafton  - [AI] Deep Learning Service Dev - Backend Engineer (3년 이상)

Krafton

Seoul, South Korea (On-Site)
5 Months ago
NVIDIA - Silicon Reliability Engineer

NVIDIA

Santa Clara, California, United States (Hybrid)
2 Months ago
ByteDance - Student Researcher (Doubao (Seed) - Generative AI)

ByteDance

San Jose, California, United States (Hybrid)
2 Weeks ago
NVIDIA - Senior Mixed Signal Designer Engineer

NVIDIA

Taipei City, Taiwan (On-Site)
3 Months ago
ByteDance - Senior Machine Learning Engineer

ByteDance

San Jose, California, United States (On-Site)
2 Weeks ago
Google - Software Engineer, AICore, Platforms and Devices

Google

Taipei City, Taiwan (On-Site)
2 Weeks ago
Meta - AI Research Scientist, Language - Generative AI

Meta

New York, New York, United States (On-Site)
5 Months ago
Google - Student Researcher, PhD, Winter/Summer 2025

Google

Ann Arbor, Michigan, United States (On-Site)
5 Months ago
Mashgin - Senior Software Engineer, Machine Learning and Artificial Intelligence

Mashgin

Palo Alto, California, United States (Hybrid)
6 Months ago
Henkel - Data Scientist-Intern

Henkel

Pune, Maharashtra, India (On-Site)
7 Months ago

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

NVIDIA - Product Marketing Manager - Inference Optimization Software

NVIDIA

Santa Clara, California, United States (On-Site)
1 Month ago
ByteDance - Research Scientist Intern (Doubao (Seed) - Music Foundation Model) - 2024 Summer (PhD)

ByteDance

San Jose, California, United States (On-Site)
6 Months ago
NVIDIA - Enterprise Software Test Development Engineer

NVIDIA

Taipei City, Taiwan (On-Site)
4 Weeks ago
Google - Machine Learning Algorithm Engineer, Silicon

Google

Mountain View, California, United States (On-Site)
2 Weeks ago
NVIDIA - Software Partner Marketing Manager

NVIDIA

Santa Clara, California, United States (On-Site)
3 Weeks ago
Niantic - 2025 R&D Software Engineering Intern (PhD, Publishing)

Niantic

London, England, United Kingdom (Hybrid)
5 Months ago
Genies - Engineering Manager, Machine Learning

Genies

Los Angeles, California, United States (On-Site)
1 Month ago
NVIDIA - Payroll Specialist

NVIDIA

Shanghai, Shanghai, China (On-Site)
2 Weeks ago
Ubisoft - Machine Learning Programmer (Character & Animation)

Ubisoft

Montreal, Quebec, Canada (On-Site)
1 Month ago
Google - Software Engineer III, Machine Learning, Google Ads

Google

Kirkland, Washington, United States (On-Site)
2 Weeks ago

Get notifed when new similar jobs are uploaded

Jobs in Texas, United States

NVIDIA - Senior Firmware Security Engineer

NVIDIA

Santa Clara, California, United States (On-Site)
1 Week ago
Gupta - Media Analyst

Gupta

Boston, Massachusetts, United States (On-Site)
1 Week ago
Google - Software Engineer III, Site Reliability Engineering

Google

San Francisco, California, United States (On-Site)
2 Weeks ago
Epic Games - Senior Development Manager

Epic Games

Cary, North Carolina, United States (On-Site)
3 Months ago
Google - Programmatic Account Manager, Food, Beverage and Restaurants

Google

Chicago, Illinois, United States (On-Site)
2 Weeks ago
Axon - Automation Technician II

Axon

Phoenix, Arizona, United States (On-Site)
7 Hours ago
Whatnot - Director, Data Science (Revenue Analytics)

Whatnot

Los Angeles, California, United States (Remote)
6 Months ago
Google - Senior Software Engineer, Generative AI, Google Cloud AI

Google

Sunnyvale, California, United States (On-Site)
2 Weeks ago
Sphere Entertainment Co - Assistant Project Manager

Sphere Entertainment Co

United States (Remote)
1 Month ago
llnl - Robotics and Automation Academic Graduate Appointee

llnl

Livermore, California, United States (On-Site)
1 Day ago

Get notifed when new similar jobs are uploaded

Artificial Intelligence Jobs

Google - Software Engineer III, AI/ML GenAI, Google Cloud Platforms

Google

Sunnyvale, California, United States (On-Site)
2 Weeks ago
Canva - GenAI Research Engineering Manager - Image Generation (m/f/x) - Canva Austria

Canva

Vienna, Vienna, Austria (Remote)
5 Months ago
NVIDIA - Principal Technical Program Manager, AI and Enterprise Apps

NVIDIA

Santa Clara, California, United States (On-Site)
2 Weeks ago
Match Group - Sr. Software Engineer, Generative AI

Match Group

Palo Alto, California, United States (Hybrid)
6 Months ago
Meta - Postdoctoral Researcher, Embodied AI (PhD)

Meta

Seattle, Washington, United States (On-Site)
5 Months ago
NVIDIA - Senior Applied LLM Engineer, AI – Chip Design

NVIDIA

Canada (On-Site)
2 Months ago
Google - Technical Program Manager III, NPI System Software, Cloud AI Systems

Google

Sunnyvale, California, United States (On-Site)
2 Weeks ago
ByteDance - Student Researcher (Doubao (Seed) - Foundation Model - Generative AI) - 2025 Start (PhD)

ByteDance

San Jose, California, United States (On-Site)
6 Months ago
Tencent - UA Manager - AI Integrated

Tencent

Shenzhen, Guangdong Province, China (On-Site)
1 Month ago
ByteDance - Student Researcher Intern (Edge Research Project for General Intelligence)

ByteDance

San Jose, California, United States (On-Site)
2 Weeks ago

Get notifed when new similar jobs are uploaded

About The Company

Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (On-Site)

Massachusetts, United States (On-Site)

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (On-Site)

Texas, United States (On-Site)

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (Hybrid)

Santa Clara, California, United States (Hybrid)

View All Jobs

Get notified when new jobs are added by NVIDIA

Level Up Your Career in Game Development!

Transform Your Passion into Profession with Our Comprehensive Courses for Aspiring Game Developers.

Job Common Plug