Senior Solutions Architect, Infiniband and Networking Ethernet

16 Minutes ago • 8 Years + • Network Engineering

Job Summary

Job Description

NVIDIA seeks a Senior Networking (ETH/IB) Solutions Architect to design and implement large-scale networking projects for AI/HPC systems. Responsibilities include building infrastructure for new and existing customers, supporting operational reliability, improving service lifecycles, and monitoring system health. The role requires strong collaboration with customers and internal teams, utilizing expertise in Infiniband and Ethernet networks, automation tools, and troubleshooting skills. The ideal candidate possesses deep knowledge of networking protocols, experience with various network platforms, and a proven ability to deliver automated network provisioning solutions.
Must have:
  • 8+ years networking experience
  • InfiniBand & Ethernet expertise
  • EVPN, BGP, OSPF, VXLAN knowledge
  • Automation skills (Ansible, Salt, Python)
  • CI/CD pipeline development
  • Customer communication skills
Good to have:
  • Cloud network familiarity (AWS, GCP, Azure)
  • Linux/Networking certifications
  • HPC architecture understanding
  • Job scheduler (Slurm, PBS) knowledge
  • Lustre management experience
  • GPU hardware/software experience
  • Mandarin communication skills

Job Details

NVIDIA is looking for Senior Networking (ETH/IB) Solutions Architect to join its NVIDIA Infrastructure Specialist Team. Academic and commercial groups around the world are using NVIDIA products to revolutionize deep learning and data analytics, and to power data centers. Join the team building many of the largest and fastest AI/HPC systems in the world! We are looking for someone with the ability to work on a dynamic customer focused team that requires excellent interpersonal skills. This role will be interacting with customers, partners and internal teams, to analyze, define and implement large scale Networking projects. The scope of these efforts includes a combination of Networking, System Design and Automation and being the face to the customer!

What you'll be doing:

  • Primary responsibilities will include building AI/HPC infrastructure for new and existing customers.
  • Support operational and reliability aspects of large-scale AI clusters, focusing on performance at scale, real-time monitoring, logging, and alerting.
  • Engage in and improve the whole lifecycle of services—from inception and design through deployment, operation, and refinement.
  • Maintain services once they are live by measuring and monitoring availability, latency, and overall system health.
  • Provide feedback to internal teams such as opening bugs, documenting workarounds, and suggesting improvements.

What we need to see:

  • BS/MS/PhD or equivalent experience in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related fields.
  • At least 8 years of professional experience in networking fundamentals, TCP/IP stack, and data center architecture
  • Proficiency in configuring, testing, validating, and resolving issues in LAN and InfiniBand networks, especially in medium to large-scale HPC/AI environments.
  • Advanced knowledge of EVPN, BGP, OSPF, VXLAN protocols.
  • Hands-on experience with network switch/router platforms like Cumulus Linux, SONiC, IOS, JunosOS, and EOS.
  • Extensive experience delivering automated network provisioning solutions using tools like Ansible, Salt, and Python.
  • Ability to develop CI/CD pipelines for network operations.
  • Strong focus on customer needs and satisfaction.
  • Self-motivated with leadership skills to work collaboratively with customers and internal teams.
  • Ability to communicate technical concepts and collaborate effectively with Mandarin-speaking customers.
  • Strong written, verbal, and listening skills in English are essential.

Ways to stand out from the crowd:

  • Familiarity with cloud networks (AWS, GCP, Azure) is a plus.
  • Linux or Networking Certifications.
  • Experience with High-performance computing architectures. Understanding of how job schedulers(Slurm, PBS) work.
  • luster management technologies knowledge (bonus credit for BCM (Base Command Manager).)
  • Experience with GPU (Graphics Processing Unit) focused hardware/software.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking individuals in the world working for us. If you're creative and autonomous, we want to hear from you.

Similar Jobs

NVIDIA - Solution Architect - CSP Cloud

NVIDIA

Beijing, Beijing, China (On-Site)
2 Months ago
Two Point Studios - Senior 3D Artist

Two Point Studios

Farnham, England, United Kingdom (Hybrid)
2 Weeks ago
AGS - American Gaming Systems - Graduate Game Designer

AGS - American Gaming Systems

Australia (On-Site)
2 Weeks ago
Amanotes - Unity Developer (Game Magic Tiles 3 - Hybrid Music Game)

Amanotes

Ho Chi Minh City, Ho Chi Minh City, Vietnam (On-Site)
3 Weeks ago
Beyond Sports  - Unity Developer

Beyond Sports

Alkmaar, North Holland, Netherlands (On-Site)
1 Week ago
The Walt Disney Company - Network Engineer (1-year contract)

The Walt Disney Company

Hong Kong (On-Site)
5 Months ago
ByteDance - Network Implementation Engineer, Physical Network Infrastructure

ByteDance

Singapore (On-Site)
3 Weeks ago
ByteDance - Security System Engineer

ByteDance

San Jose, California, United States (On-Site)
3 Weeks ago
Hologate gmbh - IT Network Specialist

Hologate gmbh

Munich, Bavaria, Germany (On-Site)
5 Days ago

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

Larian Studios - Senior Generalist Technical Animator

Larian Studios

Barcelona, Catalonia, Spain (On-Site)
2 Weeks ago
NVIDIA - Machine Learning Engineer Intern - 2025

NVIDIA

Shanghai, Shanghai, China (On-Site)
2 Months ago
N-iX - Senior C++ Engineer (High Performance Computing)

N-iX

Argentina (Remote)
3 Weeks ago
Nordcurrent - Junior Gameplay Programmer

Nordcurrent

Vilnius, Vilnius County, Lithuania (On-Site)
5 Months ago
NVIDIA - Test Engineer, Electrical

NVIDIA

Roskilde, Denmark (Hybrid)
1 Week ago
Amber - 3D Animator (Project Based)

Amber

(Remote)
3 Weeks ago
Games For Love - Mobile Game Production Mentor

Games For Love

Washington, United States (Remote)
2 Weeks ago
Meta - Software Engineer, Machine Learning

Meta

Fremont, California, United States (Remote)
5 Months ago
Genies - Senior Engineer, Core Systems

Genies

San Mateo, California, United States (On-Site)
2 Weeks ago
Nintendo - Lighting Artist [Remote Contract] (Retro Studios)

Nintendo

United States (Remote)
8 Months ago

Get notifed when new similar jobs are uploaded

Jobs in undefined

Looks like we're out of matches

Set up an alert and we'll send you similar jobs the moment they appear!

Network Engineering Jobs

ByteDance - Site Reliability Engineer, Edge Services

ByteDance

Seattle, Washington, United States (On-Site)
2 Months ago
ByteDance - AI/LLM Network Software Engineer (High Speed Network)

ByteDance

Seattle, Washington, United States (On-Site)
3 Weeks ago
Extreme Network - Systems Engineer-Scandinavia

Extreme Network

Stockholm, Stockholm County, Sweden (Remote)
5 Months ago
Tencent - Senior Cloud Network Engineer - Singapore

Tencent

Singapore (On-Site)
6 Months ago
NVIDIA - Senior Networking Architect

NVIDIA

Canada (On-Site)
2 Months ago
NVIDIA - Data Center Infrastructure Specialist

NVIDIA

(On-Site)
2 Months ago
ByteDance - Software Engineer - Network Security - San Jose

ByteDance

San Jose, California, United States (On-Site)
5 Months ago
ION - Cloud Network Engineer

ION

Italy (Hybrid)
6 Months ago
Bohemia Interactive - Engine Network Programmer Prague/Brno

Bohemia Interactive

Prague, Prague, Czechia (On-Site)
5 Months ago

Get notifed when new similar jobs are uploaded

About The Company

Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.


Amsterdam, North Holland, Netherlands (Remote)

Beijing, Beijing, China (On-Site)

Shanghai, Shanghai, China (On-Site)

Taipei City, Taiwan (On-Site)

Santa Clara, California, United States (On-Site)

Canada (On-Site)

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (On-Site)

View All Jobs

Get notified when new jobs are added by NVIDIA

Level Up Your Career in Game Development!

Transform Your Passion into Profession with Our Comprehensive Courses for Aspiring Game Developers.

Job Common Plug