Senior Solutions Architect, Infiniband and Networking Ethernet

3 Months ago • 8 Years + • Network Engineering

Job Summary

Job Description

NVIDIA seeks a Senior Networking (ETH/IB) Solutions Architect to design and implement large-scale networking projects for AI/HPC systems. Responsibilities include building infrastructure for new and existing customers, supporting operational reliability, improving service lifecycles, and monitoring system health. The role requires strong collaboration with customers and internal teams, utilizing expertise in Infiniband and Ethernet networks, automation tools, and troubleshooting skills. The ideal candidate possesses deep knowledge of networking protocols, experience with various network platforms, and a proven ability to deliver automated network provisioning solutions.
Must have:
  • 8+ years networking experience
  • InfiniBand & Ethernet expertise
  • EVPN, BGP, OSPF, VXLAN knowledge
  • Automation skills (Ansible, Salt, Python)
  • CI/CD pipeline development
  • Customer communication skills
Good to have:
  • Cloud network familiarity (AWS, GCP, Azure)
  • Linux/Networking certifications
  • HPC architecture understanding
  • Job scheduler (Slurm, PBS) knowledge
  • Lustre management experience
  • GPU hardware/software experience
  • Mandarin communication skills

Job Details

NVIDIA is looking for Senior Networking (ETH/IB) Solutions Architect to join its NVIDIA Infrastructure Specialist Team. Academic and commercial groups around the world are using NVIDIA products to revolutionize deep learning and data analytics, and to power data centers. Join the team building many of the largest and fastest AI/HPC systems in the world! We are looking for someone with the ability to work on a dynamic customer focused team that requires excellent interpersonal skills. This role will be interacting with customers, partners and internal teams, to analyze, define and implement large scale Networking projects. The scope of these efforts includes a combination of Networking, System Design and Automation and being the face to the customer!

What you'll be doing:

  • Primary responsibilities will include building AI/HPC infrastructure for new and existing customers.
  • Support operational and reliability aspects of large-scale AI clusters, focusing on performance at scale, real-time monitoring, logging, and alerting.
  • Engage in and improve the whole lifecycle of services—from inception and design through deployment, operation, and refinement.
  • Maintain services once they are live by measuring and monitoring availability, latency, and overall system health.
  • Provide feedback to internal teams such as opening bugs, documenting workarounds, and suggesting improvements.

What we need to see:

  • BS/MS/PhD or equivalent experience in Computer Science, Electrical/Computer Engineering, Physics, Mathematics, or related fields.
  • At least 8 years of professional experience in networking fundamentals, TCP/IP stack, and data center architecture
  • Proficiency in configuring, testing, validating, and resolving issues in LAN and InfiniBand networks, especially in medium to large-scale HPC/AI environments.
  • Advanced knowledge of EVPN, BGP, OSPF, VXLAN protocols.
  • Hands-on experience with network switch/router platforms like Cumulus Linux, SONiC, IOS, JunosOS, and EOS.
  • Extensive experience delivering automated network provisioning solutions using tools like Ansible, Salt, and Python.
  • Ability to develop CI/CD pipelines for network operations.
  • Strong focus on customer needs and satisfaction.
  • Self-motivated with leadership skills to work collaboratively with customers and internal teams.
  • Ability to communicate technical concepts and collaborate effectively with Mandarin-speaking customers.
  • Strong written, verbal, and listening skills in English are essential.

Ways to stand out from the crowd:

  • Familiarity with cloud networks (AWS, GCP, Azure) is a plus.
  • Linux or Networking Certifications.
  • Experience with High-performance computing architectures. Understanding of how job schedulers(Slurm, PBS) work.
  • luster management technologies knowledge (bonus credit for BCM (Base Command Manager).)
  • Experience with GPU (Graphics Processing Unit) focused hardware/software.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking individuals in the world working for us. If you're creative and autonomous, we want to hear from you.

Similar Jobs

quience - Production Manager - Home (North)

quience

Delhi, India (On-Site)
3 Weeks ago
Progress - Partner Account Manager, Senior

Progress

Japan (Remote)
3 Months ago
WebFX - Jr. Marketing Analytics Consultant

WebFX

Harrisburg, Pennsylvania, United States (On-Site)
8 Months ago
attentive - Data Scientist II

attentive

San Francisco, California, United States (Hybrid)
3 Months ago
CD PROJEKT RED - Lead Technical Animator

CD PROJEKT RED

Boston, Massachusetts, United States (On-Site)
2 Months ago
NCR Voyix - Network Engineer

NCR Voyix

Tokyo, Japan (On-Site)
2 Months ago
Nice - Lead Cloud Network Engineer

Nice

Atlanta, Georgia, United States (On-Site)
3 Weeks ago
Zones - Network Operations Engineer

Zones

Bengaluru, Karnataka, India (On-Site)
2 Months ago
Alten Technology - Network Engineer

Alten Technology

St. Cloud, Minnesota, United States (On-Site)
3 Weeks ago
Square - Network Engineer

Square

Lyon, Auvergne-Rhône-Alpes, France (On-Site)
2 Weeks ago

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

Lambda - Head of Revenue Accounting

Lambda

San Jose, California, United States (Hybrid)
3 Weeks ago
Pattern - Principal Product Manager - Operations

Pattern

Lehi, Utah, United States (On-Site)
2 Months ago
Playtouch.net - Game Developer Junior

Playtouch.net

Grand Baie, Rivière Du Rempart District, Mauritius (On-Site)
1 Year ago
Safe security - Software Development Engineer II - Frontend

Safe security

Bengaluru, Karnataka, India (On-Site)
2 Weeks ago
WebFX - Jr. SaaS Project Manager

WebFX

Ann Arbor, Michigan, United States (On-Site)
8 Months ago
entrata - Product Owner

entrata

Pune, Maharashtra, India (Hybrid)
9 Months ago
London stock Exchange - Specialist, Data Office

London stock Exchange

Bengaluru, Karnataka, India (On-Site)
2 Weeks ago
EveryMatrix - Mid/Senior Data Analyst

EveryMatrix

Yerevan, Yerevan, Armenia (On-Site)
3 Months ago
Huuuge Games - Product Management Lead

Huuuge Games

Warsaw, Masovian Voivodeship, Poland (Hybrid)
1 Week ago
NetEase Games - Global Influencer Marketing

NetEase Games

Canada (On-Site)
8 Months ago

Get notifed when new similar jobs are uploaded

Jobs in undefined

Looks like we're out of matches

Set up an alert and we'll send you similar jobs the moment they appear!

Network Engineering Jobs

Rockstar Games - Senior Network Programmer

Rockstar Games

Leeds, England, United Kingdom (On-Site)
1 Month ago
Tide - Workplace Enablement Engineer - 3 (IT Network)

Tide

Delhi, India (On-Site)
3 Weeks ago
bytedance - Network Implementation Engineer - Physical Network Infrastructure

bytedance

Bangkok, Bangkok, Thailand (On-Site)
4 Months ago
KOJIMA PRODUCTIONS - Network Programmer

KOJIMA PRODUCTIONS

Tokyo, Japan (On-Site)
8 Months ago
NinjaVan - Network Sales Specialist

NinjaVan

Makati City, Metro Manila, Philippines (Hybrid)
9 Months ago
Capgemini - Connectivity & Network Engineer

Capgemini

Hyderabad, Telangana, India (On-Site)
1 Month ago
Marvell - Principal Field Application Engineer-Networking

Marvell

Santa Clara, California, United States (On-Site)
1 Year ago
Virtuos - Network Programmer

Virtuos

Czechia (Hybrid)
4 Months ago
NVIDIA - Senior System Software Architect, HPC Networking

NVIDIA

Tel Aviv-Yafo, Tel Aviv District, Israel (On-Site)
5 Months ago
luxsoft - MS SQL Server Developer with Tableau

luxsoft

Chennai, Tamil Nadu, India (On-Site)
2 Months ago

Get notifed when new similar jobs are uploaded

About The Company

Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.

Taipei City, Taiwan (On-Site)

Beijing, Beijing, China (On-Site)

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (Hybrid)

Bengaluru, Karnataka, India (Hybrid)

Yokne'am Illit, North District, Israel (On-Site)

Yokne'am Illit, North District, Israel (On-Site)

Yokne'am Illit, North District, Israel (On-Site)

Dubai, Dubai, United Arab Emirates (On-Site)

Beijing, Beijing, China (On-Site)

View All Jobs

Get notified when new jobs are added by NVIDIA

Level Up Your Career in Game Development!

Transform Your Passion into Profession with Our Comprehensive Courses for Aspiring Game Developers.

Job Common Plug