Senior Solutions Architect - Generative AI

3 Months ago • 7 Years + • Artificial Intelligence

Job Summary

Job Description

NVIDIA seeks a Senior Solutions Architect specializing in Generative AI, particularly Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG). Responsibilities include architecting end-to-end generative AI solutions, collaborating with customers to understand their needs and design tailored solutions, supporting pre-sales activities, working with engineering teams to improve software, and leading workshops. The role requires expertise in LLM training, deployment, optimization, and RAG workflow implementation across various hardware platforms, including GPUs and cloud environments. Strong communication and leadership skills are essential.
Must have:
  • 7+ years experience in Generative AI
  • LLM training and deployment expertise
  • RAG workflow implementation
  • TensorFlow/PyTorch/Hugging Face proficiency
  • GPU cluster architecture knowledge
  • Excellent communication skills
Good to have:
  • Cloud deployment experience (AWS, Azure, GCP)
  • LLM model optimization for inference
  • Docker and Kubernetes experience
  • Experience with NVIDIA GPU technologies

Job Details

NVIDIA is seeking a dynamic and experienced Generative AI Solution Architect with specialized expertise in training Large Language Models (LLMs) and implementing workflows based on Pretraining, Finetuning LLMs & Retrieval-Augmented Generation (RAG). As a key member of our AI Solutions team, you will play a pivotal role in architecting and delivering cutting-edge solutions that leverage the power of NVIDIA's generative AI technologies. This position requires a deep understanding of language models, particularly open source LLMs, and a strong proficiency in designing and implementing RAG-based workflows.

What You Will Be Doing:

  • Architect end-to-end generative AI solutions with a focus on LLMs training , deployment and RAG workflows.

  • Collaborate closely with customers to understand their language-related business challenges and design tailored solutions.

  • Collaborate with sales and business development teams to support pre-sales activities, including technical presentations and demonstrations of LLM and RAG capabilities.

  • Work closely with NVIDIA engineering teams to provide feedback and contribute to the evolution of generative AI software.

  • Engage directly with customers/partners to understand their requirements and challenges.

  • Lead workshops and design sessions to define and refine generative AI solutions focused on LLMs and RAG workflows and lead the training and optimization of Large Language Models using NVIDIA’s hardware and software platforms.

  • Implement strategies for efficient and effective training of LLMs to achieve optimal performance.

  • Design and implement RAG-based workflows to enhance content generation and information retrieval.

  • Work closely with customers to integrate RAG workflows into their applications and systems and stay abreast of the latest developments in language models and generative AI technologies.

  • Provide technical leadership and guidance on best practices for training LLMs and implementing RAG-based solutions.

What We Need To See:

  • Master's or Ph.D. in Computer Science, Artificial Intelligence, or equivalent experience

  • 7+ years of hands-on experience in a technical AI role, specifically focusing on generative AI, with a strong emphasis on training Large Language Models (LLMs).

  • Proven track record of successfully deploying and optimizing LLM models for inference in production environments.

  • In-depth understanding of state-of-the-art language models, including but not limited to GPT-3, BERT, or similar architectures.

  • Expertise in training and fine-tuning LLMs using popular frameworks such as TensorFlow, PyTorch, or Hugging Face Transformers.

  • Proficiency in model deployment and optimization techniques for efficient inference on various hardware platforms, with a focus on GPUs.

  • Strong knowledge of GPU cluster architecture and the ability to leverage parallel processing for accelerated model training and inference.

  • Excellent communication and collaboration skills with the ability to articulate complex technical concepts to both technical and non-technical stakeholders.

  • Experience leading workshops, training sessions, and presenting technical solutions to diverse audiences.

Ways To Stand Out From The Crowd:

  • Experience in deploying LLM models in cloud environments (e.g., AWS, Azure, GCP) and on-premises infrastructure.

  • Proven ability to optimize LLM models for inference speed, memory efficiency, and resource utilization.

  • Familiarity with containerization technologies (e.g., Docker) and orchestration tools (e.g., Kubernetes) for scalable and efficient model deployment.

  • Deep understanding of GPU cluster architecture, parallel computing, and distributed computing concepts.

  • Hands-on experience with NVIDIA GPU technologies, and GPU cluster management and ability to design and implement scalable and efficient workflows for LLM training and inference on GPU clusters

With competitive salaries and a generous benefits package, we are widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us and, due to unprecedented growth, our exclusive engineering teams are rapidly growing. If you're a creative and autonomous engineer with a real passion for technology, we want to hear from you!

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Similar Jobs

ByteDance - Research Engineer Intern (Doubao (Seed) - Machine Learning System) - 2025 Summer (MS)

ByteDance

San Jose, California, United States (On-Site)
6 Months ago
Alphasense - Join AlphaSense India Talent Community

Alphasense

Pune, Maharashtra, India (On-Site)
6 Hours ago
Orion Innovation - Data Engineer-AI,ML

Orion Innovation

Chennai, Tamil Nadu, India (On-Site)
6 Months ago
Flip Fit - Senior Machine Learning Engineer

Flip Fit

(Remote)
1 Month ago
CharacterAI - Software Engineer, Machine Learning Infrastructure

CharacterAI

New York, New York, United States (On-Site)
1 Month ago
Google - Senior Software Engineer, Visual Language and Multimodal Modeling

Google

Sydney, New South Wales, Australia (On-Site)
2 Weeks ago
Google - Customer Engineer IV, AI/ML, HCLS, Google Cloud

Google

Seattle, Washington, United States (On-Site)
1 Week ago
Google - Software Engineer III, AI/ML GenAI, Google Cloud Data Management

Google

Sunnyvale, California, United States (On-Site)
2 Weeks ago
Google - Senior AI Sales Specialist, Google Cloud

Google

Tokyo, Japan (On-Site)
2 Weeks ago

Get notifed when new similar jobs are uploaded

Similar Skill Jobs

Google - Manager, gTech Ads Customer Support, Tech CoE

Google

Gurugram, Haryana, India (On-Site)
2 Days ago
Google - Applied ML Engineer for AICore

Google

Taipei City, Taiwan (On-Site)
2 Weeks ago
Armada - Senior Data Engineer

Armada

Thiruvananthapuram, Kerala, India (On-Site)
7 Months ago
HP - Machine Learning Engineer

HP

Palo Alto, California, United States (On-Site)
7 Months ago
ByteDance - Software Engineer Large Model System Graduate (Machine Learning Sys-US) - 2024 Start (BS/MS)

ByteDance

Seattle, Washington, United States (On-Site)
6 Months ago
Google - Customer Engineer, AI/ML, HCLS, Google Cloud

Google

Chicago, Illinois, United States (On-Site)
2 Weeks ago
The Walt Disney Company - Lead Applied AI Engineer

The Walt Disney Company

Santa Monica, California, United States (On-Site)
1 Month ago
NVIDIA - Principal Engineer

NVIDIA

United States (Remote)
2 Months ago

Get notifed when new similar jobs are uploaded

Jobs in Bengaluru, Karnataka, India

Sony Music Career - Manager – Royalty Audit

Sony Music Career

Mumbai, Maharashtra, India (Hybrid)
21 Hours ago
Contentstack - Senior Software Engineer II (React JS)

Contentstack

Virar, Maharashtra, India (Hybrid)
18 Hours ago
commerce iq - Director of Software Implementation

commerce iq

Bengaluru, Karnataka, India (On-Site)
16 Hours ago
Ethernovia - Senior Embedded Software Engineer

Ethernovia

Pune, Maharashtra, India (On-Site)
6 Hours ago
Dialpad AI - Senior Software Engineer, Analytics

Dialpad AI

Bengaluru, Karnataka, India (Hybrid)
19 Hours ago
Comscore - Senior Data Analyst

Comscore

Pune, Maharashtra, India (On-Site)
21 Hours ago
PwC - CRM Technical -Senior associate

PwC

Mumbai, Maharashtra, India (On-Site)
7 Months ago
Google - Software Engineering, Full Stack

Google

Hyderabad, Telangana, India (On-Site)
2 Weeks ago
CleverTap - Senior Customer Success Engineer

CleverTap

Mumbai, Maharashtra, India (On-Site)
6 Months ago
Google - International Growth Consultant

Google

Gurugram, Haryana, India (On-Site)
2 Weeks ago

Get notifed when new similar jobs are uploaded

Artificial Intelligence Jobs

Google - Technical Program Manager, Cloud ML Compute Services

Google

Sunnyvale, California, United States (On-Site)
2 Days ago
Amazon Games - Senior ML Scientist, Amazon Games AI Research

Amazon Games

San Diego, California, United States (On-Site)
4 Months ago
ByteDance - Student Researcher (Doubao (Seed) - Foundation Model - Vision Generative AI)

ByteDance

San Jose, California, United States (On-Site)
1 Month ago
Zoox - Senior Software Engineer - High Performance Computing

Zoox

Foster City, California, United States (Hybrid)
6 Months ago
Google - EDA/CAD Custom Tool Development Engineer

Google

Bengaluru, Karnataka, India (On-Site)
1 Week ago
Meta - AI Research Scientist, Language - Generative AI

Meta

Menlo Park, California, United States (On-Site)
5 Months ago
Google - Silicon AI/ML Architect, TPU

Google

Bengaluru, Karnataka, India (On-Site)
2 Weeks ago
Google - EDA/CAD Custom Tool Development Engineer

Google

Bengaluru, Karnataka, India (On-Site)
2 Days ago
Zoox - Senior Software Engineer - Simulaton Scenario Automation

Zoox

Foster City, California, United States (Hybrid)
6 Months ago
Google - Customer Engineer IV, Field CTO

Google

Austin, Texas, United States (On-Site)
2 Days ago

Get notifed when new similar jobs are uploaded

About The Company

Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (On-Site)

Massachusetts, United States (On-Site)

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (On-Site)

Texas, United States (On-Site)

Santa Clara, California, United States (On-Site)

Santa Clara, California, United States (Hybrid)

Santa Clara, California, United States (Hybrid)

View All Jobs

Get notified when new jobs are added by NVIDIA

Level Up Your Career in Game Development!

Transform Your Passion into Profession with Our Comprehensive Courses for Aspiring Game Developers.

Job Common Plug