招聘
NVIDIA is leading company of AI computing. At NVIDIA, our employees are passionate about AI, HPC, VISUAL, GAMING. Our Solutions Architect team is more focusing to bring NVIDIA new technology into difference industries. We help to design the architecture of AI computing platform, analysis the AI and HPC applications to deliver our value to customers. This role will be instrumental in leveraging NVIDIA's cutting-edge technologies to optimize open-source and proprietary large models, create AI workflows, and support our customers in implementing advanced AI solutions.
What you’ll be doing:
-
Drive the implementation and deployment of NVIDIA Inference Microservice (NIM) solutions
-
Use NVIDIA NIM Tools or Factory Pipeline to package optimized models (including LLM, VLM, Diffusion, Retriever, CV, OCR, AI4Science etc.) into containers providing standardized API access for on-prem or cloud deployment
-
Refine NIM tools for the community, help the community to build their performant NIMs
-
Design and implement agentic AI tailored to customer business scenarios using NIMs
-
Deliver technical projects, demos and client support tasks as directed by the Solution Architecture Leadership
-
Provide technical support and guidance to customers, facilitating the adoption and implementation of NVIDIA technologies and products
-
Collaborate with cross-functional teams to enhance and expand our AI solutions portfolio
-
Be an internal champion for NVIDIA software and total solutions in technical community
-
Be an industry thought leader on integrating NVIDIA technology especially inference services into LHA, business partners and whole community
-
Assist in supporting NVAIE team and driving NVAIE business in China
What we need to see:
-
3+ years working experience with Bachelor's or Master's degree in Computer Science, Artificial Intelligence, or a related field
-
Proven experience in deploying and optimizing large language models
-
Familiarity with main stream inference engines or inference framework (e.g., SGLang, vLLM, TensorRT, or ONNX Runtime, Py Torch)
-
Strong programming skills in Python or C++
-
Experience with DevOps/MLOps such as Docker, Git, and CI/CD practices
-
Excellent problem-solving skills and ability to troubleshoot complex technical issues
-
Demonstrated ability to collaborate effectively across diverse, global teams, adapting communication styles while maintaining clear, constructive professional interactions
Ways to stand out from the crowd:
-
Experience in AI performance engineering
-
Expertise in model optimization techniques
-
Knowledge of AI workflow design and implementation; experience on cluster resource management tools
-
CUDA optimization experience, extensive experience designing and deploying large scale HPC and enterprise computing systems
NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and creative people in the world working here! If you are driven by impact, passionate about enterprise AI, and inspired to shape the agentic AI at scale for NVIDIA’s Supply Chain Operations, we want to hear from you. Apply today!
Total Views
0
Apply Clicks
0
Mock Applicants
0
Scraps
0
Similar Jobs

Principal Cloud Solution Architect - Security
Microsoft · United States, Multiple Locations, Multiple Locations

Principal Solution Architect - Auckland
Microsoft · New Zealand, Auckland, Auckland

Field Solutions Architect, Applied Artificial Intelligence, Google Cloud
Google ·

Data Cloud Solution Architect
Microsoft · Germany, Multiple Locations, Multiple Locations

Field Solutions Architect, GenAI, Google Cloud
Google · placeTaipei, Taiwan
About NVIDIA

NVIDIA
PublicA computing platform company operating at the intersection of graphics, HPC, and AI.
10,001+
Employees
Santa Clara
Headquarters
$4.57T
Valuation
Reviews
4.1
10 reviews
Work Life Balance
3.5
Compensation
4.2
Culture
4.3
Career
4.5
Management
4.0
75%
Recommend to a Friend
Pros
Great culture and supportive environment
Smart colleagues and excellent people
Cutting-edge technology and learning opportunities
Cons
Team-dependent experience and outcomes
Work-life balance issues with long hours
Politics and influence over competence
Salary Ranges
47 data points
Junior/L3
Mid/L4
Junior/L3 · Analyst
7 reports
$170,275
total / year
Base
$130,981
Stock
-
Bonus
-
$155,480
$234,166
Interview Experience
7 interviews
Difficulty
3.1
/ 5
Experience
Positive 0%
Neutral 86%
Negative 14%
Interview Process
1
Application Review
2
Recruiter Screen
3
Online Assessment
4
Technical Interview
5
System Design Interview
6
Team Review
Common Questions
Coding/Algorithm
System Design
Technical Knowledge
Behavioral/STAR
News & Buzz
NVIDIA Culture Discussions
Team-dependent experience; sink-or-swim culture that rewards high performers but can be overwhelming. No politics, flat structure, but demanding workload with some teams requiring evening/weekend work.
News
·
NaNw ago
Negotiating NVIDIA's Offer
Base, stock, and sign-on negotiable. Recruiters invested in closing candidates. CEO reviews all 42K employee salaries monthly. Stock growth has made many employees millionaires.
News
·
NaNw ago
NVIDIA Company Reviews
WLB rated 3.9/5 (lowest category). 64% satisfied with WLB but 53% feel burnt out. Compensation rated 4.4-4.5/5. Experience highly team-dependent.
News
·
NaNw ago
NVIDIA Interview Discussions
Technical bar is high with 4-6 rounds. Process takes 4-8 weeks. Expect C++ questions, LeetCode medium, and system design. Difficulty rated 3.16/5.
News
·
NaNw ago