招聘
Benefits & Perks
•Equity
•Healthcare
•Equity
•Healthcare
Required Skills
Python
PyTorch
Deep Learning
Inference Optimization
We are looking for a Senior Deep Learning Engineer to help bring Cosmos World Foundation Models from research into efficient, production-grade systems. You’ll focus on optimizing and deploying models for high-performance inference on diverse GPU platforms. This role sits at the intersection of deep learning, systems, and GPU optimization - working closely with research scientists, software engineers, and hardware experts.
NVIDIA Cosmos is a platform purpose-built for physical AI, featuring powerful generative models. Developers use Cosmos to accelerate physical AI development for autonomous vehicles (AVs), robots, and video analytics AI agents by simulating and reasoning about the physical world.
What you'll be doing:
-
Improve inference speed for Cosmos WFMs on GPU platforms.
-
Effectively carry out the production deployment of Cosmos WFMs.
-
Profile and analyze deep learning workloads to identify and remove bottlenecks.
What we need to see:
-
5 years of experience.
-
MSc or PhD in CS, EE, or CSEE or equivalent experience.
-
Strong background in Deep Learning.
-
Strong programming skills in Python and Py Torch.
-
Experience with inference optimization techniques (such as quantization) and inference optimization frameworks, one of: TensorRT, TensorRT-LLM, vLLM, SGLang.
Ways to stand out from the crowd:
-
Familiarity with deploying Deep Learning models in production settings (e.g., Docker, Triton Inference Server).
-
CUDA programming experience.
-
Familiarity with diffusion models.
-
Proven experience in analyzing, modeling, and tuning the performance of GPU workloads, both inference and training.
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. For Poland: The base salary range is 221,250 PLN - 383,500 PLN for Level 3, and 292,500 PLN - 507,000 PLN for Level 4.
Total Views
0
Apply Clicks
0
Mock Applicants
0
Scraps
0
Similar Jobs

Lead Advanced Analytics, Digital & AI Products
Airbnb · Bangalore, India

Machine Learning Engineer III
Chewy · Boston, Massachusetts, United States of America

Senior Machine Learning Engineer, AI Research and Development
Robinhood · Bellevue, WA; Menlo Park, CA

Senior Machine Learning Engineer, Agentic
Robinhood · Bellevue, WA; Menlo Park, CA

Senior Staff Machine Learning Engineer, (ML Underwriting)
Affirm · Remote Canada
About NVIDIA

NVIDIA
PublicA computing platform company operating at the intersection of graphics, HPC, and AI.
10,001+
Employees
Santa Clara
Headquarters
$4.57T
Valuation
Reviews
4.1
10 reviews
Work Life Balance
3.5
Compensation
4.2
Culture
4.3
Career
4.5
Management
4.0
75%
Recommend to a Friend
Pros
Great culture and supportive environment
Smart colleagues and excellent people
Cutting-edge technology and learning opportunities
Cons
Team-dependent experience and outcomes
Work-life balance issues with long hours
Politics and influence over competence
Salary Ranges
47 data points
L3
L4
L5
L3 · Data Scientist IC2
0 reports
$177,542
total / year
Base
-
Stock
-
Bonus
-
$150,910
$204,174
Interview Experience
7 interviews
Difficulty
3.1
/ 5
Experience
Positive 0%
Neutral 86%
Negative 14%
Interview Process
1
Application Review
2
Recruiter Screen
3
Online Assessment
4
Technical Interview
5
System Design Interview
6
Team Review
Common Questions
Coding/Algorithm
System Design
Technical Knowledge
Behavioral/STAR
News & Buzz
Negotiating NVIDIA's Offer
Base, stock, and sign-on negotiable. Recruiters invested in closing candidates. CEO reviews all 42K employee salaries monthly. Stock growth has made many employees millionaires.
News
·
NaNw ago
NVIDIA Company Reviews
WLB rated 3.9/5 (lowest category). 64% satisfied with WLB but 53% feel burnt out. Compensation rated 4.4-4.5/5. Experience highly team-dependent.
News
·
NaNw ago
NVIDIA Culture Discussions
Team-dependent experience; sink-or-swim culture that rewards high performers but can be overwhelming. No politics, flat structure, but demanding workload with some teams requiring evening/weekend work.
News
·
NaNw ago
NVIDIA Interview Discussions
Technical bar is high with 4-6 rounds. Process takes 4-8 weeks. Expect C++ questions, LeetCode medium, and system design. Difficulty rated 3.16/5.
News
·
NaNw ago