热门公司

NVIDIA
NVIDIA

Pioneering accelerated computing and AI

Senior Architect, GPU Profiling System

职能工程
级别资深
地点United States, India
方式现场办公
类型全职
发布1个月前
立即申请

必备技能

Python

NVIDIA’s GPU Architecture Group is looking for architects to contribute to the design of our proprietary profiler subsystem, the apparatus embedded in every GPU that enables our profiling and monitoring tools to capture data and provide feedback for performance optimization. As a member of our team, you will need to combine skills in hardware modeling and verification with a deep understanding of GPU architecture, operating systems, and application performance analysis to innovate new methods of hardware profiling that yield more meaningful and accessible performance insights. You will have a tangible impact at a fast-paced company that is spearheading the AI revolution. Join our technically diverse team of GPU architects, software engineers and deep learning experts to push the boundaries of AI performance!

What you’ll be doing:

  • Architect and plan features in concert with software, hardware, and verification teams working across the globe to implement next generation GPU profiling features.

  • Build functional and performance models to refine and verify hardware designs.

  • Create test plans to validate the features you design and contribute to their implementation.

  • Constantly develop your skills for practical innovation by improving your understanding of the AI workloads, the GPU architecture, and the profiling software stack.

What we need to see:

  • Masters, or PhD in relevant field (Eg: Computer Science, Computer Engineering or Electrical Engineering) or equivalent experience.

  • 3+ years of relevant computer architecture, ASIC design/verification, or software development experience.

  • Strong programming skills in C++ (or similar) and Python (or similar).

  • Solid foundation in computer architecture and hardware performance analysis.

  • Experience with performance modeling and hardware simulation, ideally using SystemC.

  • Strong communication and interpersonal skills including the ability to work with a distributed interdisciplinary team.

Ways to stand out from the crowd:

  • Expertise in developing and optimizing parallel algorithms, particularly using GPUs.

  • Extensive experience as a user or developer of CPU or GPU profiling tools.

  • Background with AI and/or high-performance computing applications

  • Experience contributing to and debugging large codebases with many developers.

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until March 13, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

浏览量

0

申请点击

0

Mock Apply

0

收藏

0

关于NVIDIA

NVIDIA

NVIDIA

Public

A computing platform company operating at the intersection of graphics, HPC, and AI.

10,001+

员工数

Santa Clara

总部位置

$4.57T

企业估值

评价

10条评价

4.4

10条评价

工作生活平衡

2.8

薪酬

4.5

企业文化

4.2

职业发展

4.3

管理层

3.8

78%

推荐率

优点

Cutting-edge technology and innovation

Excellent compensation and benefits

Great team culture and collaboration

缺点

High pressure and expectations

Poor work-life balance and long hours

Fast-paced environment leading to burnout

薪资范围

79个数据点

Junior/L3

Mid/L4

Senior/L5

Junior/L3 · Analyst

7份报告

$170,275

年薪总额

基本工资

$130,981

股票

-

奖金

-

$155,480

$234,166

面试评价

5条评价

难度

3.0

/ 5

面试流程

1

Application Review

2

Recruiter Screen

3

Technical Phone Screen

4

Onsite/Virtual Interviews

5

Team Matching

6

Offer

常见问题

Coding/Algorithm

System Design

Behavioral/STAR

Technical Knowledge

Past Experience