
Accelerated Compute Systems Performance Architect Intern - 2027
Shanghai, Shanghai, ChinaFull-timePosted 4 days agoStill listed today
Most applications go out cold — see where you stand first. No sign-up to start.
Watch jobs like this. New roles like this one near Shanghai, Shanghai, by email.
Don't just apply. Show up ready.
Olive works from this exact posting.
At a glance
Olive lists jobs from US employers, including remote roles you can work from the United States.
Requirements
Credentials this posting asks for.
Job overview
NVIDIA seeks an Accelerated Computing Architect intern to contribute to high‑performance computing, AI, and automotive workloads by analyzing and optimizing GPU performance, collaborating across hardware and software teams, and communicating results through papers and blogs.
Skills & qualifications
Skills
Qualifications
Full job description
We are now looking for an Accelerated Computing Architect intern. NVIDIA is developing software and hardware system architectures for accelerated high performance computing, scientific computing, machine learning, artificial intelligence, datacenter, and automotive computing. This position offers you the opportunity to make a meaningful impact in a fast-moving, technology focused company.
What you'll be doing:
-
Performing in-depth analysis and optimization to ensure the best possible performance on current and/or next-generation NVIDIA GPUs.
-
Understanding and analyzing the interplay of hardware and software architectures on core algorithms, programming models, and applications.
-
Actively collaborating with the hardware design, software engineering, product, and research teams to guide the direction of accelerated computing.
-
Diving into accelerated computing applications to facilitate software-hardware co-design.
-
Writing up and presenting your work by writing white papers, conference publications, official blog posts, patent applications, etc. as appropriate.
What we need to see:
-
Pursuing B.Sc., M.Sc., or Ph.D. in relevant discipline (CS, EE, CE).
-
A passion for performance analysis and optimization.
-
Hands-on experience with the massively parallel GPU programming model, e.g. CUDA or OpenCL. Familiarity with APIs for multi-node communication, like MPI or OpenSHMEM/NVSHMEM, is a plus.
-
Solid background in GPU and computer systems architecture.
-
Strong knowledge of C and C++ with a solid understanding of software design, programming techniques, and algorithms. Familiarity with Python is a plus.
-
Good communication and organization skills, with a logical approach to problem solving, good time management, and task prioritization skills.
NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most brilliant and talented people in the world working for us. Are you creative and autonomous? Do you love the challenge of pushing an architecture to its limits? If so, we want to hear from you.
NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Similar jobs, posted recently
Open roles like this one, listed in the last 30 days.
2027 Shanghai Performance Researcher Summer InternshipOptiver · Shanghai, Shanghai, ChinaPosted 1 day agoPosted 1 day ago
Deep Learning Performance ArchitectNVIDIA · Shanghai, Shanghai, ChinaPosted 3w agoPosted 3w ago
Compute System Arch AI Infra Intern - 2027NVIDIA · Shanghai, Shanghai, ChinaPosted 4 days agoPosted 4 days ago
Platform Power Thermal Performance EngineerIntel · Shanghai, Shanghai, ChinaPosted 4w agoPosted 4w ago
TPC Arch Intern, GPU SM - 2027NVIDIA · Shanghai, Shanghai, ChinaPosted 6 days agoPosted 6 days ago
You've read the whole posting — now see how you match it.