NVIDIA is looking for a talented Performance Research Engineer to join our Performance group.
The ideal candidate will profile and analyze AI workloads on large GPUs and CPUs scale clusters for distributed Deep Learning LLM training focusing at the collectives communication and networking.
You will work and interact with many types of HW and platforms such as HCAs, Switches, CPUs, GPUs, and Systems.
You will experience with and develop performance analysis tools and methodologies to dive deeply into the details, understand performance expectation, limitations, and bottlenecks.
What you'll be doing:
Experience and research AI workloads and DL models specifically tailored for large-scale deep learning LLM training on NVIDIA supercomputers with a focus on High-performance networking.
Benchmarking, Profiling, and Analyzing the performance to find bottlenecks and identify areas of improvement and optimizations, with a strong emphasis on networking aspects.
Implement performance analysis tools.
Collaborating with many teams from HW to SW to provide performance analysis insights.
Define performance test planning , set performance expectations for new technologies and solutions, and work to reach the performance targets limits.
What we need to see:
B.Sc in Computer Science or Software Engineering
5+ years of experience with high-performance Networking (RDMA, MPI)
Demonstrated Performance Analysis skills and methodologies.
Experience with NVIDIA GPUs, CUDA library, deep learning frameworks like TensorFlow or PyTorch,
combined with expertise in networking collective communication libraries (such as NCCL) and protocols (such as RoCE and RDMA).Fast and self-learning capabilities with strong analytical and problem-solving skills.
Programming Languages: Python, Bash and C languages
Experience with Linux OS distros.
Team player with good communication and interpersonal skills
Ways to stand out from the crowd:
In-depth knowledge and experience with AI workloads and benchmarking for distributed LLM training.
Knowledge in CUDA, and NCCL libraries.
Knowledge in Congestion Control algorithms.
In-depth System knowledge and understanding (Intel / AMD / ARM CPUs, NVIDIA GPUs, HCA, Memory, PCI).
Strong Performance Analysis skills and methodologies using modern tools.
NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) based on race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Other Jobs from NVIDIA
Senior Deep Learning Software Engineer - 3D Pose
Software Engineering Intern - OpenBMC
DFX Software Engineer (RDSS Intern)
Similar Jobs
Principal Machine Learning Engineer (URL Filtering Data Science)
LLM Application Intern, AV Infrastructure - 2025
Deep Learning Algorithm Engineering Intern - 2025
Software Engineering Intern, DL Algorithms - 2025
Product Management Intern, Enterprise Products - Summer 2025
There are more than 50,000 engineering jobs:
Subscribe to membership and unlock all jobs
Engineering Jobs
60,000+ jobs from 4,500+ well-funded companies
Updated Daily
New jobs are added every day as companies post them
Refined Search
Use filters like skill, location, etc to narrow results
Become a member
π₯³π₯³π₯³ 401 happy customers and counting...
Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.
To try it out
For active job seekers
For those who are passive looking
Cancel anytime
Frequently Asked Questions
- We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
- We've got about 70,000 jobs from 5,000 vetted companies. No fake or sleazy jobs here!
- We aggregate jobs from 5,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
- We're the only job board *for* software engineers, *by* software engineersβ¦ in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. π οΈ
- Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. π
- Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. π―
- Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. π
What Fellow Engineers Say