AMD logo

Software Development Engineer, Collectives and Network Optimization

AMD

Hybrid
San Jose, CA
Full-time
Senior
Staff
$179k–$255kPosted 2mo ago

Real job — pulled straight from AMD’s careers page · Verified July 12, 2026 · No reposts.

Job description

AMD is hiring a Software Development Engineer, Collectives and Network Optimization — a full-time, based in San Jose, CA role ($179k–$255k). Apply directly on AMD's careers page below.



WHAT YOU DO AT AMD CHANGES EVERYTHING 

At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond. Together, we advance your career.  




THE ROLE: 

Senior level engineer who will be responsible for driving AMD’s strategy, architecture, optimization and tooling to achieve industry-leading AI Pre-training and Distributed Inference Performance on AMD GPU. You will partner across hardware architecture, AI frameworks, compilers, runtime, ROCm, developer tools and model to scale performance analysis and optimization. 

 

As an engineer of Collectives and Network performance, you will help drive the end-to-end technical performance attainment across the entire software stack focusing on getting the best performance on multiple generations of AMD GPUs with wide range of models including latest state-of-the-art AI models. You will help set the strategy and roadmap for general optimization, accelerating supporting new models and out of box performance. 

 

If you are passionate about performance optimization, getting the best out of the hardware, and shaping the future of AI acceleration, then this role is for you. 

 

 

THE PERSON: 

The ideal candidate will have deep knowledge with Network, NIC and GPU hardware architecture, software optimization, performance modeling, AI frameworks and latest trend in inference and training optimization. Hand-on experience in mapping model architecture to low level software, hardware and understanding the impact of each layer of the stack on model performance. Strong knowledge in latest generative model architecture, especially SoTA models, distributed inference and deployment at scale is crucial. 

 

KEY RESPONSIBILITIES: 

  • Help set strategy and roadmap for AMD Collectives and Network optimizations. 
  • Provide guidelines to customers on efficient network load-balancing, workload scheduling and model sharding strategies. 
  • Performance tuning, profiling and analysis of large-scale models for LLM, diffusion, multimodal, RecSys and generative AI, single node and distributed. In addition to exploring various tradeoffs and design decisions. 
  • Participate in hardware-software co-design for future hardware optimizations – especially on scale-up networks, NIC and scale-out networks. 
  • Develop and improve framework, tools and infrastructure for performance estimation, modeling and reporting. 
  • Communicate and present the results of the performance analysis and modeling to stakeholders, and senior leadership. And provide a concrete recommendation. 
  • Cross team collaboration and working across the organization to identify opportunities and develop strategies. 

 

 

PREFERRED EXPERIENCE: 

  • Multiple years of technical experience in performance optimization. 
  • Strong technical expertise and experience in performance analysis, projection, and network hardware architecture. 
  • Deep knowledge and hand-on experience of AI Frameworks such as PyTorch, JAX, vLLM, and SGLang. 
  • Strong technical leadership skills, ability to work collaboratively with cross-functional teams. 
  • Mentor, coach, and inspire a diverse and talented team of researchers and engineers. 
  • Excellent written, verbal, and presentation skills, ability to coordinate internally and externally. 

 

ACADEMIC CREDENTIALS: 

  • A PhD or master's degree in computer science, electrical engineering, or a related field. 

LOCATION:

San Jose, CA (hybrid)

 

This role is not eligible for visa sponsorship.

 

#LI-MV1




Benefits offered are described: AMD benefits at a glance.

 

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

 

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.

 

This posting is for an existing vacancy.



THE ROLE: 

Senior level engineer who will be responsible for driving AMD’s strategy, architecture, optimization and tooling to achieve industry-leading AI Pre-training and Distributed Inference Performance on AMD GPU. You will partner across hardware architecture, AI frameworks, compilers, runtime, ROCm, developer tools and model to scale performance analysis and optimization. 

 

As an engineer of Collectives and Network performance, you will help drive the end-to-end technical performance attainment across the entire software stack focusing on getting the best performance on multiple generations of AMD GPUs with wide range of models including latest state-of-the-art AI models. You will help set the strategy and roadmap for general optimization, accelerating supporting new models and out of box performance. 

 

If you are passionate about performance optimization, getting the best out of the hardware, and shaping the future of AI acceleration, then this role is for you. 

 

 

THE PERSON: 

The ideal candidate will have deep knowledge with Network, NIC and GPU hardware architecture, software optimization, performance modeling, AI frameworks and latest trend in inference and training optimization. Hand-on experience in mapping model architecture to low level software, hardware and understanding the impact of each layer of the stack on model performance. Strong knowledge in latest generative model architecture, especially SoTA models, distributed inference and deployment at scale is crucial. 

 

KEY RESPONSIBILITIES: 

  • Help set strategy and roadmap for AMD Collectives and Network optimizations. 
  • Provide guidelines to customers on efficient network load-balancing, workload scheduling and model sharding strategies. 
  • Performance tuning, profiling and analysis of large-scale models for LLM, diffusion, multimodal, RecSys and generative AI, single node and distributed. In addition to exploring various tradeoffs and design decisions. 
  • Participate in hardware-software co-design for future hardware optimizations – especially on scale-up networks, NIC and scale-out networks. 
  • Develop and improve framework, tools and infrastructure for performance estimation, modeling and reporting. 
  • Communicate and present the results of the performance analysis and modeling to stakeholders, and senior leadership. And provide a concrete recommendation. 
  • Cross team collaboration and working across the organization to identify opportunities and develop strategies. 

 

 

PREFERRED EXPERIENCE: 

  • Multiple years of technical experience in performance optimization. 
  • Strong technical expertise and experience in performance analysis, projection, and network hardware architecture. 
  • Deep knowledge and hand-on experience of AI Frameworks such as PyTorch, JAX, vLLM, and SGLang. 
  • Strong technical leadership skills, ability to work collaboratively with cross-functional teams. 
  • Mentor, coach, and inspire a diverse and talented team of researchers and engineers. 
  • Excellent written, verbal, and presentation skills, ability to coordinate internally and externally. 

 

ACADEMIC CREDENTIALS: 

  • A PhD or master's degree in computer science, electrical engineering, or a related field. 

LOCATION:

San Jose, CA (hybrid)

 

This role is not eligible for visa sponsorship.

 

#LI-MV1



Benefits offered are described:  AMD benefits at a glance.

 

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law.   We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

 

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.  AMD’s “Responsible AI Policy” is available here.

 

This posting is for an existing vacancy.



Tags: No, USD $178,500.00/Yr., USD $255,000.00/Yr., US Careers (External)

Get Software Engineer jobs like this

New roles from thousands of companies land hourly, straight from their careers pages. Get the freshest matches by email so you never miss one.

Email me new jobs
Stryker logo

Staff Engineer

Gurugram, India
✓ From careers page· 13m ago
Stryker logo

Engineer, Automations (AI/ML), Technical Support

Bangalore, India
✓ From careers page· 15m ago
TripleLift logo

Senior Software Engineer

Zürich, Zürich
✓ From careers page· 26m ago
TripleLift logo

Cloud Engineer

$105k–$140kNew York, NY
✓ From careers page· 26m ago

Frequently asked questions

What is the salary for Software Development Engineer, Collectives and Network Optimization at AMD?

The estimated salary range for Software Development Engineer, Collectives and Network Optimization at AMD is $179,000 - $255,000 USD per year.

What skills are required for Software Development Engineer, Collectives and Network Optimization at AMD?

The required skills for Software Development Engineer, Collectives and Network Optimization at AMD include: Python, PyTorch, Machine Learning, Deep Learning.

What is the seniority level for Software Development Engineer, Collectives and Network Optimization at AMD?

Software Development Engineer, Collectives and Network Optimization at AMD is a Senior / Staff level position.

How do I apply for Software Development Engineer, Collectives and Network Optimization at AMD?

You can view the full description and apply for Software Development Engineer, Collectives and Network Optimization at AMD on EchoJobs: https://echojobs.io/job/amd-sr-staff-software-development-engineer-collectives-and-network-optimization-9dqpf.