GPU Distributed System Researcher
ADVANCE YOUR CAREER. ADVANCE THE WORLD.
At AMD, we believe technology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.
Whether you’re designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger — technology that moves the world forward. Join us and, together, we’ll advance your career.
THE ROLE:
AMD Research and Advanced Development is seeking a researcher to invent and implement novel programming models and runtime environments for high performance computing systems comprised of heterogeneous processors, accelerators, and scale-up/scale-out networks.
THE PERSON:
You love inventing and implementing new features, APIs and abstractions to increase efficiency and optimize performance. You love bringing new systems to life and developing software that shows the full power of novel hardware architectures. You want to help define the software stack of AMD’s future hardware accelerators. Are you ready for the challenge? Then join us!
KEY RESPONSIBILITIES:
- Invent, design and implement novel programming models and runtime environments for high performance computing systems comprised of heterogeneous processors and accelerators.
- Implement innovative software solutions and demonstrate their effectiveness on prototype hardware.
- Work on software enhancements that aim to potentially improve programmer productivity.
- Drive the results of the research projects into product roadmaps.
- Collaborate with teams within AMD Research and Advanced Development, product groups, and vendor partners to help improve AMD’s ML/HPC ecosystem.
- Submit patentable inventions.
- Clearly communicate research findings to academic, product and commercial audiences.
PREFERRED EXPERIENCE:
- Must have experience in developing and debugging GPU and/or multi-threaded CPU applications. Distributed system experience is desired as well.
- Prior exposure to ML/HPC/datacenter software, middleware and device driver development, and familiarity with ML frameworks, OpenSHMEM, or MPI programming models are preferred.
- The position involves investigating networking architectures with an emphasis on optimizing cluster-scale applications that are bounded by memory and/or inter-GPU communication performance.
- Demonstrated experience in ideation, evaluation, and optimization of a research project.
- Excellent written and oral communication skills, ability to organize and present complex technical information.
- Strong analytical skills.
- Strong skills in C/C++, Python and/or GPU programming.
- Use of networking simulation platforms (NS-3, OMNeT++, etc.).
- Experience with High Performance Computing networking.
- Experience with the AMD ROCm software stack.
- Version control systems such as Git.
ACADEMIC CREDENTIALS:
PhD degree in Computer Engineering / Electrical Engineering preferred.
LOCATION: Bellevue, WA preferred; San Jose, CA or Austin, TX are possibilities
This role is not eligible for visa sponsorship.
Benefits offered are described: AMD benefits at a glance.
AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.
AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.
This posting is for an existing vacancy.
Compensation: Expected salary range: USD 151,900.00 per year to USD 217,000.00 per year

