GPU Performance Engineer – AI Kernels
ADVANCE YOUR CAREER. ADVANCE THE WORLD.
At AMD, we believe technology has the power to solve the world’s most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.
Whether you’re designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger — technology that moves the world forward. Join us and, together, we’ll advance your career.
GPU PERFORMANCE ENGINEER - AI KERNELS
THE ROLE:
AMD is seeking a software engineer with expertise in performance optimization who is passionate about improving the performance of key AI workloads and benchmarks. You will be part of a highly skilled team working with the latest hardware and software technologies to deliver cutting-edge solutions for AI and GPU computing.
THE PERSON:
The ideal candidate is passionate about software engineering and brings the technical leadership skills needed to drive complex challenges to resolution. This person communicates effectively, collaborates successfully across diverse teams, and thrives in a fast-paced, innovative environment.
KEY RESPONSIBILITIES:
- Continuously identify and implement improvements to machine learning kernels for AMD GPUs, with a focus on performance and power efficiency.
- Stay informed about software and hardware trends, particularly in GPU architecture and machine learning algorithms.
- Improve development workflows and CI infrastructure to enable faster and more reliable delivery.
- Design and develop innovative GPU and machine learning technologies.
- Debug and resolve existing issues, while researching and implementing more efficient alternatives.
- Build and maintain strong technical relationships with internal teams and external partners.
PREFERRED EXPERIENCE:
- Strong understanding of modern GPU architectures.
- 3+ years of GPU software development experience using HIP, CUDA, or OpenCL.
- Experience working directly with hardware ISA is highly valued.
- 5+ years of system-level programming experience in C++ (C++17 or later preferred).
- Experience with GPU profiling, debugging, benchmarking, and performance analysis tools.
- Background in high-performance computing (HPC) or other performance-critical systems.
- Familiarity with modern machine learning frameworks such as PyTorch and MIOpen.
- Experience with tile-based programming models and frameworks (e.g., Triton, CUTLASS).
- Strong written and verbal communication skills in English.
ACADEMIC CREDENTIALS:
- Bachelor's or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or a related field, or equivalent practical experience.
Benefits offered are described: AMD benefits at a glance.
AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.
AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.
This posting is for an existing vacancy.
Compensation: Expected salary range: EUR 57,750.00 per year to EUR 82,500.00 per year

