ADVANCE YOUR CAREER. ADVANCE THE WORLD.
At AMD, we believe technology can change lives for the better. It can heal us, entertain us, and make us more connected, productive, and understanding of the world around us. And we’re looking for talent who feel the same: people who want to leave the planet better than they found it, those who don’t shy away from humanity’s challenges but are determined to help solve them.
Â
AMD is powering the next generation of supercomputing, high-performance computing, cloud, and AI. Whether you’re designing next-gen processors, enabling AI breakthroughs, or creating go-to-market plans, every role at AMD contributes to something bigger — technology that moves the world forward.
THE ROLE:Â
Â
THE PERSON:Â
Â
KEY RESPONSIBILITIES:Â
- Design and develop native distributed communication and execution capabilities within the Triton AMDGPU backend, enabling scalable multi-GPU execution for large-scale AI workloads
- Design and implement Triton compiler and runtime mechanisms for native GPU-initiated communication, including collective operations, remote memory access, synchronization, and distributed execution primitives
- Drive performance optimization across compute and communication, including inter-GPU data movement, communication/computation overlap, memory hierarchy utilization, and GPU-driven scheduling efficiency
- Develop and optimize distributed Triton kernels and execution models to achieve high performance, scalability, and efficient hardware utilization for AI workloads
- Analyze, profile and debug complex cross-stack issues spanning Triton compiler, runtime, ROCm stack, and GPU hardware execution
- Collaborate closely with GPU architecture, compiler, runtime, and performance teams to co-design and enable next-generation distributed GPU programming and execution capabilities
- Contribute to open-source Triton and ROCm distributed ecosystem, driving innovation in distributed GPU computing
Â
PREFERRED EXPERIENCE:Â
- Deep experience in compiler development, GPU software, distributed systems, or performance engineering
- Familiarity or hands-on experience with Triton compiler and runtime
- Deep understanding of modern GPU architectures, including execution model, memory hierarchy (LDS, L2, HBM), scheduling, occupancy, and hardware performance characteristics
- Good understanding of GPU runtime systems, communication stacks, and multi-GPU interconnects such as XGMI, NVLink, PCIe, or InfiniBand and their performance implications Â
- Familiarity with distributed GPU communication libraries such as RCCL, NCCL, NVSHMEM, rocSHMEM, or MPI and similar technologies
- Experience developing, optimizing, and scaling workloads across multiple GPUs, including inter-GPU communication, synchronization, and communication/computation overlap
- Strong experience with GPU programming using Triton, HIP, CUDA, or similar parallel programming environments
- Strong knowledge of MLIR and/or LLVM internals
- Experience profiling, debugging, and optimizing performance across compiler, runtime, and hardware layers
- Familiarity with ROCm, HIP, CUDA, or similar GPU programming ecosystems, including performance profiling and optimization tools
- Experience optimizing large-scale AI, machine learning or HPC workloads across multi-GPU systems
- Experience contributing to open-source projects and working in collaborative, cross-functional engineering environments
- Strong problem-solving, communication, and technical leadership skills
Â
PREFERRED ACADEMIC CREDENTIALS:Â
- Bachelor’s or Master's Degree in Computer Engineering, Computer Science, Electrical Engineering or equivalent practical experience
Â
This role is not eligible for visa sponsorship.
Â
#LI-G11
Â
#LI-HYBRID
Benefits offered are described:Â AMD benefits at a glance.
Â
AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law.  We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.
Â
AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD’s “Responsible AI Policy” is available here.
Â
This posting is for an existing vacancy.
Learn more about this Employer on their Career Site
