AMD
Posted 1mo ago

GPU Compiler Development Engineer

AMD
Beijing, Beijing, China
¥412k-¥588k/yrOnsiteFull Time
Responsibilities
  • implementing lowerings
  • optimizing performance
  • debugging issues
Requirements
  • Strong C/C++ and object-oriented programming
  • MLIR/LLVM and GPU/parallel programming experience
  • Familiarity with HIP/CUDA
  • ONNX/ONNX Runtime
  • Windows and Linux development
  • Debuggers
  • GitHub and profilers
  • Bachelor’s degree in CS/CE/EE or equivalent
Technical tools mentioned
MLIRLLVMHIPCUDAONNXONNX RuntimeWindowsLinuxGitHub

Job description



WHAT YOU DO AT AMD CHANGES EVERYTHING 

At AMD, our mission is to build great products that accelerate next-generation computing experiences—from AI and data centers, to PCs, gaming and embedded systems. Grounded in a culture of innovation and collaboration, we believe real progress comes from bold ideas, human ingenuity and a shared passion to create something extraordinary. When you join AMD, you’ll discover the real differentiator is our culture. We push the limits of innovation to solve the world’s most important challenges—striving for execution excellence, while being direct, humble, collaborative, and inclusive of diverse perspectives. Join us as we shape the future of AI and beyond.  Together, we advance your career.  




THE ROLE: 

AMD is looking for a GPU compiler engineer to go deep on GPU compilers and performance within the ROCm Execution Provider (EP), which runs large language and vision models on AMD GPUs through ONNX Runtime. You will be a member of a core team of incredibly talented industry specialists and will work with the very latest hardware and software technology, implementing and optimizing operator lowerings and helping us run on more AMD GPU targets..

 

THE PERSON: 

The ideal candidate should be passionate about software engineering and possess the drive to take sophisticated issues to resolution. Able to communicate effectively and work optimally with the GPU compiler, MLIR, and kernel teams across AMD.

 

KEY RESPONSIBILITIES: 

  • Work with AMD’s architecture specialists to improve future products 
  • Implement and unit-test ONNX operator lowerings in MLIR
  • Apply a data-minded approach to close per-operator performance gaps versus competitive runtimes
  • Stay informed of software and hardware trends and innovations, especially pertaining to algorithms and architecture 
  • Design and develop new groundbreaking AMD technologies 
  • Participating in new ASIC and hardware bring ups  
  • Debug and fix existing issues and research alternative, more efficient ways to accomplish the same work
  • Develop technical relationships with peers and partners 

 

PREFERRED EXPERIENCE: 

  • Strong object-oriented programming background, C/C++ preferred 
  • Ability to write high quality code with a keen attention to detail 
  • Working knowledge of MLIR or LLVM and GPU or parallel programming
  • HIP / CUDA, ONNX / ONNX Runtime, and performance profiling and optimization is a plus
  • Experience with Windows and Linux development; familiarity with debuggers, source code control systems (GitHub), and profilers
  • Familiarity with AI coding agents / assistants and a desire to use them to improve efficiency (preferred)
  • Working proficiency in English
  • Effective communication and problem-solving skills 

 

ACADEMIC CREDENTIALS: 

  • Bachelor’s or Master's degree in Computer Science, Computer Engineering, Electrical Engineering, or equivalent

 

  • #LI-JW2



Benefits offered are described:  AMD benefits at a glance.

 

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law.   We encourage applications from all qualified candidates and will accommodate applicants’ needs under the respective laws throughout all stages of the recruitment and selection process.

 

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position.  AMD’s “Responsible AI Policy” is available here.

 

This posting is for an existing vacancy.

About AMD

Designs and manufactures computer processors and graphics technology.

Similar jobs

Compiler Engineer roles near Beijing, Beijing
3w
Save
Mark Applied
Hide
顶尖应届-AI Compiler 开发工程师-芯片
Beijing, Beijing, China
OnsiteFull Time
Xiaomi
XiaomiHong Kong Stock Exchange: 1810: Designs and manufactures smartphones, consumer electronics, and electric vehicles.
PhD in computer architecture or compiler technology; research experience in AI compiler for GPGPU/NPU; strong problem-solving and implementation skills; familiarity with LLVM and GPU/SIMT architectures.
LLVM, AutoTuning, GPGPU, NPU, Tensor Core, SIMT, NVGPU
1mo
Save
Mark Applied
Hide
硬件加速算子编译器工程师-Data
Beijing, Beijing, China
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Experience in compiler development and optimization for hardware-accelerated platforms; strong C/C++ and algorithms background; familiarity with LLVM/Clang and CPU architectures (x86_64, ARM64, RISC-V).
Clang, LLVM, GCC, C, C++, x86_64, ARM64, RISC-V
4mo
Save
Mark Applied
Hide
AI 编译器开发工程师
Beijing or Shanghai or Hangzhou
OnsiteInternship
Sunrise AI
Sunrise AI: Developer of high-performance AI inference GPU chips and solutions.
Bachelor's in CS or microelectronics, experience with Triton/TileLang/MLIR, strong C++, CUDA and Python skills, knowledge of GPGPU and heterogeneous programming, good cross-team communication.
Triton, TileLang, MLIR, TVM, LLVM, C++, CUDA, Python
8mo
Save
Mark Applied
Hide
AI编译器开发工程师-2026
Beijing or Shanghai or Chengdu or Hangzhou
OnsiteFull Time
曦望Sunrise
曦望Sunrise: Designs and produces high-performance AI inference GPU chips.
5+ YOE5+ years compiler development experience; proficient in C++, CUDA, Python; experience with Triton/TileLang, MLIR/TVM/LLVM; knowledge of GPGPU architecture and heterogeneous programming; bachelor\u0002s in CS/HPC/microelectronics.
Triton, TileLang, MLIR, TVM, LLVM, C++, CUDA, Python
11mo
Save
Mark Applied
Hide
2026校招-编译器开发工程师-北京
Beijing or Hangzhou
OnsiteFull Time
Smart Logic Technology
Smart Logic Technology: Designs and develops system-on-chip semiconductor solutions.
Master's degree in CS or related field, knowledge of computer architecture, proficiency in C/C++, experience with LLVM or GCC preferred, understanding of compiler backend and optimizations.
C, C++, LLVM, GCC
1y
Save
Mark Applied
Hide
编译器开发工程师 - 实习
Beijing, Beijing, China
OnsiteInternship
SmartLogic Tech
SmartLogic Tech: Developing AI-driven solutions for numerical calculation and life sciences.
Proficient in C/C++, good coding practices, experience with LLVM or GCC compiler frameworks; role focused on compiler and tool development.
C, C++, LLVM, GCC
1y
Save
Mark Applied
Hide
编译器开发工程师
Shanghai or Tianjin or Beijing or Hangzhou or Haikou
OnsiteFull Time
BrightChip
BrightChip: Specializes in developing high-performance AI acceleration chips and hardware.
Bachelor's degree in computer/electronic/software engineering, proficiency in C++/CUDA, experience with LLVM/GCC and compiler optimization, familiarity with GPGPU/NPU performance tuning, strong teamwork and communication.
C++, CUDA, LLVM, GCC, GPGPU
1y
Save
Mark Applied
Hide
编译器开发实习/校招岗位
Nanjing or Shanghai or Beijing
OnsiteInternship
Houmo.AI
Houmo.AI: Developer of computing-in-memory AI chips for edge computing.
Experience with C++ or Python, basic machine learning knowledge, understanding of ML model training/inference and computer architecture. Familiarity with TensorFlow/PyTorch, MLIR, or ONNX is a plus.
C++, Python, TensorFlow, Pytorch, MLIR, ONNX