100 gpu compiler developer jobs at 28 companies in United States

2mo
Save
Mark Applied
Hide
Sr. GPU Compiler Developer
San Diego, California, United States
$179k-$286k/yr OnsiteFull Time
MediaTek
MediaTekTaiwan Stock Exchange: 2454: Global fabless semiconductor providing system-on-chip solutions.
10+ YOEDegree in CS/ECE, 10+ years product compiler experience, expertise in C/C++, LLVM and compiler optimizations, strong debugging, communication and teamwork skills.
C/C++, LLVM, D3D, Vulkan/OpenGL, OpenCL, AI
1w
Save
Mark Applied
Hide
Apple GPU Compiler Backend Engineer
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designing and manufacturing consumer electronics, software, and digital services.
Design, implement, and productize backend compiler optimizations for Apple Silicon GPUs, collaborating with hardware and software teams to solve compiler and performance challenges.
1mo
Save
Mark Applied
Hide
Senior Compiler Engineer – Rust GPU
Santa Clara or Austin or Seattle or United States
$152k-$242k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOE5+ years compiler or GPU codegen experience, deep Rust and compiler internals knowledge, expertise with LLVM/MLIR/PTX/CUDA, strong parallel programming and software design skills.
Rust, rustc, Cargo, CUDA, PTX, MLIR, LLVM
1w
Save
Mark Applied
Hide
GPU Compiler Engineer
Austin or Oregon or California or Santa Clara or Redmond
$124k-$242k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
2+ YOEBachelor's degree in computer science/engineering or equivalent, 2+ years of compiler code generation experience, excellent C++ skills, software engineering expertise, and strong communication skills.
C++, LLVM, CUDA, DirectX, OpenGL, Vulkan, OpenCL, Fortran, PTX, LLVM IR, Machine IR (MIR), GlobalISel, TableGen
1mo
Save
Mark Applied
Hide
Senior Compiler Engineer – Rust GPU
Santa Clara or Austin or Seattle or California or Texas or Washington
$152k-$242k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOE5+ years compiler or GPU codegen experience, deep Rust expertise (rustc/MIR), experience with LLVM/MLIR/PTX/CUDA, strong parallel programming and systems design skills.
Rust, rustc, MIR, Cargo, LLVM IR, MLIR, PTX, CUDA, Tensor Memory Accelerator (TMA), JIT
2d
Save
Mark Applied
Hide
Member of Technical Staff, GPU Compiler
San Francisco, California, United States
$285k-$315k/yr OnsiteFull Time
SF Tensor
SF Tensor: Privately held AI infrastructure providing cross-vendor training software to enterprises and AI labs.
Deep compiler infrastructure experience; GPU architecture and low-level optimization expertise; GPU ISA experience; ML compiler familiarity; C++ or Rust systems programming; production compiler infrastructure record.
LLVM, MLIR, CUDA, ROCm, PTX, SASS, GCN, RDNA, XLA, TVM, Triton, torch.compiler, C++, Rust, StableHLO, SMT solvers
3mo
Save
Mark Applied
Hide
Compiler Engineer — LLVM Backend
Mountain View, California, United States
$180k-$320k/yr OnsiteFull Time
DensityAI
DensityAI: Semiconductor startup building full-stack AI accelerators for frontier-scale large language model inference.
5+ YOEExpert LLVM backend internals, 5+ years compiler engineering with LLVM target development, strong computer architecture and ISA design knowledge, debugging compilation correctness and performance; MLIR/GPU/linker experience optional.
LLVM, TableGen, MLIR, Microsoft Excel
3w
Save
Mark Applied
Hide
GPU Driver Developer
Sunnyvale, California, United States
$175k-$250k/yr OnsiteFull Time
Bolt Graphics
Bolt Graphics: Private semiconductor startup developing energy-efficient graphics processors for creative, gaming, and research users.
Experience designing and implementing high-performance user/kernel drivers for Windows and Linux; proficiency in C/C++; knowledge of modern GPU APIs and compiler toolchains; BS/MS in related field preferred.
RISC-V, PCIe, Vulkan, Metal, DirectX, CUDA, SYCL, LLVM, MLIR, ISPC, GCC, C, C++, Windows, Linux
2w
Save
Mark Applied
Hide
Principal Compiler Architect GPU/ML
San Diego, California, United States
$186k-$265k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Leader in high-performance computing, graphics, and visualization technologies.
Bachelor's, master's, or PhD in computer science, computer engineering, electrical engineering, or equivalent; strong C/C++ programming, LLVM, GPU architecture, concurrency, and software development experience.
C, C++, LLVM, Windows, Linux, Android, GitHub
1w
Save
Mark Applied
Hide
Quadrants: Compiler Lead
San Carlos, California, United States
OnsiteFull Time
Genesis AI
Genesis AI: Private French full-stack robotics building general-purpose robots and physical-AI systems for factories, laboratories, hospitals, and homes.
Requires GPU optimization, Python API design, compiler IR/lowering/optimization, systems thinking, testing, benchmarking, and self-directed ownership; leadership growth expected.
Python, Arm64, x86, AMDGPU, CUDA, Metal, Vulkan, NumPy, PyTorch, Nvidia Warp, Numba, Triton, IR, C, PTX, Cholesky, eigendecomposition
1mo
Save
Mark Applied
Hide
Sr/Principal Machine Learning Compiler Developer
Sunnyvale, California, United States
OnsiteFull Time
Memwize
Memwize: Private AI infrastructure startup building system-level hardware and software for data-center-scale inference systems.
Build compiler capabilities spanning PyTorch, Triton, CUDA, and machine intermediate representation for leading-edge AI accelerators.
PyTorch, Triton, CUDA
1mo
Save
Mark Applied
Hide
Machine Learning Engineer - AI Compiler Optimization
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Global technology specializing in AI-powered content platforms.
Proficient with AI compiler frameworks and GPU/NPU compilation optimization; experience with model import/conversion for PyTorch/TensorFlow and performance tuning for recommendation models.
Triton, MLIR, TVM, PyTorch, TensorFlow
2mo
Save
Mark Applied
Hide
Software Dev Engineer, Machine Learning Compilers
Sunnyvale, California, United States
$165k-$224k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Multinational technology focused on e-commerce and cloud computing.
3+ YOE3+ years professional software development, 2+ years system design/architecture, proficiency in Java/C++/C#, experience deploying LLMs on GPUs/Neuron/TPU or other AI acceleration hardware; bachelor's degree preferred; embedded/C++ and compiler experience preferred.
Java, C++, C#, GPU, Neuron, TPU
1mo
Save
Mark Applied
Hide
Software Engineer, Systems ML - Compilers / Backend
Sunnyvale, California, United States
$154k-$217k/yr OnsiteFull Time
Reality Labs
Reality LabsNASDAQ: META: ’s AR/VR and wearable-technology business unit builds immersive hardware, software, research, and content for consumers.
2+ YOEBachelor's in CS or equivalent,2+ years building compilers/toolchains or code-optimization software,experience with Python and/or C/C++,and AI framework or model-acceleration experience on GPU/TPU/custom ASICs.
Python, C, C++, PyTorch, MLIR, Tensorflow, Caffe, LLVM, GCC, MSVC, Glow, GPU, TPU
4d
Save
Mark Applied
Hide
Software Engineer, ML Infra
San Francisco or New York City
$350k-$475k/yr OnsiteFull Time
Thinking Machines Lab
Thinking Machines Lab: Private AI research and product building customizable multimodal systems for researchers and the wider public.
Hands-on competence in at least four technical stacks, including systems such as Linux kernel, networking, GPUs/CUDA, distributed systems, storage, compilers, or observability; strong incident judgment required.
Linux, Linux kernel, NCCL, CUDA, NVIDIA
2mo
Save
Mark Applied
Hide
Software Engineer, ML Infrastructure, Optimization
Mountain View, California, United States
$160k-$241k/yr OnsiteFull Time
Nuro
Nuro: Private U.S. autonomous-driving technology developing AI systems for automakers and mobility providers.
2+ YOE2+ years in ML optimization infrastructure; experience with quantization, pruning, ML compilers and GPU runtimes; proficient in Python, C++, CUDA and deep learning frameworks (PyTorch, JAX, TensorFlow, Keras).
Python, C++, CUDA, PyTorch, JAX, TensorFlow, Keras, FTL
1mo
Save
Mark Applied
Hide
ML Infrastructure Engineer
Palo Alto, California, United States
$180k-$440k/yr OnsiteFull Time
xAI
xAI: Artificial intelligence research and development.
2+ YOE2+ years building large-scale production systems or ML infrastructure; degree in CS or related field or equivalent experience; strong Python and compiled-language skills; experience with GPU and distributed systems.
Python, C++, Rust, JAX, PyTorch, NVIDIA drivers, CUDA, Linux, Slurm, Puppet, Ansible
1mo
Save
Mark Applied
Hide
Software Engineer - Model Infrastructure, TikTok Feeds
San Jose, California, United States
$156k-$388k/yr OnsiteFull Time
TikTok
TikTok: Short-form mobile video and social media platform.
Bachelor's in CS or related field; understanding of GPU architecture and CUDA/CUTLASS/Triton Lang; experience with TensorFlow or PyTorch; familiarity with large-scale model training and DNN compilers (MLIR/XLA/TVM).
CUDA, CUTLASS, Triton Lang, TensorFlow, PyTorch, MLIR, XLA, TVM
1w
Save
Mark Applied
Hide
Software Engineer II, MLOps Framework
United States
$139k-$167k/yr RemoteFull Time
Torc Robotics
Torc Robotics: Private Daimler Truck subsidiary developing self-driving truck software and integration solutions for U.S. long-haul freight operators.
4+ YOEBachelor's degree plus 4+ years of relevant experience or master's degree plus 2+ years. Requires model conversion, compilation, benchmarking, edge deployment, release registry, and cross-platform performance validation experience.
ONNX, TensorRT, torch.compile, NVIDIA Orin, PTQ, QAT, INT8, FP8, FP4, BF16/FP16, C++, CUDA
1w
Save
Mark Applied
Hide
Staff Platform Engineer - Developer Infrastructure
Houston, Texas, United States
OnsiteFull Time
Persona AI
Persona AI: Private U.S. robotics building humanoid robots for welding, fabrication, shipyards, construction, and other heavy-industrial work.
8+ YOE8+ years operating production infrastructure, with CI/CD or developer platform ownership. Requires deep Linux systems expertise, CI systems, C/C++ build systems, cross-compilation, IaC, Python, Bash, and strong written communication.
C++, Python, Rust, CUDA, GitHub Actions, GitLab CI, Buildkite, Jenkins, Bazel, CMake, Ansible, Terraform, Bash, Linux, ARM64, Jetson, Yocto, Prometheus, Grafana, OpenTelemetry, Nix, Kubernetes, systemd