25 deep learning compiler engineer jobs at 8 companies in United States

1w
Save
Mark Applied
Hide
Deep Learning Compiler Engineer
Santa Clara or Austin or Washington or Oregon or California or Redmond
$152k-$242k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
3+ YOEBachelor's, master's, or Ph.D. in computer science, computer engineering, or related field; 3+ years in compiler optimization, performance analysis, or IR design; strong C/C++ skills.
CUDA Tile, MLIR, LLVM, XLA, TVM, CUDA, OpenCL, C, C++
1w
Save
Mark Applied
Hide
Deep Learning Compiler Engineer
Santa Clara or Austin or Washington or Oregon or California or Redmond
$152k-$242k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
3+ YOEBachelor's, master's, or Ph.D. in computer science, computer engineering, or related field; 3+ years in compiler optimization, performance analysis, or IR design; strong C/C++ and software design skills.
CUDA Tile, MLIR, LLVM, XLA, TVM, CUDA, OpenCL, C, C++
1w
Save
Mark Applied
Hide
Deep Learning Compiler Engineer
Santa Clara or Austin or Oregon or Washington or California or Redmond
$152k-$242k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
3+ YOEBachelor's, master's, or Ph.D. in computer science, computer engineering, or related field; 3+ years in compiler optimization, performance analysis, or IR design; strong C/C++ and software design skills.
C, C++, CUDA, OpenCL, MLIR, LLVM, XLA, TVM
2mo
Save
Mark Applied
Hide
Staff Machine Learning Compiler Engineer
Palo Alto, California, United States
$206k-$258k/yr OnsiteFull Time
Rivian
RivianNASDAQ: RIVN: Electric adventure vehicle and technology manufacturer.
Ph.D. or M.S. in computer engineering or related field; strong C/C++ and Python skills; experience with SOC platforms for ML, deep learning models, and compiler pipeline development; familiarity with low-level IRs and hardware-aware optimizations.
C, C++, Python, SOC, Rivian Autonomy Processor (RAP1)
2w
Save
Mark Applied
Hide
Compiler Engineer, Graph Compiler Performance Optimization
Menlo Park, California, United States
$184k-$257k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Builds technologies that help people connect, find communities, and grow businesses.
3+ YOERequires C/C++ and Python, compiler or performance optimization experience, deep learning execution knowledge, and a bachelor's degree or equivalent practical experience. Advanced compiler and AI hardware experience preferred.
C/C++, Python, PyTorch, FX IR, Inductor, TorchDynamo, XLA, TVM, MLIR, Glow, TensorFlow, JAX, Triton, LLVM
2mo
Save
Mark Applied
Hide
Principal Compiler Engineer - ML Systems
United States
$200k-$275k/yr RemoteFull Time
SambaNova Systems
SambaNova Systems: AI infrastructure providing a full-stack platform for enterprises, AI labs, service providers, and sovereign AI initiatives.
5+ YOE5+ years industry experience with a Bachelor’s or Master’s in CS/CE; deep compiler fundamentals, experience building/deploying software, familiarity with deep learning frameworks and MLIR, and strong performance-debugging and cross-disciplinary collaboration skills.
PyTorch, TensorFlow, MLIR, SambaNova Suite, SN40L
3w
Save
Mark Applied
Hide
Machine Learning Backend Engineer Graduate (AML MLDev) - 2027 Start
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Global technology specializing in AI-powered content platforms.
Bachelor's or master's degree in a technical discipline; C/C++ development, assembly, CPU/GPU architecture, cache mechanisms, model compilation stacks, and deep learning framework optimization experience required.
TVM, MLIR, XLA, C, C++, TensorFlow, PyTorch, OneFlow, MxNet
3w
Save
Mark Applied
Hide
Machine Leaning Performance Engineer (Inference)
New York City, New York, United States
$200k-$300k/yr HybridFull Time
Tower Research Capital
Tower Research Capital: Proprietary quantitative trading firm employing traders, engineers, researchers, and business-support staff to trade global financial markets.
2+ YOERequires 2+ years optimizing deep learning inference, PyTorch or JAX, Python/C++, mixed-precision computation, custom GPU kernels, optimization libraries, compilers, profiling tools, and GPU microarchitecture expertise.
PyTorch, JAX, Python, C++, Triton, TensorRT, ONNX, IREE, HLS4ML, cuBLAS, CUTLASS, Nsight Systems, Nsight Compute, FPGA, ASIC
2mo
Save
Mark Applied
Hide
SR. Software Development Engineer  – GPU Kernel Development
Santa Clara, California, United States
$240k-$360k/yr OnsiteFull Time
AMD
AMDNASDAQ: AMD: Leader in high-performance computing, graphics, and visualization technologies.
Expert C++ and Python developer experienced with GPU kernel development (HIP, CUDA, ASM), deep learning frameworks (TensorFlow, PyTorch), compiler internals (LLVM/ROCm), and performance optimization in Linux; advanced degree preferred.
TensorFlow, PyTorch, C++, Python, HIP, CUDA, ASM, Compute Kernel (CK), CUTLASS, Triton, LLVM, ROCm, Linux
2mo
Save
Mark Applied
Hide
Software Engineer, ML Infrastructure, Optimization
Mountain View, California, United States
$160k-$241k/yr OnsiteFull Time
Nuro
Nuro: Private U.S. autonomous-driving technology developing AI systems for automakers and mobility providers.
2+ YOE2+ years in ML optimization infrastructure; experience with quantization, pruning, ML compilers and GPU runtimes; proficient in Python, C++, CUDA and deep learning frameworks (PyTorch, JAX, TensorFlow, Keras).
Python, C++, CUDA, PyTorch, JAX, TensorFlow, Keras, FTL