Meta
Posted 2d ago

Compiler Engineer, Graph Compiler Performance Optimization

Meta
Menlo Park, California, United States
$184k-$257k/yrOnsiteFull Time
Responsibilities
  • optimizing compilers
  • profiling models
  • collaborating with researchers
Requirements
  • Requires C/C++ and Python
  • Compiler or performance optimization experience
  • Deep learning execution knowledge, and a bachelor's degree or equivalent practical experience. Advanced compiler and AI hardware experience preferred
Technical tools mentioned
C/C++PythonPyTorchFX IRInductorTorchDynamoXLATVMMLIRGlowTensorFlowJAXTritonLLVM

Job description

In this role, you will be a member of the MTIA (Meta Training & Inference Accelerator) Software team and part of the bigger AI and Compute Foundations team. The Graph Compiler team drives the development of the top-of-stack compilation pipeline for MTIA — taking PyTorch models, tracing the model graph, lowering and optimizing through FX/Inductor, and mapping to high-performance kernels. Your specific focus will be on performance optimization within the graph compiler: designing and implementing compiler passes that maximize throughput and minimize latency for AI workloads on MTIA hardware.

You will work closely with AI researchers to understand emerging model architectures and translate performance requirements into compiler optimization strategies. You will partner with hardware design teams to drive hardware-software co-design, ensuring the compiler exploits new silicon capabilities from day one. You will also collaborate with the Triton/DSL and LLVM compiler teams to deliver cross-stack performance improvements.

Responsibilities

  • Design, implement, and validate graph-level compiler optimization passes targeting performance within the PyTorch Inductor / FX IR compilation pipeline for MTIA
  • Profile and analyze deep learning models to identify graph-level performance bottlenecks such as suboptimal fusion boundaries, excessive memory traffic, and scheduling inefficiencies — then develop compiler solutions to address them
  • Develop and extend automatic fusion strategies to unlock peak hardware utilization across MTIA chip generations
  • Implement memory optimizations including data placement strategies, memory footprint reduction, and data movement elimination to reduce latency and improve bandwidth utilization
  • Build and improve performance analysis tooling to accelerate optimization iteration cycles
  • Collaborate with hardware design teams on hardware-software co-design: informing hardware features from compiler needs and rapidly developing compiler support for new chip capabilities
  • Ensure compiler optimizations are portable and scalable across MTIA chip generations, enabling rapid software bring-up for new silicon
  • Partner with AI researchers to co-design model architectures and compiler optimizations, ensuring models run performantly on MTIA out of the box

Minimum Qualifications

  • Experience with C/C++ and Python programming
  • Experience in compiler development, performance optimization, or accelerating deep learning models on hardware architectures
  • Understanding of deep learning model execution (graphs, operators, data flow) and how compiler transformations affect end-to-end performance
  • Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience

Preferred Qualifications

  • A Bachelor's degree in Computer Science, Computer Engineering, relevant technical field and 7+ years of experience in compiler development or accelerating deep learning models on hardware architectures OR a Master's degree and 4+ years OR a PhD and 3+ years of relevant experience
  • Experience with graph-level compiler optimizations such as operator fusion, graph scheduling, memory allocation optimization, dead code elimination, or constant folding in ML compiler stacks
  • Experience with PyTorch internals, PyTorch 2.0 compilation stack (TorchDynamo, FX IR, Inductor), or similar ML compilation frameworks (XLA, TVM, MLIR, Glow)
  • Experience with performance profiling and analysis: identifying compute/memory/I/O bottlenecks, understanding roofline models, and developing systematic approaches to performance tuning
  • Experience with traditional compiler optimizations (loop transformations, vectorization, parallelization, instruction scheduling) and how they apply to ML workloads
  • Experience with AI hardware accelerator architectures (GPUs, TPUs, or custom ASICs) and understanding of how hardware constraints inform graph-level optimization decisions
  • Experience working with deep learning frameworks (PyTorch, TensorFlow, JAX) and understanding their compilation and execution models
  • Experience adhering to and implementing responsible, ethical AI practices (e.g., risk assessment, bias mitigation, quality and accuracy reviews)
  • Demonstrated ongoing AI skill development (e.g., prompt/context engineering, agent orchestration) and staying current with emerging AI technologies
  • Demonstrated ability to integrate AI tools to optimize/redesign workflows and drive measurable impact (e.g., efficiency gains, quality improvements)

Compensation

  • $183,997/year - $257,000/year; Country: US; Bonus eligible; Equity eligible

About Meta

Meta builds technologies that help people connect, find communities, and grow businesses. When Facebook launched in 2004, it changed the way people connect. Apps like Messenger, Instagram and WhatsApp further empowered billions around the world. Now, Meta is moving beyond 2D screens toward immersive experiences like augmented and virtual reality to help build the next evolution in social technology. People who choose to build their careers by building with us at Meta help shape a future that will take us beyond what digital connection makes possible today—beyond the constraints of screens, the limits of distance, and even the rules of physics.

California Notice

For those who live in or expect to work from California if hired for this position, please click here for additional information.

Equal Opportunity

Meta is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or other applicable legally protected characteristics. You may view our Equal Employment Opportunity notice here.

Accommodations

Meta is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, fill out the Accommodations request form.

About Meta

Develops social networking platforms and virtual reality technologies.

Similar jobs

Compiler Engineer roles near Menlo Park, California
15h
Save
Mark Applied
Hide
Senior Principal Compiler Engineer
Austin or San Jose
$180k-$240k/yr RemoteFull Time
SambaNova Systems
SambaNova Systems: Develops custom AI hardware and software for enterprise computing.
5+ YOEBachelor’s or master’s degree in computer science, computer engineering, or equivalent; 5–10 years of industry experience; deep compiler knowledge and software product development experience.
TensorFlow, PyTorch, MLIR
1d
Save
Mark Applied
Hide
Senior Deep Learning Compiler Engineer - XLA
Santa Clara or Austin or Texas or Washington or California or Redmond
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
4+ YOEBachelor's, master's, Ph.D., or equivalent experience; 4+ years in compiler optimization and performance analysis; strong C/C++ and hardware architecture skills; experience with distributed programming and high-performance computing.
JAX, OpenXLA, NVIDIA GPUs, MLIR, LLVM, OpenAI Triton, CUDA, OpenCL, TVM, PyTorch, TensorFlow, C, C++
5d
Save
Mark Applied
Hide
Senior Compiler Engineer Infrastructure
Santa Clara or Austin or Texas or Washington or California or Redmond
$152k-$242k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
3+ YOEBachelor's, master's, or doctoral degree in computer science, computer engineering, or related field; 3+ years with large-scale codebases; strong C++, compiler internals, open-source frameworks, and developer infrastructure experience.
LLVM, Clang, MLIR, C++, CUDA, CI
1w
Save
Mark Applied
Hide
Senior Compiler Engineer
San Jose or Pittsburgh
$160k-$210k/yr OnsiteFull Time
Efficient Computer
Efficient Computer: Developing ultra-low-power general-purpose processors for edge AI computing.
6+ YOE6+ years C++ experience, BS/MS in CS or related, familiarity with GCC/LLVM/MLIR, understanding of computer architecture, experience with GDB, strong problem-solving and communication skills.
C++, GCC, LLVM, MLIR, GDB, Verilog, SystemVerilog, VHDL
3w
Save
Mark Applied
Hide
Swift Compiler Engineer
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience with programming languages and compiler technology, passion for developer experience, collaborative mindset, and strong desire to learn.
Swift
2mo
Save
Mark Applied
Hide
Compiler Engineer (PyTorch / Triton / LLVM / MLIR)
Los Altos, California, United States
OnsiteFull Time
Majestic Labs
Majestic Labs: Developing memory-first AI server platforms for data centers.
7+ YOEBachelor's or Master's in CS/CE, 7+ years compiler engineering experience, deep LLVM/MLIR knowledge, PyTorch and Triton experience, proficiency in C/C++, performance optimization and strong communication skills.
PyTorch, PyTorch Inductor, Dynamo, Triton, LLVM, MLIR, C/C++
2mo
Save
Mark Applied
Hide
Compiler Engineer — MLIR
Mountain View, California, United States
$200k-$360k/yr OnsiteFull Time
DensityAI
DensityAI: Designing custom AI hardware accelerators for large language models.
5+ YOE5+ years compiler engineering experience with MLIR or equivalent IR design; strong C++; experience with tensor compilation, distributed/sharded execution, async/streaming dataflow, and integrating ML frameworks.
MLIR, C++, PyTorch, JAX, ONNX, TensorFlow, LLVM, Triton, IREE, XLA
3mo
Save
Mark Applied
Hide
Member of Technical Staff, Hardware, Compiler Engineer
Palo Alto or Austin
$200k-$420k/yr OnsiteFull Time
River AI
River AI: Building user-controlled, personalized AI with integrated local hardware.
5+ YOEBachelor's in EE/CE and 5+ years experience with advanced process nodes; deep experience with MLIR/XLA and PyTorch internals; strong C/C++ skills and computer architecture knowledge.
PyTorch, MLIR, XLA, LLVM, Triton, CUDA, C, C++