ByteDance
Posted 1mo ago

Machine Learning Engineer - AI Compiler Optimization

ByteDance
San Jose, California, United States
OnsiteFull Time
Responsibilities
  • building compilers
  • optimizing models
  • collaborating teams
Requirements
  • Proficient with AI compiler frameworks and GPU/NPU compilation optimization
  • Experience with model import/conversion for PyTorch/TensorFlow and performance tuning for recommendation models
Technical tools mentioned
TritonMLIRTVMPyTorchTensorFlow

Job description

The mission of our AML team is to push the next-generation AI infrastructure and recommendation platform for the ads ranking, search ranking, live & ecom ranking in our company. We also drive substantial impact on core businesses of the company. Currently, we are looking for Machine Learning Engineer in AI Compiler Optimization to join our team to support and advance that mission.

Responsibilities:
- Responsible for building and implementing the compilation optimization system for the recommendation machine learning engine. Design and implement full-stack optimization solutions at the graph, operator, and memory levels specifically for recommendation model scenarios, including but not limited to graph-operator fusion and automatic operator generation, to maximize hardware computing limits.
- Collaborate closely with hardware and algorithm teams to carry out hardware-software co-design. Optimize compilation strategies based on hardware characteristics to improve the efficiency of hardware-software synergy.
- Responsible for the compilation adaptation of recommendation models from the PyTorch framework to the engine. Optimize the entire process of model import, conversion, and code generation to simplify the model deployment process and enhance development efficiency.

Minimum Qualifications:
- Proficient in one of the mainstream AI compiler frameworks (e.g., Triton, MLIR, TVM), with practical project experience in customized compilation optimization and Pass development based on the framework.
- Experience in GPU/NPU compilation optimization, mastering core techniques such as loop optimization, memory optimization, and operator optimization, with the ability to independently perform performance bottleneck analysis and technical optimization.
- Familiar with common model structures and compilation adaptation logic of deep learning frameworks such as PyTorch and TensorFlow, capable of designing targeted optimization solutions.

Preferred Qualifications:
- Familiar with the architecture design of recommendation machine learning engines. Solid implementation experience in compilation optimization and low-latency inference optimization for large-scale recommendation systems, with the ability to handle compilation optimization needs in high-concurrency scenarios.
- Experience contributing to open-source AI compiler projects (e.g., TVM, MLIR), or possess technical expertise in the compilation adaptation of large models for recommendation scenarios and the automatic generation of sparse operators.

About ByteDance

Global technology specializing in AI-powered content platforms.

Similar jobs

Machine Learning Engineer roles near San Jose, California
2h
Save
Mark Applied
Hide
Senior Machine Learning Engineer, Multimodal Perception
Mountain View, California, United States
$213k-$263k/yr HybridFull Time
Waymo
Waymo: Autonomous driving technology and robotaxi service provider.
2+ YOE2–5+ years training and releasing ML models in autonomous driving, robotics, or spatial AI; experience across perception and planning, learned policies, PyTorch/JAX, simulation evaluation, and production releases.
PyTorch, JAX, Vision-Language-Action (VLA), World Models
4h
Save
Mark Applied
Hide
Machine Learning Engineer, Performance Tooling
London or Sunnyvale
OnsiteFull Time
Wayve
Wayve: British autonomous-driving software licensing vehicle-agnostic AI Driver technology to automakers and fleet owners.
Requires deep performance engineering, end-to-end tool or service ownership, strong Python, PyTorch model development, data analysis, profiling, optimization, and quantitative communication skills.
Python, PyTorch
22h
Save
Mark Applied
Hide
Senior Machine Learning Engineer
Fremont, California, United States
$155k-$200k/yr HybridFull Time
Velo3D
Velo3DNasdaq Capital Market: VELO: Public metal additive manufacturer providing printers, software, and engineering services to aerospace, defense, energy, and industrial customers.
Experience with machine learning pipelines for geometric data and production software engineering; strong Python or C++ skills. Manufacturing, IoT, and industrial automation knowledge preferred.
Python, C++, Flow, Assure
2d
Save
Mark Applied
Hide
Founding Machine Learning Engineer
Mountain View, California, United States
$220k-$300k/yr OnsiteFull Time
Clera
Clera: AI-powered talent agent matching candidates to startup roles.
3+ YOE3–10 years as an ML Engineer, Applied Scientist, or Research Engineer; Python and PyTorch, TensorFlow, or JAX proficiency; ML fundamentals, distributed systems, cloud ML infrastructure, and MLOps experience.
Python, PyTorch, TensorFlow, JAX, AWS, GCP, Azure, Weights & Biases, MLflow
2d
Save
Mark Applied
Hide
Machine Learning Engineer Graduate (E-Commerce Knowledge Graph) - 2027 Start
San Jose or Los Angeles or Singapore or New York City or London or Dublin or Paris or Berlin or Dubai or Jakarta or Seoul or Tokyo
$128k-$317k/yr OnsiteFull Time
TikTok
TikTok: Short-form mobile video and social media platform.
Bachelor's degree in software development, computer science, computer engineering, or related field; machine learning exposure; programming familiarity; strong teamwork and communication skills.
C++, Python, Go, Java, TensorFlow, PyTorch, Hadoop, Spark, Hive, Flink
2d
Save
Mark Applied
Hide
Machine Learning Engineer III
San Jose or Seattle
$146k-$221k/yr HybridFull Time
Expedia Group
Expedia GroupNASDAQ: EXPE: Global travel technology powering online booking platforms.
3+ YOEBachelor’s or master’s degree in a quantitative field, 3+ years of machine learning or data-driven systems experience, Python and ML framework proficiency, and experience with data pipelines and large datasets.
Python, PyTorch, TensorFlow, Spark, SQL, Databricks, AWS
3d
Save
Mark Applied
Hide
Machine Learning Engineer
San Francisco, California, United States
$260k-$300k/yr OnsiteFull Time
Hyperbound
Hyperbound: Private AI sales-coaching platform helping enterprise revenue teams practice conversations and improve performance.
Own machine learning models end to end, including training, fine-tuning, production deployment, on-device execution, evaluation frameworks, benchmarks, and regression suites.
open source models
3d
Save
Mark Applied
Hide
Machine Learning Engineer, Causal Inference, Level 5
Los Angeles or Seattle or Palo Alto or New York City or Bellevue or Santa Monica or California or Washington or New York City
$209k-$313k/yr OnsiteFull Time
Snap Inc.
Snap Inc.NYSE: SNAP: Technology focused on augmented reality and communication.
5+ YOEBachelor’s degree or equivalent practical experience and 5+ years of post-bachelor’s machine learning experience, or equivalent master’s/PhD pathways, with causal inference and experimentation expertise.
Python, pandas, NumPy, scikit-learn, CausalML, CausalM, EconML, DoWhy