48 reinforcement learning engineer jobs at 29 companies in Tracy, CA
🚀PromotedHiringCafe
Founding Machine Learning / AI Search Engineer
Cupertino, CA, US
$160k-$310k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
Pony.aiNASDAQ: PONY: Develops autonomous driving technology and operates robotaxi services.
3+ YOEM.S./Ph.D. or equivalent experience; 3+ years building production ML with strong RL experience; depth in deep learning and generative models; distributed training and large-scale data processing; strong communication.
Elorian AI: AI lab building multimodal models for advanced visual reasoning.
3+ YOE3+ years distributed systems experience, strong Python and PyTorch or JAX, multi-node GPU orchestration (Ray, SLURM, Kubernetes), experience with actor-learner architectures and RL training pipelines.
Staff Reinforcement Learning Engineer – Whole Body Control
San Jose, California, United States
$150k-$250k/yrOnsiteFull Time
Figure: Develops autonomous humanoid robots for commercial and residential tasks.
Strong dynamics/control background for legged robots; RL for robotics (PPO, SAC); tune hyperparameters; apply domain randomization and curriculum learning; lead projects and mentor.
Lead AI Infrastructure Engineer, Reinforcement Learning
Santa Clara, California, United States
$179k-$306k/yrHybridFull Time
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Design and operate distributed RL training infrastructure at scale; strong systems experience in ML platforms; proficiency with PyTorch/JAX, NCCL/MPI-style distributed training, C++/Python performance tuning; Bachelor's degree required.
Pony.aiNASDAQ: PONY: Develops autonomous driving technology for passenger and freight transportation.
MS/PhD in CS/ML/AI or related field, or equivalent experience; RL/production ML experience; deep learning, sequence modeling, generative models; publish or ship impactful ML systems; large-scale training and data processing; lead ambiguous work.
Research Scientist / Engineer – Reinforcement Learning Infrastructure
Redwood City, California, United States
HybridFull Time
Luma AI: Develops multimodal AI for video generation and creative production.
Experience operating post-trained LLMs with reinforcement learning at scale, distributed PyTorch training, building RL environments and reward infrastructure, and debugging large asynchronous rollout pipelines.
Research Engineer – Reinforcement Learning (RL) Systems & Infrastructure (Seed Infra)
San Jose, California, United States
OnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
Design and build scalable RL systems and infrastructure for large-scale model training; expertise in distributed systems, GPU optimization, Python/C++, and RL workflows.
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOEBachelor's in CS (or equivalent), 5+ years relevant experience (including 3 years engineering), strong communication/presentation skills, RL and LLM pipeline experience, and proficiency with PyTorch, JAX, and NeMo-RL.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOEBachelor's in CS or related field; 5+ years exp including 3+ in engineering; strong communication; experience presenting to technical audiences; RL pipelines; PyTorch/JAX/NeMo-RL; able to work on training pipelines and demos.
UnityNYSE: U: Provides software for creating real-time 3D interactive content.
Expertise in machine learning, reinforcement learning, control theory, large-scale data analysis, online experimentation; strong software engineering in Python and deep learning (PyTorch).
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
3+ YOE BS with 3+ years in ML/AI engineering; knowledge of transformers, reinforcement learning; PyTorch; end-to-end ML pipelines; multimodal sensing integration.
PyTorch, Multimodal Sensing, ML Pipelines, On-device/Edge ML
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
Master's or PhD in CS/robotics/ML, deep reinforcement learning experience, strong C++ and Python skills, hands-on with JAX or PyTorch, experience bridging sim-to-real fidelity gaps and working with GPU/CUDA environments.
Research Scientist / Engineer – Reinforcement Learning Infrastructure
Redwood City or United States or London or Europe
$188k-$395k/yrHybridFull Time
Luma AI: Generative AI for high-quality video and 3D modeling.
Hands-on experience post-training LLMs with RL at scale, distributed PyTorch expertise, building RL environments/reward/verifier systems, and strong GPU cluster and networking knowledge.
Machine Learning Engineer, E-commerce Recommendation Foundation - USDS
San Jose, California, United States
$137k-$360k/yrOnsiteFull Time
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
Strong ML and deep learning foundation, proficiency in Python and PyTorch, experience with large-scale recommendation or model training, research mindset for LLMs/multimodal/reinforcement learning.
Lawrence Livermore National Laboratory: Develops science and technology for United States national security.
Ability to develop ML and LLM-centric software for HPC and multi-modal sensor data; Python/C++/Java proficiency; familiarity with reinforcement learning, agentic AI, GNNs, MCP; ability to obtain DOE Q clearance; BS required, MS preferred.
Python, C++, JAVA, Flux, SLURM, Large Language Model (LLM), Model-Context-Protocol (MCP)
Staff Robotics Engineer / Tech Lead – Whole-Body Control & Robot Learning
Santa Clara or Mountain View
$215k-$364k/yrOnsiteFull Time
XPengNew York Stock Exchange: XPEV: Designs and manufactures smart electric vehicles and autonomous technology.
5+ YOEPhD or Master's in robotics/engineering; 5+ years in reinforcement learning, robotics, or related areas; strong control, learning, and imitation learning skills; cross-disciplinary communication and leadership experience favored.
Clera: AI talent agent matching professionals with high-growth startup roles
3+ YOE3+ years ML/robotics experience, proficiency in Python and C++, PyTorch or TensorFlow, ROS/ROS2, simulation experience, reinforcement learning familiarity, and real-time/embedded deployment experience.
LTA Research & Exploration: Develops and builds next-generation zero-emission rigid airships.
Degree in robotics/mechatronics/engineering, expertise in humanoid kinematics and dynamic control, proficiency in C++ and Python, experience with ROS 2 and simulation, background in ML/Reinforcement Learning preferred, background check and drug test required.
C++, Python, ROS 2, Gazebo, NVIDIA Isaac Sim, SST, TTS