34 reinforcement learning engineer jobs at 19 companies in Novato, CA

4w
Save
Mark Applied
Hide
Reinforcement Learning Engineer (Cybersecurity)
United States or San Francisco or New Hampshire
$176k-$243k/yr RemoteFull Time
Bugcrowd
Bugcrowd: Provides a crowdsourced platform for security vulnerability testing.
Experience with reinforcement learning workflows, Linux ML environments, vulnerability research/binary exploitation, proficiency in Python and C, DevOps pipelines and reproducible builds, and low-level debugging.
Mayhem, GitHub Actions, docker, buildkit, nix, Python, C, Rust, Linux
3mo
Save
Mark Applied
Hide
(Senior) Machine Learning Engineer - Reinforcement Learning
Fremont, California, United States
$150k-$250k/yr OnsiteFull Time
Pony.ai
Pony.aiNASDAQ: PONY: Develops autonomous driving technology and operates robotaxi services.
3+ YOEM.S./Ph.D. or equivalent experience; 3+ years building production ML with strong RL experience; depth in deep learning and generative models; distributed training and large-scale data processing; strong communication.
Python, PyTorch, LLM, VLM, dLLM, VLA
1w
Save
Mark Applied
Hide
Senior/Staff Deep Reinforcement Learning Engineer
San Francisco, California, United States
$168k-$247k/yr OnsiteFull Time
DoorDash
DoorDashNYSE: DASH: Local food delivery and on-demand logistics platform.
BS, MS, or PhD in CS, EE, Robotics, or related field; deep RL and deep learning expertise; large-scale RL training; JAX or similar framework; GPU simulation and software development experience.
JAX, Claude Code, Codex, Cursor, NeurIPS, ICML, ICLR, CoRL, RSS, ICRA
1mo
Save
Mark Applied
Hide
Member of Technical Staff – Senior Engineer, Reinforcement Learning – Policy Post-Training
Cambridge or San Francisco
$255k-$340k/yr OnsiteFull Time
Walden Robotics
Walden Robotics: Builds general-purpose robots and advances robot manipulation through research and development.
Hands-on RL experience for manipulation, sim-to-real transfer, reward and curriculum design, strong software engineering, and ability to run large-scale training and experiments.
Isaac, MuJoCo
2mo
Save
Mark Applied
Hide
Machine Learning Engineer - Reinforcement Learning
Fremont, California, United States
$150k-$250k/yr OnsiteFull Time
Pony.ai
Pony.aiNASDAQ: PONY: Develops autonomous driving technology for passenger and freight transportation.
MS/PhD in CS/ML/AI or related field, or equivalent experience; RL/production ML experience; deep learning, sequence modeling, generative models; publish or ship impactful ML systems; large-scale training and data processing; lead ambiguous work.
Python, PyTorch, Reinforcement Learning
2w
Save
Mark Applied
Hide
Research Scientist / Engineer – Reinforcement Learning Infrastructure
Redwood City, California, United States
HybridFull Time
Luma AI
Luma AI: Develops multimodal AI for video generation and creative production.
Experience operating post-trained LLMs with reinforcement learning at scale, distributed PyTorch training, building RL environments and reward infrastructure, and debugging large asynchronous rollout pipelines.
vLLM, SGLang, PyTorch, FSDP, Tensor Parallel, Pipeline Parallel, Expert Parallel, veRL, OpenRLHF, TRL, Ray, NCCL, MPI, Kubernetes
2mo
Save
Mark Applied
Hide
Machine Learning Engineer
Ann Arbor or San Francisco or Houston
$120k-$160k/yr OnsiteFull Time
Mariana Minerals
Mariana Minerals: Building software-first infrastructure to produce and refine critical minerals.
0+ YOE0–4 years ML or scientific computing experience; strong ML fundamentals and deep learning exposure; reinforcement learning a plus; proficiency in Python; ability to debug codebases and collaborate with process/chemistry experts.
Python
1mo
Save
Mark Applied
Hide
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco or New York City
$500k-$850k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
Bachelor's or equivalent, expertise in ASIC/FPGA design and EDA tools, experience with RTL, verification (UVM, formal methods), physical design and tapeout experience; RL experience and tooling experience preferred.
EDA tools, UVM, formal methods, place-and-route
1w
Save
Mark Applied
Hide
Machine Learning Engineer, Drive
San Francisco or Sunnyvale or Seattle
$137k-$202k/yr OnsiteFull Time
DoorDash
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOERequires 5+ years building production ML systems, Python, PyTorch, Spark, Airflow, modern ML infrastructure, and expertise in deep learning, reinforcement learning, optimization, LLMs, or VLMs.
PyTorch, Spark, Airflow, Python, Claude Code, Codex, Cursor, large language models (LLMs), vision-language models (VLMs)
1mo
Save
Mark Applied
Hide
Machine Learning Engineer - Simulation Framework
Foster City or Seattle
$151k-$257k/yr HybridFull Time
Zoox
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
Master's or PhD in CS/robotics/ML, deep reinforcement learning experience, strong C++ and Python skills, hands-on with JAX or PyTorch, experience bridging sim-to-real fidelity gaps and working with GPU/CUDA environments.
JAX, PyTorch, C++, Python, CUDA
1w
Save
Mark Applied
Hide
Applied ML Engineer
San Francisco, California, United States
$170k-$280k/yr OnsiteFull Time
Macroscope
Macroscope: Provides AI-powered codebase analysis and automated code reviews.
3+ YOE3+ years applied ML experience; experience building, training, fine-tuning, or evaluating ML models; dataset creation and evaluation; experiment design; familiarity with LLMs and reinforcement learning techniques.
TypeScript, React, Golang, Temporal, Google Cloud (GCP), Postgres, Terraform, Swift, Python, Rust
2mo
Save
Mark Applied
Hide
Senior Robotics/Physical AI Manipulation Engineer
Menlo Park, California, United States
HybridFull Time
Autonomique
Autonomique: Developing hardware-agnostic physical AI software for industrial robotics.
3+ YOEMS or PhD in Robotics/CS, 3+ years shipping code for physical robots, expertise in manipulation planning, imitation or reinforcement learning, production Python and C++ experience, familiarity with sim-to-real.
Python, C++, Isaac Lab, MuJoCo
1mo
Save
Mark Applied
Hide
Research Engineer / Research Scientist - RE / RS - Proactivity
San Francisco, California, United States
$295k-$555k/yr HybridFull Time
OpenAI
OpenAI: Develops artificial intelligence models and generative AI software services.
Strong ML engineering and research experience with LLM post-training, reinforcement learning, dataset creation, evaluations, and ability to work in a large ML codebase.
6d
Save
Mark Applied
Hide
Staff ML Engineer, Perception Research
Mountain View or New York City or Kirkland or San Francisco
$251k-$310k/yr HybridFull Time
Waymo
Waymo: Autonomous driving technology for ride-hailing and logistics.
4+ YOEMaster's or PhD in a technical field and 4+ years of industry or postdoctoral research experience in reinforcement learning or foundation models. Requires distributed training, JAX, Flax, and Transformer optimization expertise.
JAX, Flax, TensorFlow, PyTorch, FSDP, xprof, Gemax, XManager
1mo
Save
Mark Applied
Hide
Staff Research Engineer, Data Agents
San Francisco, California, United States
$190k-$270k/yr OnsiteFull Time
Databricks
Databricks: A unified platform for data analytics and artificial intelligence.
2+ YOEBS/MS/PhD in CS or related; 2+ years applied research with shipped prototypes; experience with LLMs, agents, reinforcement learning, and post-training workflows; strong communication and cross-functional collaboration.
Apache Spark, Delta Lake, MLflow, Genie
2mo
Save
Mark Applied
Hide
Founding Engineer, Applied Research
San Francisco, California, United States
OnsiteFull Time
Backbone
Backbone: AI-powered infrastructure for healthcare payments and authorizations.
0+ YOEExperience with ML research, language models, NLP, evals, reinforcement learning, agentic systems, and model improvement; demonstrated technical depth via papers/projects/internships; new grads through ~5-6 years experience considered.
1mo
Save
Mark Applied
Hide
Member of Technical Staff, Platform Engineering
San Francisco, California, United States
OnsiteFull Time
Abundant
Abundant: Building simulation infrastructure for AI model training.
Deep technical fluency in evaluations, reinforcement learning, LLMs and agents; research-oriented with ability to read SOTA papers; highly proficient coding agents (e.g., Codex, Claude Code); ownership mindset and experience building scalable systems.
Codex, Claude Code
1mo
Save
Mark Applied
Hide
Member of Technical Staff, Post-Training, RL Environments
San Francisco, California, United States
$350k-$500k/yr OnsiteFull Time
Mirendil
Mirendil: Developing frontier artificial intelligence models to accelerate scientific research.
Build and own data systems and execution environments for reinforcement learning; experience with long-horizon RL tasks, scalable sandboxed environments, and preventing reward hacking; strong collaboration and system-design skills.
2w
Save
Mark Applied
Hide
Member of Technical Staff — Agent Post-Training
San Francisco, California, United States
OnsiteFull Time
Moonlake
Moonlake: Generates interactive 3D world simulations using artificial intelligence.
Experience training large language, vision-language, multimodal, or code models; strong reinforcement learning and post-training expertise; distributed training and high-throughput systems experience; Python and PyTorch/JAX proficiency.
Python, PyTorch, JAX
3w
Save
Mark Applied
Hide
Founding Research Scientist
San Francisco, California, United States
OnsiteFull Time
Xterra AI
Xterra AI: A stealth-stage AI building agents and foundation models to solve complex scientific problems.
Hands-on ML experience training large models, reinforcement learning (RLHF/RLAIF), reward modeling, and ability to work across research and engineering.

Explore Jobs

Expand Your Job Search