27 reinforcement learning engineer jobs at 15 companies in Cotati, CA

3w
Save
Mark Applied
Hide
Reinforcement Learning Engineer (Cybersecurity)
United States or San Francisco or New Hampshire
$176k-$243k/yr RemoteFull Time
Bugcrowd
Bugcrowd: Provides a crowdsourced platform for security vulnerability testing.
Experience with reinforcement learning workflows, Linux ML environments, vulnerability research/binary exploitation, proficiency in Python and C, DevOps pipelines and reproducible builds, and low-level debugging.
Mayhem, GitHub Actions, docker, buildkit, nix, Python, C, Rust, Linux
3w
Save
Mark Applied
Hide
Member of Technical Staff – Senior Engineer, Reinforcement Learning – Policy Post-Training
Cambridge or San Francisco
$255k-$340k/yr OnsiteFull Time
Walden Robotics
Walden Robotics: Builds general-purpose robots and advances robot manipulation through research and development.
Hands-on RL experience for manipulation, sim-to-real transfer, reward and curriculum design, strong software engineering, and ability to run large-scale training and experiments.
Isaac, MuJoCo
2mo
Save
Mark Applied
Hide
Machine Learning Engineer
Ann Arbor or San Francisco or Houston
$120k-$160k/yr OnsiteFull Time
Mariana Minerals
Mariana Minerals: Building software-first infrastructure to produce and refine critical minerals.
0+ YOE0–4 years ML or scientific computing experience; strong ML fundamentals and deep learning exposure; reinforcement learning a plus; proficiency in Python; ability to debug codebases and collaborate with process/chemistry experts.
Python
3w
Save
Mark Applied
Hide
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco or New York City
$500k-$850k/yr HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
Bachelor's or equivalent, expertise in ASIC/FPGA design and EDA tools, experience with RTL, verification (UVM, formal methods), physical design and tapeout experience; RL experience and tooling experience preferred.
EDA tools, UVM, formal methods, place-and-route
3d
Save
Mark Applied
Hide
Machine Learning Engineer, Drive
San Francisco or Sunnyvale or Seattle
$137k-$202k/yr OnsiteFull Time
DoorDash
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOERequires 5+ years building production ML systems, Python, PyTorch, Spark, Airflow, modern ML infrastructure, and expertise in deep learning, reinforcement learning, optimization, LLMs, or VLMs.
PyTorch, Spark, Airflow, Python, Claude Code, Codex, Cursor, large language models (LLMs), vision-language models (VLMs)
1w
Save
Mark Applied
Hide
Founding AI Engineer
San Francisco, California, United States
$150k-$200k/yr OnsiteFull Time
Clera
Clera: AI talent agent matching professionals with high-growth startup roles
Eligible to work in the US without sponsorship, up to 4 years building production AI/ML systems, experience with matching/personalization, reinforcement learning, multi-channel agent orchestration, and cloud infrastructure.
Instagram, iMessage, WhatsApp, TikTok, AWS, GCP, Azure
6d
Save
Mark Applied
Hide
Applied ML Engineer
San Francisco, California, United States
$170k-$280k/yr OnsiteFull Time
Macroscope
Macroscope: Provides AI-powered codebase analysis and automated code reviews.
3+ YOE3+ years applied ML experience; experience building, training, fine-tuning, or evaluating ML models; dataset creation and evaluation; experiment design; familiarity with LLMs and reinforcement learning techniques.
TypeScript, React, Golang, Temporal, Google Cloud (GCP), Postgres, Terraform, Swift, Python, Rust
1mo
Save
Mark Applied
Hide
Research Engineer / Research Scientist - RE / RS - Proactivity
San Francisco, California, United States
$295k-$555k/yr HybridFull Time
OpenAI
OpenAI: Develops artificial intelligence models and generative AI software services.
Strong ML engineering and research experience with LLM post-training, reinforcement learning, dataset creation, evaluations, and ability to work in a large ML codebase.
2d
Save
Mark Applied
Hide
Senior ML Engineer, Perception Research
Mountain View or New York City or Kirkland or San Francisco
$213k-$263k/yr HybridFull Time
Waymo
Waymo: Autonomous driving technology for ride-hailing and logistics.
2+ YOEMaster's or PhD in a technical field plus 2+ years of industry or postdoctoral research experience in reinforcement learning or foundation models; scalable distributed model training experience required.
Multimodal LLMs, World Models, Camera, LiDAR, Radar, Transformer, Data Parallel, FSDP, Diffusion, Autoregressive Models
1mo
Save
Mark Applied
Hide
Staff Research Engineer, Data Agents
San Francisco, California, United States
$190k-$270k/yr OnsiteFull Time
Databricks
Databricks: A unified platform for data analytics and artificial intelligence.
2+ YOEBS/MS/PhD in CS or related; 2+ years applied research with shipped prototypes; experience with LLMs, agents, reinforcement learning, and post-training workflows; strong communication and cross-functional collaboration.
Apache Spark, Delta Lake, MLflow, Genie
2mo
Save
Mark Applied
Hide
Founding Engineer, Applied Research
San Francisco, California, United States
OnsiteFull Time
Backbone
Backbone: AI-powered infrastructure for healthcare payments and authorizations.
0+ YOEExperience with ML research, language models, NLP, evals, reinforcement learning, agentic systems, and model improvement; demonstrated technical depth via papers/projects/internships; new grads through ~5-6 years experience considered.
1mo
Save
Mark Applied
Hide
Member of Technical Staff, Platform Engineering
San Francisco, California, United States
OnsiteFull Time
Abundant
Abundant: Building simulation infrastructure for AI model training.
Deep technical fluency in evaluations, reinforcement learning, LLMs and agents; research-oriented with ability to read SOTA papers; highly proficient coding agents (e.g., Codex, Claude Code); ownership mindset and experience building scalable systems.
Codex, Claude Code
1mo
Save
Mark Applied
Hide
Member of Technical Staff, Post-Training, RL Environments
San Francisco, California, United States
$350k-$500k/yr OnsiteFull Time
Mirendil
Mirendil: Developing frontier artificial intelligence models to accelerate scientific research.
Build and own data systems and execution environments for reinforcement learning; experience with long-horizon RL tasks, scalable sandboxed environments, and preventing reward hacking; strong collaboration and system-design skills.
1w
Save
Mark Applied
Hide
Member of Technical Staff — Agent Post-Training
San Francisco, California, United States
OnsiteFull Time
Moonlake
Moonlake: Generates interactive 3D world simulations using artificial intelligence.
Experience training large language, vision-language, multimodal, or code models; strong reinforcement learning and post-training expertise; distributed training and high-throughput systems experience; Python and PyTorch/JAX proficiency.
Python, PyTorch, JAX
2w
Save
Mark Applied
Hide
Founding Research Scientist
San Francisco, California, United States
OnsiteFull Time
Xterra AI
Xterra AI: A stealth-stage AI building agents and foundation models to solve complex scientific problems.
Hands-on ML experience training large models, reinforcement learning (RLHF/RLAIF), reward modeling, and ability to work across research and engineering.