69 reinforcement learning engineer jobs at 37 companies in Concord, CA
3w
Save
Mark Applied
Hide
3w
Reinforcement Learning Engineer (Cybersecurity)
United States or San Francisco or New Hampshire
$176k-$243k/yrRemoteFull Time
Bugcrowd: Provides a crowdsourced platform for security vulnerability testing.
Experience with reinforcement learning workflows, Linux ML environments, vulnerability research/binary exploitation, proficiency in Python and C, DevOps pipelines and reproducible builds, and low-level debugging.
Mayhem, GitHub Actions, docker, buildkit, nix, Python, C, Rust, Linux
Pony.aiNASDAQ: PONY: Develops autonomous driving technology and operates robotaxi services.
3+ YOEM.S./Ph.D. or equivalent experience; 3+ years building production ML with strong RL experience; depth in deep learning and generative models; distributed training and large-scale data processing; strong communication.
Elorian AI: AI lab building multimodal models for advanced visual reasoning.
3+ YOE3+ years distributed systems experience, strong Python and PyTorch or JAX, multi-node GPU orchestration (Ray, SLURM, Kubernetes), experience with actor-learner architectures and RL training pipelines.
XPengNew York Stock Exchange: XPEV: Designs and manufactures smart electric vehicles and autonomous technology.
1+ YOEAdvanced degree preferred but open to fresh graduates; proficiency in Python; 1+ years experience with deep learning frameworks such as PyTorch; strong RL and imitation learning knowledge; experience with PPO/DQN/SAC; C++ and hardware experience preferred.
Lead AI Infrastructure Engineer, Reinforcement Learning
Santa Clara, California, United States
$179k-$306k/yrHybridFull Time
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Design and operate distributed RL training infrastructure at scale; strong systems experience in ML platforms; proficiency with PyTorch/JAX, NCCL/MPI-style distributed training, C++/Python performance tuning; Bachelor's degree required.
Member of Technical Staff – Senior Engineer, Reinforcement Learning – Policy Post-Training
Cambridge or San Francisco
$255k-$340k/yrOnsiteFull Time
Walden Robotics: Builds general-purpose robots and advances robot manipulation through research and development.
Hands-on RL experience for manipulation, sim-to-real transfer, reward and curriculum design, strong software engineering, and ability to run large-scale training and experiments.
Pony.aiNASDAQ: PONY: Develops autonomous driving technology for passenger and freight transportation.
MS/PhD in CS/ML/AI or related field, or equivalent experience; RL/production ML experience; deep learning, sequence modeling, generative models; publish or ship impactful ML systems; large-scale training and data processing; lead ambiguous work.
Research Scientist / Engineer – Reinforcement Learning Infrastructure
Redwood City, California, United States
HybridFull Time
Luma AI: Develops multimodal AI for video generation and creative production.
Experience operating post-trained LLMs with reinforcement learning at scale, distributed PyTorch training, building RL environments and reward infrastructure, and debugging large asynchronous rollout pipelines.
Research Engineer – Reinforcement Learning (RL) Systems & Infrastructure (Seed Infra)
San Jose, California, United States
OnsiteFull Time
ByteDance: Developing AI-driven content platforms and mobile applications.
Design and build scalable RL systems and infrastructure for large-scale model training; expertise in distributed systems, GPU optimization, Python/C++, and RL workflows.
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOEBachelor's in CS (or equivalent), 5+ years relevant experience (including 3 years engineering), strong communication/presentation skills, RL and LLM pipeline experience, and proficiency with PyTorch, JAX, and NeMo-RL.
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOEBachelor's in CS or related field; 5+ years exp including 3+ in engineering; strong communication; experience presenting to technical audiences; RL pipelines; PyTorch/JAX/NeMo-RL; able to work on training pipelines and demos.
Mariana Minerals: Building software-first infrastructure to produce and refine critical minerals.
0+ YOE0–4 years ML or scientific computing experience; strong ML fundamentals and deep learning exposure; reinforcement learning a plus; proficiency in Python; ability to debug codebases and collaborate with process/chemistry experts.
UnityNYSE: U: Provides software for creating real-time 3D interactive content.
Expertise in machine learning, reinforcement learning, control theory, large-scale data analysis, online experimentation; strong software engineering in Python and deep learning (PyTorch).
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
5+ YOERequires 5+ years building production ML systems, Python, PyTorch, Spark, Airflow, modern ML infrastructure, and expertise in deep learning, reinforcement learning, optimization, LLMs, or VLMs.
PyTorch, Spark, Airflow, Python, Claude Code, Codex, Cursor, large language models (LLMs), vision-language models (VLMs)
Research Engineer, Chip Design RL (Reinforcement Learning)
San Francisco or New York City
$500k-$850k/yrHybridFull Time
Anthropic: Developing safe and reliable artificial intelligence systems.
Bachelor's or equivalent, expertise in ASIC/FPGA design and EDA tools, experience with RTL, verification (UVM, formal methods), physical design and tapeout experience; RL experience and tooling experience preferred.
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
Master's or PhD in CS/robotics/ML, deep reinforcement learning experience, strong C++ and Python skills, hands-on with JAX or PyTorch, experience bridging sim-to-real fidelity gaps and working with GPU/CUDA environments.
LTA Research & Exploration: Develops and builds next-generation zero-emission rigid airships.
Degree in robotics/mechatronics/engineering, expertise in humanoid kinematics and dynamic control, proficiency in C++ and Python, experience with ROS 2 and simulation, background in ML/Reinforcement Learning preferred, background check and drug test required.
C++, Python, ROS 2, Gazebo, NVIDIA Isaac Sim, SST, TTS
Imagry: Mapless autonomous driving software for vehicle manufacturers.
3+ YOEM.Sc. (or equivalent) in a quantitative field, 3+ years algorithm engineering experience, expertise in motion planning/decision making/optimization/Reinforcement Learning or LLM fine-tuning, and proficiency in Python and C++.
Eightfold.ai: AI-native platform for talent management and workforce optimization.
2+ YOEExperience in machine learning, LLMs, agentic AI, reinforcement learning; proficiency with Python, TensorFlow/PyTorch, AWS, Docker, Kubernetes; 2+ years relevant ML experience; strong coding, algorithms, and communication skills.