251 reinforcement learning jobs at 152 companies in United States

PromotedHiringCafe
Founding Machine Learning / AI Search Engineer
Cupertino, CA, US
$160k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
Python, PyTorch, Elasticsearch, LLMs
1w
Save
Mark Applied
Hide
Reinforcement Learning Engineer (Cybersecurity)
United States or San Francisco or New Hampshire
$176k-$243k/yr RemoteFull Time
Bugcrowd
Bugcrowd: Provides a crowdsourced platform for security vulnerability testing.
Experience with reinforcement learning workflows, Linux ML environments, vulnerability research/binary exploitation, proficiency in Python and C, DevOps pipelines and reproducible builds, and low-level debugging.
Mayhem, GitHub Actions, docker, buildkit, nix, Python, C, Rust, Linux
1mo
Save
Mark Applied
Hide
Reinforcement Learning Engineer, Grasping
Houston, Texas, United States
OnsiteFull Time
Persona AI
Persona AI: Developing rugged humanoid robots for industrial labor automation.
2+ YOEBS/MS/PhD in Robotics/CS/ML; 2+ years RL experience for robotic manipulation (or exceptional recent grads); proficiency in Python, PyTorch, JAX; experience with MuJoCo/Isaac Sim, sim-to-real, reward shaping, and RL libraries (rsl_rl, skrl).
Python, PyTorch, JAX, rsl_rl, skrl, MuJoCo, Isaac Lab, Isaac Sim
3mo
Save
Mark Applied
Hide
Reinforcement Learning Engineer - Locomanipulation
Cambridge or Boston
$200k-$350k/yr OnsiteFull Time
Humanoid
Humanoid: Developing commercial humanoid robots for industrial applications.
MS or PhD in Robotics, ML, CS or related field; strong RL experience; robotics experience; Python/C++ skills; deploy RL on real robots.
Python, C++, MuJoCo, Isaac Lab, ROS
2mo
Save
Mark Applied
Hide
(Senior) Machine Learning Engineer - Reinforcement Learning
Fremont, California, United States
$150k-$250k/yr OnsiteFull Time
Pony.ai
Pony.aiNASDAQ: PONY: Develops autonomous driving technology and operates robotaxi services.
3+ YOEM.S./Ph.D. or equivalent experience; 3+ years building production ML with strong RL experience; depth in deep learning and generative models; distributed training and large-scale data processing; strong communication.
Python, PyTorch, LLM, VLM, dLLM, VLA
2mo
Save
Mark Applied
Hide
Reinforcement Learning Engineer ($400k - $800k salary)
New York or London
$400k-$800k/yr OnsiteFull Time
Baton Corporation
Baton Corporation: Development building and operating the pump.fun blockchain platform.
Experience deploying autonomous learning systems with real capital; design risk limits; build evaluation loops; lead ML system end-to-end.
Python, ML frameworks, Simulation tools
3w
Save
Mark Applied
Hide
Research Scientist, Reinforcement Learning (LLM) and Post-training
Santa Clara, California, United States
$179k-$306k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Expertise in reinforcement learning theory and practice for post-training LLMs; strong publication record; hands-on experience training RL/preference-optimized models at scale; PhD strongly preferred.
3mo
Save
Mark Applied
Hide
Anthropic Fellows Program — Reinforcement Learning
London or Berkeley or San Francisco or London or Ontario or British Columbia or Remote-Friendly US (Travel Required)
$4k/wk HybridFull Time
Anthropic
Anthropic: Developing safe and reliable artificial intelligence systems.
Fluent in Python; available to work full-time; strong software engineering; interest in empirical AI research.
Python, Distributed Systems, Machine Learning, Research Tools
2d
Save
Mark Applied
Hide
Reinforcement Learning Infrastructure Engineer
Palo Alto, California, United States
$275k-$475k/yr OnsiteFull Time
Elorian AI
Elorian AI: AI lab building multimodal models for advanced visual reasoning.
3+ YOE3+ years distributed systems experience, strong Python and PyTorch or JAX, multi-node GPU orchestration (Ray, SLURM, Kubernetes), experience with actor-learner architectures and RL training pipelines.
Python, PyTorch, JAX, Ray, SLURM, Kubernetes
3d
Save
Mark Applied
Hide
Reinforcement Learning AI Engineer
Huntsville, Alabama, United States
$99k-$225k/yr HybridFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Consulting and technology services for government and commercial clients
3+ YOE3+ years developing and training RL agents, experience with Gym/PettingZoo, PyTorch/TensorFlow/JAX, Python/Rust/C++, secret clearance eligibility, bachelor's in CS/AI/Engineering, up to 25% travel.
Python, Gym, PettingZoo, PyTorch, TensorFlow, JAX, Rust, C++, C, AFSIM, CUDA, RAPID, Kubernetes
3d
Save
Mark Applied
Hide
Reinforcement Learning AI Engineer
Huntsville or McLean or El Segundo or Colorado Springs
$99k-$225k/yr OnsiteFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Provides technology and management consulting services to diverse organizations.
3+ YOE3+ years developing and training RL agents; experience with Gym/PettingZoo, PyTorch/TensorFlow/JAX, Python and optionally C++/Rust; ability to obtain/hold Secret clearance and travel up to 25%.
Python, Gym, PettingZoo, PyTorch, TensorFlow, JAX, Rust, C++, C, CUDA, RAPID, Kubernetes, AFSIM
1mo
Save
Mark Applied
Hide
Research Scientist, Reinforcement Learning
Fremont, California, United States
OnsiteFull Time
DeepRoute.ai
DeepRoute.ai: Develops full-stack autonomous driving software and VLA models.
Proficiency in modern RL and RLHF algorithms, experience with reward model training and LLM/VLM fine-tuning, distributed RL training and massively parallel simulation, sim-to-real transfer, Python and C++, PyTorch, and distributed training frameworks.
DQN, PPO, SAC, TD3, DPO, GRPO, LLM, VLM, VLA, Python, C++, PyTorch, Ray, Horovod, CUDA
3mo
Save
Mark Applied
Hide
Staff Reinforcement Learning Engineer – Whole Body Control
San Jose, California, United States
$150k-$250k/yr OnsiteFull Time
Figure
Figure: Develops autonomous humanoid robots for commercial and residential tasks.
Strong dynamics/control background for legged robots; RL for robotics (PPO, SAC); tune hyperparameters; apply domain randomization and curriculum learning; lead projects and mentor.
Python, PyTorch, ROS
2mo
Save
Mark Applied
Hide
Staff Reinforcement Learning Research Engineer
Waltham, Massachusetts, United States
$155k-$200k/yr OnsiteFull Time
Boston Dynamics
Boston Dynamics: Building advanced mobile robots for industrial and warehouse automation.
3+ YOEMS with 3+ years or PhD in ML, Robotics; deployed policies on physical robots; RL toolbox and simulation expertise; PyTorch/JAX; software fundamentals
Reinforcement Learning Toolboxes (RSL-RL, CleanRL, RLlib, Stable Baselines), Simulation and Rendering Tools (Isaac Lab, MuJoCo, MjWarp, MjLab), PyTorch, JAX, ONNX, Triton, TensorRT, Bazel, Docker, CI/CD
2w
Save
Mark Applied
Hide
Reinforcement Learning Engineer - Ingénieur(e) en apprentissage par renforcement
Montréal or New York
HybridFull Time
NBCUniversal
NBCUniversalNASDAQ: CMCSA: Produces and distributes entertainment content and theme park experiences.
Graduate degree in robotics/CS/AI, proven RL engineering experience, strong math, fluency in Python/Git/Unix, experience with RL frameworks and simulators, and sim-to-real expertise.
Python, Git, Unix shell, Ray RLlib, Stable Baselines3, CleanRL, MuJoCo, Bullet, Unity, Unreal, Isaac Sim, Jira, Confluence, Slack
1mo
Save
Mark Applied
Hide
Machine Learning Engineer - Reinforcement Learning
Fremont, California, United States
$150k-$250k/yr OnsiteFull Time
Pony.ai
Pony.aiNASDAQ: PONY: Develops autonomous driving technology for passenger and freight transportation.
MS/PhD in CS/ML/AI or related field, or equivalent experience; RL/production ML experience; deep learning, sequence modeling, generative models; publish or ship impactful ML systems; large-scale training and data processing; lead ambiguous work.
Python, PyTorch, Reinforcement Learning
1w
Save
Mark Applied
Hide
Member of Technical Staff – Senior Engineer, Reinforcement Learning for Wholebody Control
Cambridge, Massachusetts, United States
$180k-$240k/yr OnsiteFull Time
Walden Robotics
Walden Robotics: A developing general-purpose robots and related control and simulation systems to improve quality of life.
Proven RL experience for continuous control, sim-to-real transfer on real robots, control theory knowledge, large-scale GPU simulation, strong Python and C++ skills, and software engineering for training/deployment pipelines.
Python, C++
1mo
Save
Mark Applied
Hide
Machine Learning Scientist, Reinforcement Learning
Emeryville or Hybrid (2-3 days on-site)
$200k-$330k/yr HybridFull Time
Profluent
Profluent: Designs functional proteins using deep generative AI models.
PhD or equivalent in CS/ML/NLP/Applied Math/Computational Biology/Statistics; experience with ML/RL; publications; PyTorch or Jax experience.
PyTorch, Jax
3mo
Save
Mark Applied
Hide
Applied Scientist II, Reinforcement Learning
North Reading, Massachusetts, United States
$143k-$193k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
2+ YOEPhD or Masters with 2+ years of applied robotics research; expertise in reinforcement learning, imitation learning, whole-body control, and model-based control.
Python, C++, Java, Matlab, ROS, Mujoco, Drake, IsaacLab, Deep Learning
1w
Save
Mark Applied
Hide
Senior Machine Learning Engineer - Reinforcement Learning
Columbus, Ohio, United States
HybridFull Time
Path Robotics
Path Robotics: Developing autonomous robotic welding systems for manufacturing.
Master’s/PhD or equivalent experience in CS/Robotics/ML, strong RL experience for real-world systems, proficiency in Python and PyTorch/TensorFlow, simulation and sim-to-real experience, production ML deployment skills.
Python, PyTorch, TensorFlow, Isaac Gym, Gazebo, MuJoCo, PyBullet
2w
Save
Mark Applied
Hide
Research Engineer – Reinforcement Learning (RL) Systems & Infrastructure (Seed Infra)
San Jose, California, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
Design and build scalable RL systems and infrastructure for large-scale model training; expertise in distributed systems, GPU optimization, Python/C++, and RL workflows.
Python, C++, PyTorch