Research Scientist, Reinforcement Learning

Apply Now โ†—

About this role

We are building next-generation end-to-end autonomous driving systems powered by reinforcement learning.

You will work on applying RL in closed-loop, safety-critical environments, leveraging large-scale simulation and real-world driving data to improve safety, comfort, and robustness.

  • Train and deploy RL policies in closed-loop driving environments
  • Scale RL training using massively parallel simulation systems
  • Design and optimize reward functions for complex driving behaviors
  • Improve sim-to-real transfer for real-world robustness
  • Collaborate with cross-functional teams to integrate models into production systems

Core Technical Skills

  • Proficiency in modern RL algorithms: DQN, PPO, SAC, TD3, etc.
  • Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc.
  • Hands-on experience training reward models and finetuning LLM/VLM/VLA
  • Knowledge of distributed RL training at scale
  • Proficiency with massively parallel simulation environments
  • Knowledge of sim-to-real transfer techniques and domain randomization
  • Proficiency in Python, comfortable with C++
  • Proficiency in deep learning frameworks such as PyTorch
  • Experience with distributed training frameworks (Ray, Horovod, etc.)
  • Knowledge of model optimization (quantization, pruning) and CUDA is a plus
  • Knowledge of traffic rules, driving behavior modeling

Preferred Qualifications

  • Publications in top-tier venues (ICML, NeurIPS, ICLR, CVPR, ICCV, ECCV, ICRA, IROS, etc.)
  • Open-source contributions to RL libraries or autonomous driving projects
  • Previous experience with LLM fine-tuning using RLHF
  • Knowledge of safe RL, interpretable AI, or robustness techniques
  • Familiarity with autonomous vehicle regulations and safety standards

Frequently Asked Questions

Is the salary disclosed for the Research Scientist, Reinforcement Learning position at Deeproute.ai?
The salary for this Research Scientist, Reinforcement Learning role at Deeproute.ai is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Where is the Research Scientist, Reinforcement Learning position at Deeproute.ai located?
This Research Scientist, Reinforcement Learning role at Deeproute.ai is based in Fremont, California, United States. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
How do I apply for the Research Scientist, Reinforcement Learning position at Deeproute.ai?
Click the "Apply Now" button on this page. You will be redirected to Deeproute.ai's official application portal hosted on workable where you can submit your application directly.
When was the Research Scientist, Reinforcement Learning job at Deeproute.ai posted?
This Research Scientist, Reinforcement Learning position at Deeproute.ai was posted on Jun 6, 2026. Apply as soon as possible โ€” early applications are often reviewed first.
Research Scientist, Reinforcement Learning
Deeproute.ai
Apply for this role โ†—

You'll be redirected to Deeproute.ai's official application page on workable.