D
Deeproute.ai
Research Scientist

Research Scientist, Reinforcement Learning

On-siteSeniorResearch Scientistposted 2mo ago
Role summaryAI-generated

Deeproute.ai seeks a research scientist to develop and deploy reinforcement learning policies for autonomous driving, focusing on closed-loop safety-critical simulations and sim-to-real transfer. The role involves scaling training on massive parallel simulators, designing reward functions, and integrating models into production systems.

Skills required

About this role

We are building next-generation end-to-end autonomous driving systems powered by reinforcement learning.

You will work on applying RL in closed-loop, safety-critical environments, leveraging large-scale simulation and real-world driving data to improve safety, comfort, and robustness.

  • Train and deploy RL policies in closed-loop driving environments
  • Scale RL training using massively parallel simulation systems
  • Design and optimize reward functions for complex driving behaviors
  • Improve sim-to-real transfer for real-world robustness
  • Collaborate with cross-functional teams to integrate models into production systems

Requirements

Core Technical Skills

  • Proficiency in modern RL algorithms: DQN, PPO, SAC, TD3, etc.
  • Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc.
  • Hands-on experience training reward models and finetuning LLM/VLM/VLA
  • Knowledge of distributed RL training at scale
  • Proficiency with massively parallel simulation environments
  • Knowledge of sim-to-real transfer techniques and domain randomization
  • Proficiency in Python, comfortable with C++
  • Proficiency in deep learning frameworks such as PyTorch
  • Experience with distributed training frameworks (Ray, Horovod, etc.)
  • Knowledge of model optimization (quantization, pruning) and CUDA is a plus
  • Knowledge of traffic rules, driving behavior modeling

Preferred Qualifications

  • Publications in top-tier venues (ICML, NeurIPS, ICLR, CVPR, ICCV, ECCV, ICRA, IROS, etc.)
  • Open-source contributions to RL libraries or autonomous driving projects
  • Previous experience with LLM fine-tuning using RLHF
  • Knowledge of safe RL, interpretable AI, or robustness techniques
  • Familiarity with autonomous vehicle regulations and safety standards
✕ position closed

This role is no longer accepting applications. It’s kept here for reference — check out the similar open roles below.

Similar open roles

G
NEW

Senior Applied Research Scientist

GEICO·Bethesda, MD; New York City, NY; Palo Alto, CA
On-siteSeniorResearch Scientist
$115k – $230k USD
2d ago
TS
NEW

Senior Applied Machine Learning Scientist, Predictive AI(B3617)

TD Securities·Toronto, Ontario; Montréal, Québec
On-siteSeniorResearch Scientist
$157k – $190k USD
2d ago
SS
NEW

Quantitative Investment Researcher (Assistant Vice President)

State Street·Cambridge, Massachusetts
On-siteSeniorResearch Scientist
$175k USD
2d ago
Z
NEW

Staff Machine Learning Scientist

Zendesk·Tallinn, Estonia; Krakow, Poland
On-siteStaffResearch Scientist
2d ago
H

Research Scientist for Power Electronics

Hitachi·Vaesteras, Vastmanland County, Sweden
On-siteSeniorResearch Scientist
5d ago