Research Scientist, Reinforcement Learning
Data ScientistData
0 views0 saves0 applied
Quick Summary
Requirements Summary
DQN, PPO, SAC, TD3, etc. Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc.
Technical Tools
Data ScientistData
We are building next-generation end-to-end autonomous driving systems powered by reinforcement learning.
You will work on applying RL in closed-loop, safety-critical environments, leveraging large-scale simulation and real-world driving data to improve safety, comfort, and robustness.
- Train and deploy RL policies in closed-loop driving environments
- Scale RL training using massively parallel simulation systems
- Design and optimize reward functions for complex driving behaviors
- Improve sim-to-real transfer for real-world robustness
- Collaborate with cross-functional teams to integrate models into production systems
Requirements
~1 min read- Publications in top-tier venues (ICML, NeurIPS, ICLR, CVPR, ICCV, ECCV, ICRA, IROS, etc.)
- Open-source contributions to RL libraries or autonomous driving projects
- Previous experience with LLM fine-tuning using RLHF
- Knowledge of safe RL, interpretable AI, or robustness techniques
- Familiarity with autonomous vehicle regulations and safety standards
- Proficiency in modern RL algorithms: DQN, PPO, SAC, TD3, etc.
- Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc.
- Hands-on experience training reward models and finetuning LLM/VLM/VLA
- Knowledge of distributed RL training at scale
- Proficiency with massively parallel simulation environments
- Knowledge of sim-to-real transfer techniques and domain randomization
- Proficiency in Python, comfortable with C++
- Proficiency in deep learning frameworks such as PyTorch
- Experience with distributed training frameworks (Ray, Horovod, etc.)
- Knowledge of model optimization (quantization, pruning) and CUDA is a plus
- Knowledge of traffic rules, driving behavior modeling
Location & Eligibility
Where is the job
Fremont, United States
On-site at the office
Listing Details
- Posted
- June 6, 2026
- First seen
- September 29, 2026
- Last seen
- September 29, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 13%
- Scored at
- September 29, 2026
Signal breakdown
freshnesssource trustcontent trustemployer trust
External application
Similar Data Scientist jobs
View all →Senior Robotics Data Engineer/Data Scientist
Radiochemist – Research Scientist / Tenure Track Faculty Position
Assistant Professor/Research Scientist Computer Science/Cybersecurity (with AI/ML/Data Science Focus)
Assistant Professor of Computer Science/Cybersecurity-Research Scientist
Remote
Data Scientist III
Full-TimeRemote
Remote
Data Scientist II
Full-TimeRemote
Newsletter
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
A
B
C
D
No spam. Unsubscribe at any time.