Why This Role Stands Out
Advance your career in autonomous driving by applying cutting-edge reinforcement learning to safety-critical systems, offering significant growth and impact. You'll thrive here if you possess strong RL expertise and a passion for developing innovative solutions in a collaborative, on-site environment.
Quick Overview
Seniority
Mid Senior
Work mode
On Site
Location
Fremont, California, United States
Posted
3 months ago
CUDADeep LearningAutonomous DrivingC++LLMPyTorchPython
Job Description
We are building next-generation end-to-end autonomous driving systems powered by reinforcement learning.
You will work on applying RL in closed-loop, safety-critical environments, leveraging large-scale simulation and real-world driving data to improve safety, comfort, and robustness.
- Train and deploy RL policies in closed-loop driving environments
- Scale RL training using massively parallel simulation systems
- Design and optimize reward functions for complex driving behaviors
- Improve sim-to-real transfer for real-world robustness
- Collaborate with cross-functional teams to integrate models into production systems
Core Technical Skills
- Proficiency in modern RL algorithms: DQN, PPO, SAC, TD3, etc.
- Proficiency in modern RLHF algorithms: PPO, DPO, GRPO, etc.
- Hands-on experience training reward models and finetuning LLM/VLM/VLA
- Knowledge of distributed RL training at scale
- Proficiency with massively parallel simulation environments
- Knowledge of sim-to-real transfer techniques and domain randomization
- Proficiency in Python, comfortable with C++
- Proficiency in deep learning frameworks such as PyTorch
- Experience with distributed training frameworks (Ray, Horovod, etc.)
- Knowledge of model optimization (quantization, pruning) and CUDA is a plus
- Knowledge of traffic rules, driving behavior modeling
Preferred Qualifications
- Publications in top-tier venues (ICML, NeurIPS, ICLR, CVPR, ICCV, ECCV, ICRA, IROS, etc.)
- Open-source contributions to RL libraries or autonomous driving projects
- Previous experience with LLM fine-tuning using RLHF
- Knowledge of safe RL, interpretable AI, or robustness techniques
- Familiarity with autonomous vehicle regulations and safety standards
Similar jobs
- ST
Food Scientist
NewStefanini
East Hanover, NJ🇺🇸On-site2 days agoComplianceTalent Acquisition - QU
AI Researcher
NewQualcomm
San Diego, California🇺🇸Hybrid4 hours ago5GMLOpsDeep Learning+4 - BS
Exploitation Specialist/Imagery Scientist SAR with Security Clearance
BTS Software Solutions
Springfield, VA🇺🇸$150k - $195k/yrHybrid1 week ago401kMachine Learning - AB
Member of Technical Staff, Research (Intern)
NewAbundant
San Francisco🇺🇸$10k/mo20 hours agoAWSLinearNLP+1 - IT
Research Scientist I/II, Purification Sciences
NewIambic Therapeutics
San Diego HQ🇺🇸18 hours ago401kClinical TrialsHTTP - 2K
Researcher
New2K
Frisco🇺🇸19 hours agoMicrosoft Excel