Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Quick Summary
PlusAI is a Physical AI company pioneering AI-based virtual driver software for factory-built autonomous trucks. Headquartered in Silicon Valley with operations in the United States and Europe,
You will join our core AI team at the frontier of autonomous decision-making, building the Vision-Language-Action (VLA) models that form SuperDrive's reasoning layer. You'll train VLA models that generate high-level driving decisions and trajectory guidance for on-board strategic decision-making, and design the knowledge distillation and compression techniques that transition large models onto on-board compute.
- M.S. minimum, Ph.D. preferred in CS, EE, Mathematics, Statistics, or a related field.
- 3+ years implementing and training models in a deep learning framework (PyTorch, TensorFlow, or JAX).
- Direct, hands-on experience training vision-language / vision-language-action models.
- Hands-on experience with model training, evaluation, and deployment in production.
- Thorough understanding of state-of-the-art vision-language / VLA models, diffusion, flow matching, and transformers.
- Experience with large-scale / distributed model training.
Location & Eligibility
Listing Details
- Posted
- September 15, 2026
- First seen
- September 15, 2026
- Last seen
- October 6, 2026
Posting Health
- Days active
- 20
- Repost count
- 0
- Trust Level
- 44%
- Scored at
- October 6, 2026
Signal breakdown
Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.