Papers
33
Total Citations
1,791
H-Index
15
About
Quan Vuong is a prominent robotics and machine learning researcher whose work sits at the cutting edge of embodied AI, robot learning, and large-scale foundation models for real-world control. His research focuses on developing transformer-based and vision-language-action (VLA) models that enable robots to generalize across diverse tasks by leveraging internet-scale data and multimodal learning. Vuong has been a key contributor to several landmark systems in the field. His involvement in RT-1 (512 citations) and RT-2 (267 citations) helped establish the paradigm of training large robotics transformers on broad, task-agnostic datasets to achieve robust real-world manipulation. His co-authorship on PaLM-E (350 citations) advanced grounded multimodal reasoning for embodied agents, while contributions to π₀, OpenVLA, and FAST reflect his continued leadership in making VLA models more capable, accessible, and efficient. The DROID dataset and Octo policy further demonstrate his commitment to open, large-scale infrastructure for the robotics community. Earlier work on constrained reinforcement learning (2020) shows his foundational grounding in safe and principled policy optimization. With over 1,500 cumulative citations and multiple highly influential papers published in just a few years, Vuong has emerged as a defining voice in the next generation of generalizable robotic intelligence.
Research Focus
Key Achievements
Top Papers
- 1RT-1: Robotics Transformer for Real-World Control at Scale512 citations · 2023
- 2PaLM-E: An Embodied Multimodal Language Model350 citations · 2023
- 3RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control267 citations · 2023
- 4π₀: A Vision-Language-Action Flow Model for General Robot Control127 citations · 2025
- 5DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset108 citations · 2024
- 6Octo: An Open-Source Generalist Robot Policy66 citations · 2024
- 7First Order Constrained Optimization in Policy Space52 citations · 2020
- 8OpenVLA: An Open-Source Vision-Language-Action Model39 citations · 2024
- 9RT-1: Robotics Transformer for Real-World Control at Scale38 citations · 2022
- 10FAST: Efficient Action Tokenization for Vision-Language-Action Models28 citations · 2025