Kevin Black
Papers
10
Total Citations
359
H-Index
8
About
Kevin Black is a leading researcher at the intersection of robotics, computer vision, and machine learning, whose work is driving the frontier of general-purpose robot control. His primary research areas include vision-language-action (VLA) models, large-scale robot learning datasets, and foundation models for manipulation and navigation. Black’s most impactful contribution is the development of π₀, a groundbreaking vision-language-action flow model for general robot control, which has already garnered 127 citations since its 2025 release. He is also the driving force behind the DROID dataset (108 citations), a massive in-the-wild robot manipulation dataset that has become a critical resource for the community. His work on Octo, an open-source generalist robot policy, and ViNT, a foundation model for visual navigation, has further cemented his reputation for creating scalable, reusable robotic systems. Notably, Black’s research emphasizes open-source tools and large-scale data—such as BridgeData V2—to democratize robot learning. With over 350 total citations across his top papers, Kevin Black is shaping the future of how robots learn to interact with the physical world, making his profile essential reading for anyone interested in the next generation of intelligent, generalist robots.
Research Focus
Key Achievements
Top Papers
- 1π₀: A Vision-Language-Action Flow Model for General Robot Control127 citations · 2025
- 2DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset108 citations · 2024
- 3Octo: An Open-Source Generalist Robot Policy66 citations · 2024
- 4ViNT: A Foundation Model for Visual Navigation15 citations · 2023
- 5BridgeData V2: A Dataset for Robot Learning at Scale12 citations · 2023
- 6
- 7Octo: An Open-Source Generalist Robot Policy8 citations · 2024
- 8$π_0$: A Vision-Language-Action Flow Model for General Robot Control8 citations · 2024
- 9DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset3 citations · 2024
- 10$π_{0.5}$: a Vision-Language-Action Model with Open-World Generalization2 citations · 2025