Papers

2

Total Citations

7

H-Index

2

About

Bo Zhao is an emerging researcher working at the intersection of robotics, computer vision, and multimodal artificial intelligence. His work spans two compelling frontiers: the development of intelligent robotic systems capable of real-world object recognition, and the advancement of large multimodal models that unify learning across diverse data types. His 2025 paper on a robotic moving target system demonstrates a practical commitment to bridging perception and action in dynamic environments, earning early citations that signal growing interest from the robotics community. More ambitiously, his 2026 contribution tackles one of artificial intelligence's most fundamental challenges — extending next-token prediction, the engine behind transformative large language models, into truly multimodal domains encompassing text, images, and video. This work represents a significant step toward unified generative models capable of both understanding and producing content across modalities. Though early in citation accumulation, the foundational nature of these contributions positions Zhao as a researcher to watch in the rapidly evolving landscape of embodied AI and multimodal learning. Students exploring robotics-AI integration or generative model architectures will find his work particularly relevant and forward-looking.

Research Focus

Key Achievements

2
H-Index
2
Papers
7
Total Citations
4
Avg Citations/Paper
🏆 Most Cited Paper
Development and implementation of a robotic moving target system for object recognition testing
4 citations · 2025
📈 Most Prolific Year: 2025 (1 Papers)
🤝 Key Collaborators: 26
🏛 Institutions: Xi'an Jiaotong University, Beijing Academy of Artificial Intelligence

Top Papers

  1. 1
  2. 2

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 14 days ago