Ajmal Mian
Papers
30
Total Citations
1,439
H-Index
14
About
Ajmal Mian is a prominent researcher whose work spans natural language processing, computer vision, and robotics, with particular expertise in large language models, video understanding, and 3D scene analysis. His most influential contribution, "A Comprehensive Overview of Large Language Models," has accumulated over 800 citations across its versions, establishing him as a leading authority on LLM architectures, capabilities, and emerging research directions — a testament to the field's rapid growth and the survey's value as a foundational reference. Mian's work on video description, including a widely cited 2018 survey and its expanded 2019 follow-up (collectively exceeding 240 citations), helped consolidate methodologies for automatic natural language generation from video content, with meaningful implications for accessibility and human-robot interaction. His contributions to 3D scene understanding are equally significant, encompassing deep learning-based 3D segmentation surveys, 3D scene graph prediction from RGB-D sequences, and point cloud-based place recognition — all critical capabilities for autonomous robotics and SLAM systems. His recent innovations in hyperrectangle embedding and history-enhanced scene graph reasoning reflect a commitment to pushing the boundaries of spatial reasoning and long-term robot autonomy. Across diverse domains, Mian's survey work and original research have made him an indispensable guide for students and practitioners navigating rapidly evolving fields.
Research Focus
Key Achievements
Top Papers
- 1A Comprehensive Overview of Large Language Models465 citations · 2025
- 2A Comprehensive Overview of Large Language Models357 citations · 2023
- 3Video Description146 citations · 2019
- 4Video Description: A Survey of Methods, Datasets and Evaluation Metrics95 citations · 2018
- 5Deep Learning Based 3D Segmentation: A Survey44 citations · 2021
- 6
- 7Deep learning based 3D segmentation in computer vision: A survey33 citations · 2024
- 83D point cloud-based place recognition: a survey30 citations · 2024
- 9
- 10History-Enhanced 3D Scene Graph Reasoning From RGB-D Sequences28 citations · 2025