Dahua Lin

Active 2005–2026

Also published as
Dahua, Lin
521
Papers
60,697
Citations
102
h-index
342
i10-index

Citations

Citations per year for Dahua Lin1971: 1 citations1978: 1 citations1988: 1 citations1990: 1 citations1994: 1 citations1998: 1 citations1999: 1 citations2000: 1 citations2002: 1 citations2004: 1 citations2005: 5 citations2006: 17 citations2007: 33 citations2008: 36 citations2009: 34 citations2010: 40 citations2011: 55 citations2012: 67 citations2013: 91 citations2014: 128 citations2015: 162 citations2016: 194 citations2017: 373 citations2018: 669 citations2019: 1,203 citations2020: 2,107 citations2021: 2,968 citations2022: 2,720 citations2023: 3,090 citations2024: 5,566 citations2025: 10,177 citations2026: 5,092 citations1972–1977: no citations, so these years are not shown1979–1987: no citations, so these years are not shown1989: no citations, so this year is not shown1991–1993: no citations, so these years are not shown1995–1997: no citations, so these years are not shown2001: no citations, so this year is not shown2003: no citations, so this year is not shown

Citation sources

Countries

World map of the countries and regions citing this authorChina: 10,154 citing papers, 40% of this breakdownUnited States: 4,141 citing papers, 16.3% of this breakdownHong Kong: 1,186 citing papers, 4.7% of this breakdownUnited Kingdom: 1,174 citing papers, 4.6% of this breakdownSingapore: 872 citing papers, 3.4% of this breakdownAustralia: 853 citing papers, 3.4% of this breakdownGermany: 724 citing papers, 2.8% of this breakdownSouth Korea: 673 citing papers, 2.6% of this breakdownCanada: 576 citing papers, 2.3% of this breakdownIndia: 491 citing papers, 1.9% of this breakdownJapan: 490 citing papers, 1.9% of this breakdownFrance: 394 citing papers, 1.6% of this breakdown
0%40%Other 14.5%

Fields

  • Computer Science80.6%
  • Engineering9%
  • Medicine2.1%
  • Social Sciences1.1%
  • Psychology1.1%
  • Neuroscience0.9%
  • Other5.2%

Topics

  • Multimodal Machine Learning Applications10%
  • Human Pose and Action Recognition6.8%
  • Domain Adaptation and Few-Shot Learning6.2%
  • Advanced Neural Network Applications5.5%
  • Advanced Image and Video Retrieval Techniques3.7%
  • Generative Adversarial Networks and Image Synthesis3.6%
  • Other64.2%

Coauthors

All papers

Open in search
  1. InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Wenqi Shao, Junjun He, Yingtong Xiong, Wenwen Qu, Peng Sun, Penglong Jiao, Han Lv, Lijun Wu, Kaipeng Zhang, Huipeng Deng, Jiaye Ge, Kai Chen, Limin Wang, Min Dou, Lewei Lu, Xizhou Zhu, Tong Lu, Dahua Lin, Yu Qiao, Jifeng Dai, Wenhai Wang - ArXiv.org, CoRR 2025 cited by 1,446

  2. Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Jiaye Ge, Kai Chen, Zhang, Kaipeng, Wang, Limin, Min Dou, Lewei Lu, Zhu, Xizhou, Tong Lü, Dahua Lin, Yu Qiao, Jifeng Dai, Wenhai Wang - arXiv (Cornell University), CoRR 2024 cited by 1,368

  3. InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Tianyi Zhang, Songze Li, Xiangyu Zhao, Haodong Duan, Nianchen Deng, Bin Fu, Yinan He, Yi Wang, Conghui He, Botian Shi, Junjun He, Yingtong Xiong, Han Lv, Lijun Wu, Wenqi Shao, Kaipeng Zhang, Huipeng Deng, Biqing Qi, Jiaye Ge, Qipeng Guo, Wenwei Zhang, Songyang Zhang, Maosong Cao, Junyao Lin, Kexian Tang, Jianfei Gao, Haian Huang, Yuzhe Gu, Chengqi Lyu, Huanze Tang, Rui Wang, Haijun Lv, Wanli Ouyang, Limin Wang, Min Dou, Xizhou Zhu, Tong Lu, Dahua Lin, Jifeng Dai, Weijie Su, Bowen Zhou, Kai Chen, Yu Qiao, Wenhai Wang, Gen Luo - ArXiv.org, CoRR 2025 cited by 1,013

  4. Unsupervised Feature Learning via Non-Parametric Instance Discrimination

    Authors: , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2018 cited by 3,580

  5. AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

    Authors: , , , , , , , , - ICLR 2024 cited by 1,624

  6. MMBench: Is Your Multi-modal Model an All-Around Player?

    Authors: , , , , , , , , , , , - Lecture notes in computer science, ECCV (6) 2024 cited by 889

  7. How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Tong Lu, Dahua Lin, Yu Qiao, Jifeng Dai, Wenhai Wang, Wenhai Wang, Wenhai Wang - Science China Information Sciences, Sci. China Inf. Sci. 2024 cited by 701

  8. MMDetection: Open MMLab Detection Toolbox and Benchmark

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2019 cited by 1,692

  9. InternLM2 Technical Report

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Linyang Li, Shuaibin Li, Wei Li, Yining Li, Hongwei Liu, Jiangning Liu, Jiawei Hong, Kaiwen Liu, Kuikun Liu, Xiaoran Liu, Chengqi Lv, Haijun Lv, Kai Lv, Li Ma, Runyuan Ma, Zerun Ma, Wenchang Ning, Linke Ouyang, Jiantao Qiu, Yuan Qu, Fukai Shang, Yunfan Shao, Demin Song, Zifan Song, Zhihao Sui, Peng Sun, Yu Sun, Huanze Tang, Bin Wang, Guoteng Wang, Jiaqi Wang, Jiayu Wang, Rui Wang, Yudong Wang, Ziyi Wang, Xingjian Wei, Qizhen Weng, Fan Wu, Yingtong Xiong, Chao Xu, Ruiliang Xu, Hang Yan, Yirong Yan, Xiaogui Yang, Haochen Ye, Huaiyuan Ying, Jia Yu, Jing Yu, Yuhang Zang, Chuyu Zhang, Li Zhang, Pan Zhang, Peng Zhang, Ruijie Zhang, Shuo Zhang, Song‐Yang Zhang, Wenjian Zhang, Wenwei Zhang, Xingcheng Zhang, Xinyue Zhang, Hui Zhao, Qian Zhao, Xiaomeng Zhao, Fengzhe Zhou, Zaida Zhou, Jingming Zhuo, Yicheng Zou, Xipeng Qiu, Yu Qiao, Dahua Lin - arXiv (Cornell University), CoRR 2024 cited by 591

  10. Temporal Segment Networks: Towards Good Practices for Deep Action Recognition

    Authors: , , , , , , - Lecture notes in computer science, ECCV (8) 2016 cited by 3,924

  11. ShareGPT4V: Improving Large Multi-Modal Models with Better Captions

    Authors: , , , , , , , - Lecture notes in computer science, ECCV (17) 2024 cited by 561

  12. Are We on the Right Way for Evaluating Large Vision-Language Models?

    Authors: , , , , , , , , , , - Advances in Neural Information Processing Systems 37, NeurIPS 2024 cited by 890

  13. Learning a Unified Classifier Incrementally via Rebalancing

    Authors: , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2019 cited by 1,113

  14. Visual-RFT: Visual Reinforcement Fine-Tuning

    Authors: , , , , , , , - IEEE/CVF International Conference on Computer Vision (ICCV) 2025 cited by 378

  15. InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model

    Authors: , , , , , , , , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 322

  16. LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models

    Authors: , , , , , , , , , , , , , , , , , , , , , - International Journal of Computer Vision, Int. J. Comput. Vis. 2024 cited by 274

  17. InternLM-XComposer: A Vision-Language Large Model for Advanced Text-image Comprehension and Composition

    Authors: , , , , , , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2023 cited by 273

  18. PSANet: Point-wise Spatial Attention Network for Scene Parsing

    Authors: , , , , , , - Lecture notes in computer science, ECCV (9) 2018 cited by 1,257

  19. Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models

    Authors: , , , , , , - arXiv (Cornell University), CoRR 2023 cited by 220

  20. InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 202

  21. MinerU: An Open-Source Solution for Precise Document Content Extraction

    Authors: , , , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 197

  22. SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models

    Authors: , , , , , , , - Findings of the Association for Computational Linguistics ACL 2024, ACL (Findings) 2024 cited by 193

  23. ShareGPT4Video: Improving Video Understanding and Generation with Better Captions

    Authors: , , , , , , , , , , , , , , - Advances in Neural Information Processing Systems 37, NeurIPS 2024 cited by 454

  24. PyramidDrop: Accelerating Your Large Vision-Language Models via Pyramid Visual Redundancy Reduction

    Authors: , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 159