Junnan Li

Active 2008–2026

214
Papers
33,682
Citations
53
h-index
134
i10-index

Citations

Citations per year for Junnan Li1988: 1 citations1995: 1 citations2004: 2 citations2005: 1 citations2008: 1 citations2009: 2 citations2010: 2 citations2011: 1 citations2013: 1 citations2014: 4 citations2015: 6 citations2016: 12 citations2017: 30 citations2018: 43 citations2019: 153 citations2020: 327 citations2021: 525 citations2022: 801 citations2023: 2,121 citations2024: 3,668 citations2025: 2,680 citations2026: 905 citations2027: 2 citations1989–1994: no citations, so these years are not shown1996–2003: no citations, so these years are not shown2006–2007: no citations, so these years are not shown2012: no citations, so this year is not shown

Citation sources

Countries

World map of the countries and regions citing this authorChina: 3,572 citing papers, 37.6% of this breakdownUnited States: 1,524 citing papers, 16.1% of this breakdownUnited Kingdom: 446 citing papers, 4.7% of this breakdownHong Kong: 398 citing papers, 4.2% of this breakdownSingapore: 358 citing papers, 3.8% of this breakdownAustralia: 305 citing papers, 3.2% of this breakdownGermany: 260 citing papers, 2.7% of this breakdownSouth Korea: 248 citing papers, 2.6% of this breakdownCanada: 210 citing papers, 2.2% of this breakdownIndia: 205 citing papers, 2.2% of this breakdownJapan: 188 citing papers, 2% of this breakdownFrance: 140 citing papers, 1.5% of this breakdown
0%37.6%Other 17.2%

Fields

  • Computer Science73.7%
  • Engineering5.2%
  • Biochemistry, Genetics and Molecular Biology5%
  • Medicine4.6%
  • Materials Science2.3%
  • Social Sciences1.7%
  • Other7.5%

Topics

  • Multimodal Machine Learning Applications11.6%
  • Topic Modeling5.9%
  • Domain Adaptation and Few-Shot Learning5.7%
  • Advanced Image and Video Retrieval Techniques4%
  • Natural Language Processing Techniques3.9%
  • Advanced Neural Network Applications2.6%
  • Other66.3%

Coauthors

All papers

Open in search
  1. BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

    Authors: , , , - ICML 2023 cited by 8,663

  2. InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

    Authors: , , , , , , , , - Advances in Neural Information Processing Systems 36, NeurIPS 2023 cited by 3,733

  3. BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation

    Authors: , , , - ICML 2022 cited by 7,076

  4. Align before Fuse: Vision and Language Representation Learning with Momentum Distillation

    Authors: , , , , , - Neural Information Processing Systems, NeurIPS 2021 cited by 2,847

  5. CodeT5+: Open Code Large Language Models for Code Understanding and Generation

    Authors: , , , , , - Conference on Empirical Methods in Natural Language Processing, EMNLP 2023 cited by 347

  6. DivideMix: Learning with Noisy Labels as Semi-supervised Learning

    Authors: , , - ICLR 2020 cited by 1,393

  7. Prototypical Contrastive Learning of Unsupervised Representations

    Authors: , , , , - ICLR 2021 cited by 1,201

  8. CoMatch: Semi-supervised Learning with Contrastive Graph Regularization

    Authors: , , - IEEE/CVF International Conference on Computer Vision (ICCV) 2021 cited by 265

  9. LongVideoBench: A Benchmark for Long-context Interleaved Video-Language Understanding

    Authors: , , , - Advances in Neural Information Processing Systems 37, NeurIPS 2024 cited by 659

  10. A Minimalist Approach to LLM Reasoning: from Rejection Sampling to Reinforce

    Authors: , , , , , , , , , , - ArXiv.org, CoRR 2025 cited by 124

  11. Aria: An Open Multimodal Native Mixture-of-Experts Model

    Authors: , , , , , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 124

  12. BLIP-Diffusion: Pre-trained Subject Representation for Controllable Text-to-Image Generation and Editing

    Authors: , , - Advances in Neural Information Processing Systems 36, NeurIPS 2023 cited by 126

  13. From Images to Textual Prompts: Zero-shot Visual Question Answering with Frozen Large Language Models

    Authors: , , , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2023 cited by 131

  14. Aria-UI: Visual Grounding for GUI Instructions

    Authors: , , , , , , - Findings of the Association for Computational Linguistics: ACL 2025, ACL (Findings) 2025 cited by 93

  15. GTA1: GUI Test-time Scaling Agent

    Authors: , , , , , , , , , , , , , , - ArXiv.org, CoRR 2025 cited by 78

  16. Reward-Guided Speculative Decoding for Efficient LLM Reasoning

    Authors: , , , , , , , - ICML 2025 cited by 109

  17. Moirai-MoE: Empowering Time Series Foundation Models with Sparse Mixture of Experts

    Authors: , , , , , , , , , , - ICML 2025 cited by 107

  18. X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning

    Authors: , , , , , , , , , - arXiv (Cornell University), CoRR 2023 cited by 69

  19. MCP-Universe: Benchmarking Large Language Models with Real-World Model Context Protocol Servers

    Authors: , , , , , , , , , - ArXiv.org, CoRR 2025 cited by 63

  20. ULIP-2: Towards Scalable Multimodal Pre-Training for 3D Understanding

    Authors: , , , , , , , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 cited by 100

  21. Learning from Noisy Data with Robust Representation Learning

    Authors: , , - IEEE/CVF International Conference on Computer Vision (ICCV) 2021 cited by 108

  22. Mechanism of pH-switchable peroxidase and catalase-like activities of gold, silver, platinum and palladium

    Authors: , , , - Biomaterials 2015 cited by 557

  23. A novel oversampling technique for class-imbalanced learning based on SMOTE and natural neighbors

    Authors: , , , - Information Sciences, Inf. Sci. 2021 cited by 182

  24. Open Vocabulary Object Detection with Pseudo Bounding-Box Labels

    Authors: , , , , , , - Lecture notes in computer science, ECCV (10) 2022 cited by 77