Nan Duan

Active 2008–2026

332
Papers
25,034
Citations
68
h-index
194
i10-index

Citations

Citations per year for Nan Duan1979: 1 citations1993: 1 citations2001: 1 citations2006: 1 citations2007: 1 citations2008: 2 citations2009: 2 citations2010: 8 citations2011: 10 citations2012: 11 citations2013: 10 citations2014: 7 citations2015: 29 citations2016: 38 citations2017: 59 citations2018: 139 citations2019: 281 citations2020: 634 citations2021: 1,166 citations2022: 1,455 citations2023: 2,878 citations2024: 3,805 citations2025: 4,172 citations2026: 1,446 citations1980–1992: no citations, so these years are not shown1994–2000: no citations, so these years are not shown2002–2005: no citations, so these years are not shown

Citation sources

Countries

World map of the countries and regions citing this authorChina: 4,371 citing papers, 34.5% of this breakdownUnited States: 2,387 citing papers, 18.8% of this breakdownUnited Kingdom: 682 citing papers, 5.4% of this breakdownHong Kong: 442 citing papers, 3.5% of this breakdownGermany: 416 citing papers, 3.3% of this breakdownCanada: 415 citing papers, 3.3% of this breakdownSingapore: 414 citing papers, 3.3% of this breakdownAustralia: 408 citing papers, 3.2% of this breakdownIndia: 294 citing papers, 2.3% of this breakdownSouth Korea: 259 citing papers, 2% of this breakdownJapan: 225 citing papers, 1.8% of this breakdownItaly: 188 citing papers, 1.5% of this breakdown
0%34.5%Other 17.1%

Fields

  • Computer Science83.2%
  • Biochemistry, Genetics and Molecular Biology3.4%
  • Engineering3.1%
  • Social Sciences3%
  • Medicine1.6%
  • Neuroscience1.1%
  • Other4.6%

Topics

  • Topic Modeling14.1%
  • Natural Language Processing Techniques9.3%
  • Multimodal Machine Learning Applications8.7%
  • Software Engineering Research4.5%
  • Domain Adaptation and Few-Shot Learning2.9%
  • Advanced Image and Video Retrieval Techniques2.5%
  • Other58%

Coauthors

All papers

Open in search
  1. CodeBERT: A Pre-Trained Model for Programming and Natural Languages

    Authors: , , , , , , , , , , - Findings of the Association for Computational Linguistics: EMNLP 2020, EMNLP (Findings) 2020 cited by 2,434

  2. Findings of the Association for Computational Linguistics: EMNLP 2021

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Wenhu Chen, Xifeng Yan, ;, Weiwen Xu, Yang Deng, Huihui Zhang, Deng Cai, Wai Lam, . . . . . . . . . . . . . . . . . . . . . ; Bowei, Zhifeng Zou, Manuel Li, Jan-David Plank, Terry Krieger, Bela Ruas, Akiko Gipp, Aizawa, Jiangnan Li, Zheng Lin, Peng Fu, Weiping Wang ; Yanyan, Hainan Zou, Hongshen Zhang, Zhuoye Chen, Caixia Ding, Xiaojie Yuan, Wang, Guoxin Yu, Jiwei Li, Ling Luo, Yuxian Meng, Xiang Ao, Qing He, Zixuan Zhang, Hongwei Wang, Han Zhao, Hanghang Tong, Heng Ji ; Guangrun, Hang Wang, Jiefeng Xu, Bert, Tong Overkill ; Bryan Mccann, Nazneen Niu, Rajani, Shirish Nitish, Thamar Keskar, Solorio, Shifeng Liu, Yifang Sun, Bing Li, Wei Wang, Florence Bourgeois, Adam Dunn, . ; Meishan, Zhenghua Zhang, Min ; Meng Li, Oyvind Huang, Chao Tafjord, Zhao, Yunlong Liang, Fandong Meng, Jinchao Zhang, Yufeng Chen, Jinan Xu, Jie Zhou, Yang Zhong, Jingfeng Yang, Wei Xu, Diyi Yang ; Yuchen, Chengyu Zhai, Minghui Wang and 296 more - 2021 cited by 896

  3. UniXcoder: Unified Cross-Modal Pre-training for Code Representation

    Authors: , , , , , - Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ECOOP 2022 cited by 678

  4. Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation Models

    Authors: , , , , , - arXiv (Cornell University), CoRR 2023 cited by 688

  5. CLIP4Clip: An empirical study of CLIP for end to end video clip retrieval and captioning

    Authors: , , , , , , - Neurocomputing 2022 cited by 677

  6. scGPT: toward building a foundation model for single-cell multi-omics using generative AI

    Authors: , , , , , , - Nature Methods 2024 cited by 1,044

  7. AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

    Authors: , , , , , , , , - Findings of the Association for Computational Linguistics: NAACL 2024, NAACL-HLT (Findings) 2024 cited by 485

  8. CodeXGLUE: A Machine Learning Benchmark Dataset for Code Understanding and Generation

    Authors: , , , , , , , , , , , , , , , , , , , , , - NeurIPS Datasets and Benchmarks 2021 cited by 1,543

  9. GraphCodeBERT: Pre-training Code Representations with Data Flow

    Authors: , , , , , , , , , , , , , , , , , - ICLR 2021 cited by 1,779

  10. CRITIC: Large Language Models Can Self-Correct with Tool-Interactive Critiquing

    Authors: , , , , , , - ICLR 2024 cited by 816

  11. CMMLU: Measuring massive multitask language understanding in Chinese

    Authors: , , , , , , , - Findings of the Association for Computational Linguistics ACL 2024, ACL (Findings) 2024 cited by 301

  12. DragNUWA: Fine-grained Control in Video Generation by Integrating Text, Image, and Trajectory

    Authors: , , , , , , - arXiv (Cornell University), CoRR 2023 cited by 236

  13. ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving

    Authors: , , , , , , , - ICLR 2024 cited by 314

  14. GODIVA: Generating Open-DomaIn Videos from nAtural Descriptions

    Authors: , , , , , , , - arXiv (Cornell University), CoRR 2021 cited by 245

  15. Enhancing Retrieval-Augmented Large Language Models with Iterative Retrieval-Generation Synergy

    Authors: , , , , , - Findings of the Association for Computational Linguistics: EMNLP 2023, EMNLP (Findings) 2023 cited by 176

  16. Unicoder-VL: A Universal Encoder for Vision and Language by Cross-Modal Pre-Training

    Authors: , , , , - AAAI Conference on Artificial Intelligence 2020 cited by 750

  17. Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Chenfei Wu, Chenguang Yu, Dapeng Shi, Dingyuan Hu, Liu, Enle, Yu, Gang, Ge Yang, Guangnan Huang, Yan, Gulin, Haiyang Feng, Hao Nie, Haonan Jia, Hanpeng Hu, Hanqi Chen, Haolong Yan, Heng Wang, Hongcheng Guo, Huilin Xiong, Huixin Xiong, Jiahao Gong, Jianchang Wu, Jiao-Ren Wu, Jie Wu, Jie Yang, Jiashuai Liu, Jiashuo Li, Jingyang Zhang, Jiajie Guo, Junzhe Lin, Kaixiang Li, Lei Liu, Lei Xia, Liang Zhao, Liguo Tan, Liwen Huang, Liying Shi, Ming Li, Mingliang Li, Mu‐hua Cheng, Na Wang, Qiaohui Chen, Qinglin He, Qiuyan Liang, Quan Sun, Ran Sun, Rui Wang, Shaoliang Pang, Shiliang Yang, Sitong Liu, Siqi Liu, Shuli Gao, Tiancheng Cao, Tianyu Wang, Weipeng Ming, Wenqing He, Xu Zhao, Xuelin Zhang, Xianfang Zeng, Xiaojia Liu, Xuan Yang, Dai, Yaqi, Yanbo Yu, Yang Li, Deng, Yineng, Ying-Ming Wang, Yilei Wang, Yuanwei Lu, Yu Chen, Yu Luo, Yi Luo and 15 more - ArXiv.org, CoRR 2025 cited by 152

  18. Query Rewriting for Retrieval-Augmented Large Language Models

    Authors: , , , , - arXiv (Cornell University), CoRR 2023 cited by 157

  19. Query Rewriting in Retrieval-Augmented Large Language Models

    Authors: , , , , - Conference on Empirical Methods in Natural Language Processing, EMNLP 2023 cited by 209

  20. Baize: An Open-Source Chat Model with Parameter-Efficient Tuning on Self-Chat Data

    Authors: , , , - Conference on Empirical Methods in Natural Language Processing, EMNLP 2023 cited by 193

  21. Rho-1: Not All Tokens Are What You Need

    Authors: , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 124

  22. UniViLM: A Unified Video and Language Pre-Training Model for Multimodal Understanding and Generation

    Authors: , , , , , , , , - arXiv (Cornell University), CoRR 2020 cited by 306

  23. AnnoLLM: Making Large Language Models to Be Better Crowdsourced Annotators

    Authors: , , , , , , , , , - Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 6: Industry Track), NAACL (Industry Track) 2024 cited by 136

  24. NUWA-XL: Diffusion over Diffusion for eXtremely Long Video Generation

    Authors: , , , , , , , , , , , , , , , - Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL (1) 2023 cited by 114