Tiejun Huang

Active 2000–2026

Also published as
TieJun Huang
499
Papers
20,211
Citations
70
h-index
262
i10-index

Citations

Citations per year for Tiejun Huang1905: 1 citations1998: 1 citations2001: 1 citations2004: 3 citations2005: 7 citations2006: 12 citations2007: 11 citations2008: 16 citations2009: 29 citations2010: 39 citations2011: 74 citations2012: 81 citations2013: 104 citations2014: 164 citations2015: 232 citations2016: 267 citations2017: 372 citations2018: 545 citations2019: 723 citations2020: 811 citations2021: 840 citations2022: 814 citations2023: 1,089 citations2024: 2,035 citations2025: 2,941 citations2026: 1,218 citations1906–1997: no citations, so these years are not shown1999–2000: no citations, so these years are not shown2002–2003: no citations, so these years are not shown

Citation sources

Countries

World map of the countries and regions citing this authorChina: 4,441 citing papers, 44% of this breakdownUnited States: 1,135 citing papers, 11.2% of this breakdownUnited Kingdom: 464 citing papers, 4.6% of this breakdownHong Kong: 345 citing papers, 3.4% of this breakdownIndia: 325 citing papers, 3.2% of this breakdownAustralia: 315 citing papers, 3.1% of this breakdownSingapore: 304 citing papers, 3% of this breakdownSouth Korea: 250 citing papers, 2.5% of this breakdownGermany: 222 citing papers, 2.2% of this breakdownCanada: 203 citing papers, 2% of this breakdownJapan: 180 citing papers, 1.8% of this breakdownFrance: 179 citing papers, 1.8% of this breakdown
0%44%Other 17.2%

Fields

  • Computer Science72%
  • Engineering14.3%
  • Psychology4.4%
  • Medicine1.9%
  • Neuroscience1.9%
  • Social Sciences1.1%
  • Other4.4%

Topics

  • Multimodal Machine Learning Applications6.2%
  • Video Surveillance and Tracking Methods5.2%
  • Advanced Image and Video Retrieval Techniques4.8%
  • Advanced Neural Network Applications4.1%
  • Generative Adversarial Networks and Image Synthesis3.5%
  • Human Pose and Action Recognition3.4%
  • Other72.8%

Coauthors

All papers

Open in search
  1. Emu3: Next-Token Prediction is All You Need

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 723

  2. OmniGen2: Exploration to Advanced Multimodal Generation

    Authors: , , , , , , , , , , , , , , , , , , , , , - ArXiv.org, CoRR 2025 cited by 347

  3. MLVU: A Comprehensive Benchmark for Multi-Task Long Video Understanding

    Authors: , , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 333

  4. Emu: Generative Pretraining in Multimodality

    Authors: , , , , , , , , , - ICLR 2024 cited by 249

  5. SegGPT: Towards Segmenting Everything In Context

    Authors: , , , , , - IEEE/CVF International Conference on Computer Vision (ICCV) 2023 cited by 271

  6. OmniGen: Unified Image Generation

    Authors: , , , , , , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 cited by 177

  7. SegGPT: Segmenting Everything In Context

    Authors: , , , , , - arXiv (Cornell University), CoRR 2023 cited by 208

  8. M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

    Authors: , , , , - arXiv (Cornell University), CoRR 2024 cited by 146

  9. EVA-02: A Visual Representation for Neon Genesis

    Authors: , , , , , - Image and Vision Computing, Image Vis. Comput. 2024 cited by 170

  10. SVIT: Scaling up Visual Instruction Tuning

    Authors: , , , - arXiv (Cornell University), CoRR 2023 cited by 144

  11. Efficient Multimodal Learning from Data-centric Perspective

    Authors: , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 120

  12. Generative Multimodal Models are In-Context Learners

    Authors: , , , , , , , , , , - Advances in computer vision and pattern recognition 2025 cited by 110

  13. Deep Relative Distance Learning: Tell the Difference between Similar Vehicles

    Authors: , , , , - IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2016 cited by 807

  14. Emu3.5: Native Multimodal Models are World Learners

    Authors: , , , , , , , , , , , , , , , , , , , , , , - ArXiv.org, CoRR 2025 cited by 98

  15. MemoRAG: Boosting Long Context Processing with Global Memory-Enhanced Retrieval Augmentation

    Authors: , , , , , , - on Web Conference 2025, WWW 2025 cited by 98

  16. Optimal ANN-SNN Conversion for High-accuracy and Ultra-low-latency Spiking Neural Networks

    Authors: , , , , , - ICLR 2022 cited by 295

  17. RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics

    Authors: , , , , , , , , , , - NeurIPS 2025 cited by 96

  18. Learning Open Set Network with Discriminative Reciprocal Points

    Authors: , , , , , , , - Lecture notes in computer science, ECCV (3) 2020 cited by 227

  19. Uni3D: Exploring Unified 3D Representation at Scale

    Authors: , , , , , - ICLR 2024 cited by 248

  20. Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding

    Authors: , , , , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 cited by 86

  21. Asynchronous Spatio-Temporal Memory Network for Continuous Event-Based Object Detection

    Authors: , , , , , - IEEE Transactions on Image Processing, IEEE Trans. Image Process. 2022 cited by 121

  22. Generative Multimodal Models are In-Context Learners

    Authors: , , , , , , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2024 cited by 75

  23. EVA: Exploring the Limits of Masked Visual Representation Learning at Scale

    Authors: , , , , , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2023 cited by 447

  24. Speech Emotion Recognition Using Deep Convolutional Neural Network and Discriminant Temporal Pyramid Matching

    Authors: , , , - IEEE Transactions on Multimedia, IEEE Trans. Multim. 2017 cited by 437