Jan Kautz

Active 1998–2026

426
Papers
53,138
Citations
101
h-index
303
i10-index

Citations

Citations per year for Jan Kautz1972: 1 citations1988: 3 citations1990: 2 citations1994: 2 citations1998: 1 citations1999: 1 citations2000: 7 citations2001: 18 citations2002: 49 citations2003: 118 citations2004: 123 citations2005: 121 citations2006: 186 citations2007: 195 citations2008: 178 citations2009: 201 citations2010: 193 citations2011: 181 citations2012: 223 citations2013: 291 citations2014: 280 citations2015: 384 citations2016: 391 citations2017: 525 citations2018: 979 citations2019: 1,542 citations2020: 1,915 citations2021: 2,148 citations2022: 1,645 citations2023: 1,680 citations2024: 2,091 citations2025: 3,518 citations2026: 2,454 citations1973–1987: no citations, so these years are not shown1989: no citations, so this year is not shown1991–1993: no citations, so these years are not shown1995–1997: no citations, so these years are not shown

Citation sources

Countries

World map of the countries and regions citing this authorChina: 5,445 citing papers, 28.1% of this breakdownUnited States: 3,641 citing papers, 18.8% of this breakdownUnited Kingdom: 1,184 citing papers, 6.1% of this breakdownGermany: 975 citing papers, 5% of this breakdownSouth Korea: 698 citing papers, 3.6% of this breakdownCanada: 617 citing papers, 3.2% of this breakdownJapan: 579 citing papers, 3% of this breakdownHong Kong: 578 citing papers, 3% of this breakdownFrance: 558 citing papers, 2.9% of this breakdownSingapore: 437 citing papers, 2.3% of this breakdownIndia: 430 citing papers, 2.2% of this breakdownAustralia: 407 citing papers, 2.1% of this breakdown
0%28.1%Other 19.7%

Fields

  • Computer Science73.5%
  • Engineering15.9%
  • Medicine2.6%
  • Physics and Astronomy1.5%
  • Earth and Planetary Sciences1.1%
  • Environmental Science0.9%
  • Other4.5%

Topics

  • Advanced Vision and Imaging7.1%
  • Generative Adversarial Networks and Image Synthesis4.8%
  • Computer Graphics and Visualization Techniques4.7%
  • Image Enhancement Techniques4.4%
  • Advanced Image Processing Techniques4.1%
  • Multimodal Machine Learning Applications3.9%
  • Other71%

Coauthors

All papers

Open in search
  1. GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

    Authors: , :, , , , , , , , , , , , , , , , , , , , , , , , , , , , , Jing Wang, Qi Wang, Jiannan Xiang, Yuqi Xie, Yinzhen Xu, Zhenjia Xu, Seonghyeon Ye, Zhiding Yu, Ao Zhang, Hao Zhang, Yizhou Zhao, Ruijie Zheng, Yuke Zhu - ArXiv.org, CoRR 2025 cited by 741

  2. Loss Functions for Image Restoration With Neural Networks

    Authors: , , , - IEEE Transactions on Computational Imaging, IEEE Trans. Computational Imaging 2016 cited by 2,633

  3. LongVILA: Scaling Long-Context Visual Language Models for Long Videos

    Authors: , , , , , , , , , , , , , , , , , - ICLR 2025 cited by 307

  4. Gated Delta Networks: Improving Mamba2 with Delta Rule

    Authors: , , - ICLR 2025 cited by 386

  5. A-ViT: Adaptive Tokens for Efficient Vision Transformer

    Authors: , , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022 cited by 299

  6. ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models

    Authors: , , , , , , , - NeurIPS 2025 cited by 180

  7. Hydra-MDP: End-to-end Multimodal Planning with Multi-target Hydra-Distillation

    Authors: , , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 176

  8. An Empirical Study of Mamba-based Language Models

    Authors: , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 168

  9. Pruning Convolutional Neural Networks for Resource Efficient Transfer Learning

    Authors: , , , , - International Conference on Learning Representations, ICLR (Poster) 2016 cited by 2,293

  10. FB-OCC: 3D Occupancy Prediction based on Forward-Backward View Transformation

    Authors: , , , , , , - arXiv (Cornell University), CoRR 2023 cited by 147

  11. NaVILA: Legged Robot Vision-Language-Action Model for Navigation

    Authors: , , , , , , , , , - Robotics: Science and Systems XXI 2025 cited by 130

  12. CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation

    Authors: , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 124

  13. NVILA: Efficient Frontier Visual Language Models

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , - ArXiv.org, CoRR 2024 cited by 123

  14. World Action Models are Zero-shot Policies

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Yevgen Chebotar, Scott Reed, Jan Kautz, Yuke Zhu, Linxi "Jim" Fan, Joel Jang - ArXiv.org, CoRR 2026 cited by 116

  15. MambaVision: A Hybrid Mamba-Transformer Vision Backbone

    Authors: , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 cited by 195

  16. World Simulation with Video Foundation Models for Physical AI

    Authors: , :, , , , , , , , , , , , , , , , , , , , , , , , , , , , , Siddharth Gururani, Imad El Hanafi, Ali Hassani, Hao, Zekun, J. B. Huffman, Joel Jang, Pooya Jannaty, Jan Kautz, Grace Y. Lam, Xuan Li, Zhaoshuo Li, Maosheng Liao, Chen-Hsuan Lin, Tsung‐Yi Lin, Yen-Chen Lin, Huan Ling, Ming-Yu Liu, Xian Liu, Yifan Lu, Ancheng Luo, Qianli Ma, Hanzi Mao, Kaichun Mo, Seungjun Nah, Yashraj Narang, Abhijeet Panaskar, Lindsey Pavao, Trung Kien Pham, Morteza Ramezanali, Fitsum A. Reda, Scott Reed, Xuanchi Ren, Haonan Shao, Yue Shen, Stella Shi, Shuran Song, Bartosz Stefaniak, Shangkun Sun, Shitao Tang, Sameena Tasmeen, Lyne P. Tchapmi, Wei‐Cheng Tseng, Jibin Varghese, Andrew Z. Wang, Hao Wang, Haoxiang Wang, Heng Wang, Ting-Chun Wang, Fangyin Wei, Jiashu Xu, Yang, Dinghao, Xiaodong Yang, Haotian Ye, Seonghyeon Ye, Xiaohui Zeng, Jing Zhang, Qinsheng Zhang, Kaiwen Zheng, Andrew X. Zhu, Yuke Zhu - ArXiv.org, CoRR 2025 cited by 113

  17. SpatialRGPT: Grounded Spatial Reasoning in Vision-Language Models

    Authors: , , , , , , , - Advances in Neural Information Processing Systems 37, NeurIPS 2024 cited by 111

  18. Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders

    Authors: , , , , , , , , , , , , , , , - ICLR 2025 cited by 152

  19. VILA: On Pre-training for Visual Language Models

    Authors: , , , , , , , , , - arXiv (Cornell University), CoRR 2023 cited by 109

  20. OmniDrive: A Holistic LLM-Agent Framework for Autonomous Driving with 3D Perception, Reasoning and Planning

    Authors: , , , , , , , , - IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025 cited by 108

  21. Online Detection and Classification of Dynamic Hand Gestures with Recurrent 3D Convolutional Neural Networks

    Authors: , , , , , - IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2016 cited by 675

  22. A Variational Perspective on Solving Inverse Problems with Diffusion Models

    Authors: , , , - ICLR 2024 cited by 269

  23. NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model

    Authors: , :, , , , , , , , , , , , , , , , , , , , , , , , , , , , , Banghua Zhu, Barnaby Simkin, Bilal Kartal, Bita Darvish Rouhani, B Chen, Boris Ginsburg, Brandon Norick, Brian Yu, Bryan Catanzaro, Charles Wang, Truong, Charlie, Chetan Mungekar, Chintan Patel, Chris Alexiuk, Christian Munley, Christopher Parisien, Dan Su, Daniel Afrimi, Daniel Korzekwa, Daniel Rohrer, Daria Gitman, David Mosallanezhad, Deepak Narayanan, Dima Rekesh, Dina Yared, Dmytro Pykhtar, Dong Uk Ahn, Duncan Riach, Eileen Long, Elliott Ning, Eric S. Chung, Erick Galinkin, Evelina Bakhturina, Prasad, Gargi, Gerald Shen, Haifeng Qian, Haim Elisha, Harsh Sharma, Hayley Ross, Helen L. Ngo, Herman Sahota, Hexin Wang, Hoo Chang Shin, Hua Huang, Cunningham, Iain, Igor Gitman, Ivan Moshkov, Jaehun Jung, Jan Kautz, Jane Polak Scowcroft, Jared Casper, Jian Zhang, Jiaqi Zeng, Jimmy Zhang, Jinze Xue, Jocelyn Huang, J. S. Conway, John Kamalu, Jonathan D. Cohen, Joseph Jennings, J. Vialard, Jiun-Hung Yi, Jupinder Parmar, Kari Briski, Katherine Cheung, Katherine Luna, Keith Wyss, Keshav Santhanam, Kezhi Kong, Krzysztof Pawelec and 117 more - ArXiv.org, CoRR 2025 cited by 91

  24. Nemotron-H: A Family of Accurate and Efficient Hybrid Mamba-Transformer Models

    Authors: , :, , , , , , , , , , , , , , , , , , , , , , , , , , , , , Yin, Danny, Daria Gitman, David Mosallanezhad, Deepak Narayanan, Denys Fridman, Dima Rekesh, Ding Ma, Dmytro Pykhtar, Dong H. Ahn, Duncan Riach, Dušan Stošić, Eileen Long, Elad Segal, Ellie Evans, Eric S. Chung, Erick Galinkin, Evelina Bakhturina, Ewa Dobrowolska, Fei Jia, Fuxiao Liu, Prasad, Gargi, Gerald Shen, Guilin Liu, Chen Guo, Haifeng Qian, Helen L. Ngo, Hongbin Liu, Hui Li, Igor Gitman, Ilia Karmanov, Ivan Moshkov, Izik Golan, Jan Kautz, Jane Polak Scowcroft, Jared Casper, Jarno Seppänen, Jason Lu, Jason Sewall, Zeng, Jiaqi, You, Jiaxuan, Jimmy Zhang, Jing Zhang, Jining Huang, Jinze Xue, Jocelyn Huang, J. S. Conway, John Kamalu, Jon Barker, Jonathan Cohen, Joseph Jennings, Jupinder Parmar, Karan Sapra, Kari Briski, Kateryna Chumachenko, Katherine Luna, Keshav Santhanam, Kezhi Kong, Kirthi Sivamani, Krzysztof Pawelec, Kumar Anik, Kunlun Li, Lawrence McAfee, Leon Derczynski, Lindsey Pavao, L.A. Vega, Lukas Voegtle, Bala, Maciej, Maer Rodrigues de Melo, Makesh Narsimhan Sreedhar, Marcin Chochowski and 101 more - ArXiv.org, CoRR 2025 cited by 89