Martin Jaggi

Active 2000–2026

Also published as
Martin Jäggi
212
Papers
25,195
Citations
58
h-index
133
i10-index

Citations

Citations per year for Martin Jaggi1971: 1 citations1989: 1 citations2001: 7 citations2002: 6 citations2003: 4 citations2004: 6 citations2005: 3 citations2006: 3 citations2007: 2 citations2008: 10 citations2009: 7 citations2010: 10 citations2011: 21 citations2012: 37 citations2013: 43 citations2014: 84 citations2015: 142 citations2016: 152 citations2017: 224 citations2018: 318 citations2019: 529 citations2020: 1,172 citations2021: 1,803 citations2022: 1,588 citations2023: 1,736 citations2024: 2,041 citations2025: 2,074 citations2026: 755 citations1972–1988: no citations, so these years are not shown1990–2000: no citations, so these years are not shown

Citation sources

Countries

World map of the countries and regions citing this authorChina: 2,770 citing papers, 23.2% of this breakdownUnited States: 2,556 citing papers, 21.4% of this breakdownUnited Kingdom: 651 citing papers, 5.4% of this breakdownGermany: 436 citing papers, 3.6% of this breakdownAustralia: 382 citing papers, 3.2% of this breakdownHong Kong: 376 citing papers, 3.1% of this breakdownCanada: 366 citing papers, 3.1% of this breakdownFrance: 365 citing papers, 3% of this breakdownSwitzerland: 306 citing papers, 2.6% of this breakdownIndia: 288 citing papers, 2.4% of this breakdownSingapore: 272 citing papers, 2.3% of this breakdownSouth Korea: 267 citing papers, 2.2% of this breakdown
0%23.2%Other 24.5%

Fields

  • Computer Science78.6%
  • Engineering9%
  • Medicine2.7%
  • Biochemistry, Genetics and Molecular Biology1.4%
  • Mathematics1.3%
  • Decision Sciences1.3%
  • Other5.7%

Topics

  • Privacy-Preserving Technologies in Data14.2%
  • Stochastic Gradient Optimization Techniques7.3%
  • Topic Modeling3.8%
  • Cryptography and Data Security3.7%
  • Sparse and Compressive Sensing Techniques2.9%
  • Advanced Neural Network Applications2.5%
  • Other65.6%

Coauthors

All papers

Open in search
  1. Advances and Open Problems in Federated Learning

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Jakub Konecný, Aleksandra Korolova, Farinaz Koushanfar, Sanmi Koyejo, Tancrède Lepoint, Yang Liu, Prateek Mittal, Mehryar Mohri, Richard Nock, Ayfer Özgür, Rasmus Pagh, Hang Qi, Daniel Ramage, Ramesh Raskar, Mariana Raykova, Dawn Song, Weikang Song, Sebastian U. Stich, Ziteng Sun, Ananda Theertha Suresh, Florian Tramèr, Praneeth Vepakomma, Jianyu Wang, Li Xiong, Zheng Xu, Qiang Yang, Felix X. Yu, Han Yu, Sen Zhao - Foundations and Trends® in Machine Learning, Found. Trends Mach. Learn. 2020 cited by 4,744

  2. MEDITRON-70B: Scaling Medical Pretraining for Large Language Models

    Authors: , , , , , , , , , , , , , , , , , , , - arXiv (Cornell University), CoRR 2023 cited by 428

  3. QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs

    Authors: , , , , , , , , - Advances in Neural Information Processing Systems 37, NeurIPS 2024 cited by 582

  4. Advances and Open Problems in Federated Learning

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Aleksandra Korolova, Farinaz Koushanfar, Sanmi Koyejo, Tancrède Lepoint, Yang Liu, Prateek Mittal, Mehryar Mohri, Richard Nock, Ayfer Özgür, Rasmus Pagh, Mariana Raykova, Hang Qi, Daniel Ramage, Ramesh Raskar, Dawn Song, Weikang Song, Sebastian U. Stich, Ziteng Sun, Ananda Theertha Suresh, Florian Tramèr, Praneeth Vepakomma, Jianyu Wang, Li Xiong, Zheng Xu, Qiang Yang, Felix X. Yu, Han Yu, Sen Zhao - CoRR 2019 cited by 1,103

  5. Ensemble Distillation for Robust Model Fusion in Federated Learning

    Authors: , , , - NeurIPS 2020 cited by 1,528

  6. A Field Guide to Federated Optimization

    Authors: , , , , , , , , , , , , , , , , , , , , , , , , , , , , , , Luyang Liu, Mehryar Mohri, Hang Qi, Sashank J. Reddi, Peter Richtárik, Karan Singhal, Virginia Smith, Mahdi Soltanolkotabi, Weikang Song, Ananda Theertha Suresh, Sebastian U. Stich, Ameet Talwalkar, Hongyi Wang, Blake E. Woodworth, Shanshan Wu, Felix X. Yu, Honglin Yuan, Manzil Zaheer, Mi Zhang, Tong Zhang, Chunxiang Zheng, Chen Zhu, Wennan Zhu - arXiv (Cornell University), CoRR 2021 cited by 378

  7. Landmark Attention: Random-Access Infinite Context Length for Transformers

    Authors: , - arXiv (Cornell University), CoRR 2023 cited by 102

  8. FineWeb2: One Pipeline to Scale Them All - Adapting Pre-Training Data Processing to Every Language

    Authors: , , , , , , , , , - ArXiv.org, CoRR 2025 cited by 76

  9. Mime: Mimicking Centralized Stochastic Algorithms in Federated Learning

    Authors: , , , , , , - arXiv (Cornell University), CoRR 2020 cited by 181

  10. Scaling Laws and Compute-Optimal Training Beyond Fixed Training Durations

    Authors: , , , , , - Advances in Neural Information Processing Systems 37, NeurIPS 2024 cited by 146

  11. Unsupervised Learning of Sentence Embeddings Using Compositional n-Gram Features

    Authors: , , - Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers), NAACL-HLT 2018 cited by 700

  12. On the Relationship between Self-Attention and Convolutional Layers

    Authors: , , - ICLR 2020 cited by 647

  13. DOGE: Domain Reweighting with Generalization Estimation

    Authors: , , - ICML 2024 cited by 97

  14. Multi-Head Attention: Collaborate Instead of Concatenate

    Authors: , , - arXiv (Cornell University), CoRR 2020 cited by 91

  15. Generating Steganographic Text with LSTMs

    Authors: , , - ACL 2017, Student Research Workshop, ACL (Student Research Workshop) 2017 cited by 141

  16. Benchmarking Optimizers for Large Language Model Pretraining

    Authors: , , - ArXiv.org, CoRR 2025 cited by 45

  17. FLamby: Datasets and Benchmarks for Cross-Silo Federated Learning in Realistic Healthcare Settings

    Authors: , , , , , , , , , , , , , , , , , , , , , , , - NeurIPS 2022 cited by 238

  18. Attention with Markov: A Framework for Principled Analysis of Transformers via Markov Chains

    Authors: , , , , , , - arXiv (Cornell University), CoRR 2024 cited by 41

  19. Learning Aerial Image Segmentation From Online Maps

    Authors: , , , , , - IEEE Transactions on Geoscience and Remote Sensing, IEEE Trans. Geosci. Remote. Sens. 2017 cited by 278

  20. Byzantine-Robust Decentralized Learning via Self-Centered Clipping

    Authors: , , - CoRR 2022 cited by 51

  21. Don't Use Large Mini-Batches, Use Local SGD

    Authors: , , , - ICLR 2020 cited by 470

  22. Masking as an Efficient Alternative to Finetuning for Pretrained Language Models

    Authors: , , , , - Conference on Empirical Methods in Natural Language Processing (EMNLP), EMNLP (1) 2020 cited by 75

  23. Byzantine-Robust Learning on Heterogeneous Datasets via Resampling

    Authors: , , - ICLR 2022 cited by 200

  24. Distributed Optimization with Arbitrary Local Solvers

    Authors: , , , , , , - Optimization methods & software, Optim. Methods Softw. 2017 cited by 195