Satinder Singh
1965–2026 年に発表
- 297
- 論文数
- 32,955
- 被引用数
- 74
- h 指数
- 205
- i10 指数
被引用数
引用元
国・地域
機関
分野
- Computer Science67.7%
- Engineering11.9%
- Decision Sciences6.6%
- Neuroscience2.9%
- Psychology2.1%
- Medicine1.9%
- その他6.9%
トピック
- Reinforcement Learning in Robotics15.5%
- Advanced Bandit Algorithms Research3%
- Evolutionary Algorithms and Applications2.8%
- Topic Modeling2.4%
- Robot Manipulation and Learning2.1%
- Adaptive Dynamic Programming Control2%
- その他72.2%
共著者
- Richard L. Lewis33
- Edmund H. Durfee19
- Tom Zahavy19
- Michael J. Kearns18
- Michael P. Wellman17
- Doina Precup16
- Junhyuk Oh16
- Georges Kaddoum14
- Honglak Lee13
- Richard S. Sutton13
- David Silver12
- Vivek Veeriah12
- Michael L. Littman11
- Nan Jiang11
- André Barreto10
- Hado van Hasselt10
- Sahil Garg10
- Yevgeniy Vorobeychik10
- Peter Stone9
- Qi Zhang9
- Vishal Soni9
- Jonathan Sorg8
- Matteo Hessel8
- Michael R. James8
全論文
- Between MDPs and Semi-MDPs: A Framework for Temporal Abstraction in Reinforcement Learning
著者: Richard S. Sutton, Doina Precup, Satinder Singh - Artificial Intelligence, Artif. Intell. 1999 被引用: 3,146
- Policy Gradient Methods for Reinforcement Learning with Function Approximation
著者: Richard S. Sutton, David A. McAllester, Satinder Singh, Yishay Mansour - http://www.cis.upenn.edu/~mkearns/finread/Sutton.pdf, NIPS 1999 被引用: 7,816
- Reward is enough
著者: David Silver, Satinder Singh, Doina Precup, Richard S. Sutton - Artificial Intelligence, Artif. Intell. 2021 被引用: 385
- Genie: Generative Interactive Environments
著者: Jake Bruce, Michael D. Dennis, Ashley Edwards, Jack Parker-Holder, Yuge Shi, Edward Hughes, Matthew Lai, Aditi Mavalankar, Richie Steigerwald, Chris Apps, Yusuf Aytar, Sarah Bechtle, Feryal M. P. Behbahani, Stephanie C. Y. Chan, Nicolas Heess, Lucy Gonzalez, Simon Osindero, Sherjil Ozair, Scott E. Reed, Jingwei Zhang, Konrad Zolna, Jeff Clune, Nando de Freitas, Satinder Singh, Tim Rocktäschel - ICML 2024 被引用: 709
- In-context Reinforcement Learning with Algorithm Distillation
著者: Michael Laskin, Luyu Wang, Junhyuk Oh, Emilio Parisotto, Stephen Spencer, Richie Steigerwald, DJ Strouse, Steven Stenberg Hansen, Angelos Filos, Ethan Brooks, Maxime Gazeau, Himanshu Sahni, Satinder Singh, Volodymyr Mnih - ICLR 2023 被引用: 212
- Blockchain and Deep Learning for Secure Communication in Digital Twin Empowered Industrial IoT Network
著者: Prabhat Kumar, Randhir Kumar, Abhinav Kumar, A. Antony Franklin, Sahil Garg, Satinder Singh - IEEE Transactions on Network Science and Engineering, IEEE Trans. Netw. Sci. Eng. 2022 被引用: 109
- Convergence Results for Single-Step On-Policy Reinforcement-Learning Algorithms
著者: Satinder Singh, Tommi S. Jaakkola, Michael L. Littman, Csaba Szepesvári - Machine Learning, Mach. Learn. 1998 被引用: 625
- Near-Optimal Reinforcement Learning in Polynomial Time
著者: Michael J. Kearns, Satinder Singh - Machine Learning, Mach. Learn. 1998 被引用: 858
- Computational Rationality: Linking Mechanism and Behavior Through Bounded Utility Maximization
著者: Richard L. Lewis, Andrew Howes, Satinder Singh - Topics in Cognitive Science, Top. Cogn. Sci. 2014 被引用: 233
- Structured State Space Models for In-Context Reinforcement Learning
著者: Chris Lu, Yannick Schroecker, Albert Gu, Emilio Parisotto, Jakob N. Foerster, Satinder Singh, Feryal M. P. Behbahani - Advances in Neural Information Processing Systems 36, NeurIPS 2023 被引用: 152
- Communication-Efficient Personalized Federated Meta-Learning in Edge Networks
著者: Feng Yu, Hui Lin, Xiaoding Wang, Sahil Garg, Georges Kaddoum, Satinder Singh, Mohammad Mehedi Hassan - IEEE Transactions on Network and Service Management, IEEE Trans. Netw. Serv. Manag. 2023 被引用: 34
- Mastering Board Games by External and Internal Planning with Language Models
著者: John Schultz, Jakub Adámek, Matej Jusup, Marc Lanctot, Michael Kaisers, Sarah Perrin, Daniel Hennes, Jeremy Shar, Cannada A. Lewis, Anian Ruoss, Tom Zahavy, Petar Velickovic, Laurel Prince, Satinder Singh, Eric Malmi, Nenad Tomasev - ICML 2025 被引用: 31
- Intrinsically Motivated Reinforcement Learning: An Evolutionary Perspective
著者: Satinder Singh, Richard L. Lewis, Andrew G. Barto, Jonathan Sorg - IEEE Transactions on Autonomous Mental Development, IEEE Trans. Auton. Ment. Dev. 2010 被引用: 412
- Behaviour Suite for Reinforcement Learning
著者: Ian Osband, Yotam Doron, Matteo Hessel, John Aslanides, Eren Sezener, Andre Saraiva, Katrina McKinney, Tor Lattimore, Csaba Szepesvári, Satinder Singh, Benjamin Van Roy, Richard S. Sutton, David Silver, Hado van Hasselt - ICLR 2020 被引用: 212
- Intrinsically Motivated Reinforcement Learning
著者: Satinder Singh, Andrew G. Barto, Nuttapong Chentanez - OSD or Non-Service DoD Agency, NIPS 2004 被引用: 639
- Softwarized Resource Management and Allocation With Autonomous Awareness for 6G-Enabled Cooperative Intelligent Transportation Systems
著者: Haotong Cao, Sahil Garg, Georges Kaddoum, Satinder Singh, M. Shamim Hossain - IEEE Transactions on Intelligent Transportation Systems, IEEE Trans. Intell. Transp. Syst. 2022 被引用: 33
- Action-Conditional Video Prediction using Deep Networks in Atari Games
著者: Junhyuk Oh, Xiaoxiao Guo, Honglak Lee, Richard L. Lewis, Satinder Singh - http://www-personal.umich.edu/%7Erickl/pubs/oh-et-al-2015-NIPS.pdf 2015 被引用: 909
- Diversifying AI: Towards Creative Chess with AlphaZero
著者: Tom Zahavy, Vivek Veeriah, Shaobo Hou, Kevin Waugh, Matthew Lai, Edouard Leurent, Nenad Tomasev, Lisa Schut, Demis Hassabis, Satinder Singh - arXiv (Cornell University), CoRR 2023 被引用: 16
- A Definition of Continual Reinforcement Learning
著者: David Abel, André Barreto, Benjamin Van Roy, Doina Precup, Hado Philip van Hasselt, Satinder Singh - Advances in Neural Information Processing Systems 36, NeurIPS 2023 被引用: 150
- Discovering Policies with DOMiNO: Diversity Optimization Maintaining Near Optimality
著者: Tom Zahavy, Yannick Schroecker, Feryal M. P. Behbahani, Kate Baumli, Sebastian Flennerhag, Shaobo Hou, Satinder Singh - ICLR 2023 被引用: 49
- Eligibility Traces for Off-Policy Policy Evaluation
著者: Doina Precup, Richard S. Sutton, Satinder Singh - ICML 2000 被引用: 932
- Self-Imitation Learning
著者: Junhyuk Oh, Yijie Guo, Satinder Singh, Honglak Lee - International Conference on Machine Learning, ICML 2018 被引用: 308
- Discovering Reinforcement Learning Algorithms
著者: Junhyuk Oh, Matteo Hessel, Wojciech M. Czarnecki, Zhongwen Xu, Hado van Hasselt, Satinder Singh, David Silver - Neural Information Processing Systems, NeurIPS 2020 被引用: 150
- Code World Models for General Game Playing
著者: Wolfgang Lehrach, Daniel Hennes, Miguel Lázaro-Gredilla, Xinghua Lou, Carter Wendelken, Zun Li, Antoine Dedieu, Jordi Grau-Moya, Marc Lanctot, Atil Iscen, John Schultz, Marcus Chiam, Ian Gemp, Piotr Zielinski, Satinder Singh, Kevin P. Murphy - ArXiv.org, CoRR 2025 被引用: 12
