On the Binding Problem in Artificial Neural Networks

Contemporary neural networks still fall short of human-level generalization, which extends far beyond our direct experiences. In this paper, we argue that the underlying cause for this shortcoming is their inability to dynamically and flexibly bind information that is distributed throughout the network. This binding problem affects their capacity to acquire a compositional understanding of the world in terms of symbol-like entities (like objects), which is crucial for generalizing in predictable and systematic ways. To address this issue, we propose a unifying framework that revolves around forming meaningful entities from unstructured sensory inputs (segregation), maintaining this separation of information at a representational level (representation), and using these entities to construct new inferences, predictions, and behaviors (composition). Our analysis draws inspiration from a wealth of research in neuroscience and cognitive psychology, and surveys relevant mechanisms from the machine learning literature, to help identify a combination of inductive biases that allow symbolic information processing to emerge naturally in neural networks. We believe that a compositional approach to AI, in terms of grounded symbol-like representations, is of fundamental importance for realizing human-level generalization, and we hope that this paper may contribute towards that goal as a reference and inspiration.

A Model ofSaliency-Based Visual…A Model of Saliency-Based Visual Attention for Rapid Scene AnalysisThe time dimension forscene analysisThe time dimension for scene analysisWhat is an object?What is an object?Recurrent Models ofVisual AttentionRecurrent Models of Visual AttentionNeural Turing MachinesNeural Turing MachinesShow, Attend and Tell:Neural Image Caption…Show, Attend and Tell: Neural Image Caption Generation with Visual AttentionWeakly Supervised MemoryNetworksWeakly Supervised Memory NetworksPicture: A probabilisticprogramming language fo…Picture: A probabilistic programming language for scene perceptionNeural Message Passingfor Quantum ChemistryNeural Message Passing for Quantum ChemistryBERT: Pre-training ofDeep Bidirectional…BERT: Pre-training of Deep Bidirectional Transformers for Language UnderstandingRecurrent IndependentMechanismsRecurrent Independent MechanismsUnderstanding the Impactof Value Selection…Understanding the Impact of Value Selection Heuristics in Scheduling ProblemsZero-Shot Text-to-ImageGenerationZero-Shot Text-to-Image GenerationLinear Transformers AreSecretly Fast Weight…Linear Transformers Are Secretly Fast Weight Memory SystemsSIMONe: View-Invariant,Temporally-Abstracted…SIMONe: View-Invariant, Temporally-Abstracted Object Representations via Unsupervised Video DecompositionConditionalObject-Centric Learning…Conditional Object-Centric Learning from VideoSimple UnsupervisedObject-Centric Learning…Simple Unsupervised Object-Centric Learning for Complex and Naturalistic VideosComplex-ValuedAutoencoders for Object…Complex-Valued Autoencoders for Object DiscoverySAVi++: TowardsEnd-to-End…SAVi++: Towards End-to-End Object-Centric Learning from Real-World VideosLearning with Capsules:A SurveyLearning with Capsules: A SurveySelf-Supervised VisualRepresentation Learning…Self-Supervised Visual Representation Learning with Semantic GroupingBridging the Gap toReal-World…Bridging the Gap to Real-World Object-Centric LearningAn Investigation intoPre-Training…An Investigation into Pre-Training Object-Centric Representations for Reinforcement LearningContrastive Training ofComplex-Valued…Contrastive Training of Complex-Valued Autoencoders for Object DiscoveryOn the Binding Problemin Artificial Neural…On the Binding Problem in Artificial Neural NetworksEarlier referencesFocus paperCiting papersOlderNewer

Click a node to pin it, click the empty canvas to go back to this paper, or hover to preview. Open a node’s page from its title.