Designing and Interpreting Probes with Control Tasks

Probes, supervised models trained to predict properties (like parts-of-speech) from representations (like ELMo), have achieved high accuracy on a range of linguistic tasks. But does this mean that the representations encode linguistic structure or just that the probe has learned the linguistic task? In this paper, we propose control tasks, which associate word types with random outputs, to complement linguistic tasks. By construction, these tasks can only be learned by the probe itself. So a good probe, (one that reflects the representation), should be selective, achieving high linguistic task accuracy and low control task accuracy. The selectivity of a probe puts linguistic task accuracy in context with the probe's capacity to memorize from word types. We construct control tasks for English part-of-speech tagging and dependency edge prediction, and show that popular probes on ELMo representations are not selective. We also find that dropout, commonly used to control probe complexity, is ineffective for improving selectivity of MLPs, but that other forms of regularization are effective. Finally, we find that while probes on the first layer of ELMo yield slightly better part-of-speech tagging accuracy than the second, probes on the second layer are substantially more selective, which raises the question of which layer better represents parts-of-speech.

openalex_id:w3102226577openalex_id:w3102226577Does String-Based NeuralMT Learn Source Syntax?Does String-Based Neural MT Learn Source Syntax?Probing for semanticevidence of composition…Probing for semantic evidence of composition by means of simple classification tasksFine-grained Analysis ofSentence Embeddings…Fine-grained Analysis of Sentence Embeddings Using Auxiliary Prediction TasksEvaluating Layers ofRepresentation in Neura…Evaluating Layers of Representation in Neural Machine Translation on Part-of-Speech and Semantic Tagging TasksDissecting ContextualWord Embeddings…Dissecting Contextual Word Embeddings: Architecture and RepresentationQualitative SpatialReasoning over Question…Qualitative Spatial Reasoning over Questions (Short Paper)What you can cram into asingle $&!#* vector…What you can cram into a single $&!#* vector: Probing sentence embeddings for linguistic propertiesLanguage ModelingTeaches You More Syntax…Language Modeling Teaches You More Syntax than Translation Does: Lessons Learned Through Auxiliary Task AnalysisA Structural Probe forFinding Syntax in Word…A Structural Probe for Finding Syntax in Word RepresentationsProbing What DifferentNLP Tasks Teach Machine…Probing What Different NLP Tasks Teach Machines about Function Word ComprehensionBERT: Pre-training ofDeep Bidirectional…BERT: Pre-training of Deep Bidirectional Transformers for Language UnderstandingInformation-TheoreticProbing for Linguistic…Information-Theoretic Probing for Linguistic StructureIntrinsic Probingthrough Dimension…Intrinsic Probing through Dimension SelectionOn the Systematicity ofProbing Contextualized…On the Systematicity of Probing Contextualized Word Representations: The Case of Hypernymy in BERTContext Analysis forPre-trained Masked…Context Analysis for Pre-trained Masked Language ModelsConditional probing:measuring usable…Conditional probing: measuring usable information beyond a baselineWhen Do You NeedBillions of Words of…When Do You Need Billions of Words of Pretraining Data?What Artificial NeuralNetworks Can Tell Us…What Artificial Neural Networks Can Tell Us About Human Language AcquisitionProbing with Noise:Unpicking the Warp and…Probing with Noise: Unpicking the Warp and Weft of EmbeddingsThe ArchitecturalBottleneck PrincipleThe Architectural Bottleneck PrincipleLearning ChessBlindfolded: Evaluating…Learning Chess Blindfolded: Evaluating Language Models on State TrackingHelping Cancer Patientsto Choose the Best…Helping Cancer Patients to Choose the Best Treatment: Towards Automated Data-Driven and Personalized Information Presentation of Cancer Treatment OptionsLinguisticInterpretability of…Linguistic Interpretability of Transformer-based Language Models: a systematic reviewDesigning andInterpreting Probes wit…Designing and Interpreting Probes with Control TasksEarlier referencesFocus paperCiting papersOlderNewer

Click a node to pin it, click the empty canvas to go back to this paper, or hover to preview. Open a node’s page from its title.