Who’s on First?: Probing the Learning and Representation Capabilities of Language Models on Deterministic Closed Domains

David Demeter; Doug Downey

2021 CONLL CoNLL 2021

Who’s on First?: Probing the Learning and Representation Capabilities of Language Models on Deterministic Closed Domains

Abstract

AbstractThe capabilities of today’s natural language processing systems are typically evaluated using large datasets of curated questions and answers. While these are critical benchmarks of progress, they also suffer from weakness due to artificial distributions and incomplete knowledge. Artifacts arising from artificial distributions can overstate language model performance, while incomplete knowledge limits fine-grained analysis. In this work, we introduce a complementary benchmarking approach based on SimPlified Language Activity Traces (SPLAT). SPLATs are corpora of language encodings of activity in some closed domain (we study traces from chess and baseball games in this work). SPLAT datasets use naturally-arising distributions, allow the generation of question-answer pairs at scale, and afford complete knowledge in their closed domains. We show that language models of three different architectures can answer questions about world states using only verb-like encodings of activity. Our approach is extensible to new language models and additional question-answering tasks.

❓ The Questioner

🌉 Interdisciplinary Bridge — Artificial Intelligence and Natural Language Processing

🧭 Keyword Pioneer — closed domain

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics, Security & Privacy, Speech & Audio

Authors

David Demeter , Doug Downey

Topics

Artificial Intelligence > Core AI > Interpretability Natural Language Processing > Applications > Question Answering

Keywords

representation learning question answering closed domain language model probing

Download PDF

Related papers

BabyBERTa: Learning More Grammar With Small-Scale Child-Directed Language 2021

“It’s our fault!”: Insights Into Users’ Understanding and Interaction With an Explanatory Collaborative Dialog System 2021

VQA-MHUG: A Gaze Dataset to Study Multimodal Neural Attention in Visual Question Answering 2021

“It seemed like an annoying woman”: On the Perception and Ethical Considerations of Affective Language in Text-Based Conversational Agents 2021

Generalising to German Plural Noun Classes, from the Perspective of a Recurrent Neural Network 2021