Pre-training Multi-task Contrastive Learning Models for Scientific Literature Understanding

Yu Zhang; Hao Cheng; Zhihong Shen; Xiaodong Liu; Ye-Yi Wang; Jianfeng Gao

2023 EMNLP EMNLP 2023

Pre-training Multi-task Contrastive Learning Models for Scientific Literature Understanding

Abstract

AbstractScientific literature understanding tasks have gained significant attention due to their potential to accelerate scientific discovery. Pre-trained language models (LMs) have shown effectiveness in these tasks, especially when tuned via contrastive learning. However, jointly utilizing pre-training data across multiple heterogeneous tasks (e.g., extreme multi-label paper classification, citation prediction, and literature search) remains largely unexplored. To bridge this gap, we propose a multi-task contrastive learning framework, SciMult, with a focus on facilitating common knowledge sharing across different scientific literature understanding tasks while preventing task-specific skills from interfering with each other. To be specific, we explore two techniques – task-aware specialization and instruction tuning. The former adopts a Mixture-of-Experts Transformer architecture with task-aware sub-layers; the latter prepends task-specific instructions to the input text so as to produce task-aware outputs. Extensive experiments on a comprehensive collection of benchmark datasets verify the effectiveness of our task-aware specialization strategy, where we outperform state-of-the-art scientific pre-trained LMs. Code, datasets, and pre-trained models can be found at https://scimult.github.io/.

🌉 Interdisciplinary Bridge — Deep Learning and Machine Learning and Natural Language Processing

🧭 Keyword Pioneer — multi-task contrastive learning

🐣 Hot Topic Early Bird — scientific literature

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics, Security & Privacy, Speech & Audio

Authors

Yu Zhang , Hao Cheng , Zhihong Shen , Xiaodong Liu , Ye-Yi Wang , Jianfeng Gao

Topics

Machine Learning > Learning Types > Contrastive Learning Natural Language Processing > Resources & Methods > Large Language Models Machine Learning > Learning Types > Multi-Task Learning Deep Learning > Learning Types > Contrastive Learning Deep Learning > Learning Types > Multi-Task Learning

Keywords

contrastive learning multi-task learning text representation instruction tuning language model mixture of expert pre-trained language model citation prediction scientific literature extreme multi-label classification multi-task contrastive learning scientific literature understanding

Download PDF

Related papers

Exploring Linguistic Probes for Morphological Generalization 2023

NameGuess: Column Name Expansion for Tabular Data 2023

Vision-Enhanced Semantic Entity Recognition in Document Images via Visually-Asymmetric Consistency Learning 2023

Improving Conversational Recommendation Systems via Bias Analysis and Language-Model-Enhanced Data Augmentation 2023

On the Calibration of Large Language Models and Alignment 2023