Open-source Large Language Models are Strong Zero-shot Query Likelihood Models for Document Ranking

Shengyao Zhuang; Bing Liu; Bevan Koopman; Guido Zuccon

2023 EMNLP EMNLP 2023

Open-source Large Language Models are Strong Zero-shot Query Likelihood Models for Document Ranking

Abstract

AbstractIn the field of information retrieval, Query Likelihood Models (QLMs) rank documents based on the probability of generating the query given the content of a document. Recently, advanced large language models (LLMs) have emerged as effective QLMs, showcasing promising ranking capabilities. This paper focuses on investigating the genuine zero-shot ranking effectiveness of recent LLMs, which are solely pre-trained on unstructured text data without supervised instruction fine-tuning. Our findings reveal the robust zero-shot ranking ability of such LLMs, highlighting that additional instruction fine-tuning may hinder effectiveness unless a question generation task is present in the fine-tuning dataset. Furthermore, we introduce a novel state-of-the-art ranking system that integrates LLM-based QLMs with a hybrid zero-shot retriever, demonstrating exceptional effectiveness in both zero-shot and few-shot scenarios. We make our codebase publicly available at https://github.com/ielab/llm-qlm.

🌉 Interdisciplinary Bridge — Computer Science and Data Science & Analytics and Deep Learning and Machine Learning and Natural Language Processing

🧭 Keyword Pioneer — query likelihood

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics, Security & Privacy, Speech & Audio

Authors

Shengyao Zhuang , Bing Liu , Bevan Koopman , Guido Zuccon

Topics

Machine Learning > Learning Types > Zero-Shot Learning Natural Language Processing > Generation > Language Modeling Natural Language Processing > Applications > Information Retrieval Computer Science > Applications > Information Retrieval Data Science & Analytics > Applications > Information Retrieval Deep Learning > Models > Large Language Models Machine Learning > Learning Paradigms > Zero-Shot Learning

Keywords

zero-shot learning information retrieval document ranking large language model query likelihood model query likelihood

Download PDF

Related papers

Exploring Linguistic Probes for Morphological Generalization 2023

NameGuess: Column Name Expansion for Tabular Data 2023

Vision-Enhanced Semantic Entity Recognition in Document Images via Visually-Asymmetric Consistency Learning 2023

Improving Conversational Recommendation Systems via Bias Analysis and Language-Model-Enhanced Data Augmentation 2023

On the Calibration of Large Language Models and Alignment 2023