Fine-tuning Language Models for Triple Extraction with Data Augmentation

Yujia Zhang; Tyler Sadler; Mohammad Reza Taesiri; Wenjie Xu; Marek Reformat

2024 ACL ACL 2024

Fine-tuning Language Models for Triple Extraction with Data Augmentation

Abstract

AbstractAdvanced language models with impressive capabilities to process textual information can more effectively extract high-quality triples, which are the building blocks of knowledge graphs. Our work examines language models’ abilities to extract entities and the relationships between them. We use a diverse data augmentation process to fine-tune large language models to extract triples from the text. Fine-tuning is performed using a mix of trainers from HuggingFace and five public datasets, such as different variations of the WebNLG, SKE, DocRed, FewRel, and KELM. Evaluation involves comparing model outputs with test-set triples based on several criteria, such as type, partial, exact, and strict accuracy.The obtained results outperform ChatGPT and even match or exceed the performance of GPT-4.

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Healthcare & Medicine, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics, Security & Privacy, Speech & Audio

Authors

Yujia Zhang , Tyler Sadler , Mohammad Reza Taesiri , Wenjie Xu , Marek Reformat

Topics

Natural Language Processing > Applications > Information Extraction Natural Language Processing > Resources & Methods > Large Language Models

Keywords

data augmentation relation extraction knowledge graph language model fine-tuning entity extraction triple extraction

Download PDF

Related papers

Reinforcement Learning-Driven LLM Agent for Automated Attacks on LLMs 2024

EtymoLink: A Structured English Etymology Dataset 2024

Turkish Delights: A Dataset on Turkish Euphemisms 2024

Subjectivity Detection in English News using Large Language Models 2024

Does DetectGPT Fully Utilize Perturbation? Bridging Selective Perturbation to Fine-tuned Contrastive Learning Detector would be Better 2024