Never Lost in the Middle: Improving Large Language Models via Attention
  Strengthening Question Answering

Dong, Xiaoqun; He, Junqing; Liang, Yuxin; Liu, Yibo; Pan, Kunhao; Song, Zhuoyang; Sun, Qianguo; Wang, Hao; Xie, Zejian; Zhang, Jiaxing; Zhang, Songxin

Never Lost in the Middle: Improving Large Language Models via Attention Strengthening Question Answering

Authors: Xiaoqun Dong
Junqing He
Yuxin Liang
Yibo Liu
Kunhao Pan
Zhuoyang Song
Qianguo Sun
Hao Wang
Zejian Xie
Jiaxing Zhang
Songxin Zhang
Publication date: 15 November 2023
Publisher

Abstract

While large language models (LLMs) are equipped with longer text input capabilities than before, they are struggling to seek correct information in long contexts. The "lost in the middle" problem challenges most LLMs, referring to the dramatic decline in accuracy when correct information is located in the middle. To overcome this crucial issue, this paper proposes to enhance the information searching and reflection ability of LLMs in long contexts via specially designed tasks called Attention Strengthening Multi-doc QA (ASM QA). Following these tasks, our model excels in focusing more precisely on the desired information. Experimental results show substantial improvement in Multi-doc QA and other benchmarks, superior to state-of-the-art models by 13.7% absolute gain in shuffled settings, by 21.5% in passage retrieval task. We release our model, Ziya-Reader to promote related research in the community

Similar works

Full text

Available Versions

arXiv.org e-Print Archive

oai:arXiv.org:2311.09198

Last time updated on 10/02/2024