Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Generation-Augmented Retrieval for Open-domain Question Answering

Sep 17, 2020

Yuning Mao, Pengcheng He, Xiaodong Liu, Yelong Shen, Jianfeng Gao, Jiawei Han, Weizhu Chen

Figure 1 for Generation-Augmented Retrieval for Open-domain Question Answering

Figure 2 for Generation-Augmented Retrieval for Open-domain Question Answering

Figure 3 for Generation-Augmented Retrieval for Open-domain Question Answering

Figure 4 for Generation-Augmented Retrieval for Open-domain Question Answering

Share this with someone who'll enjoy it:

Abstract:Conventional sparse retrieval methods such as TF-IDF and BM25 are simple and efficient, but solely rely on lexical overlap and fail to conduct semantic matching. Recent dense retrieval methods learn latent representations to tackle the lexical mismatch problem, while being more computationally expensive and sometimes insufficient for exact matching as they embed the entire text sequence into a single vector with limited capacity. In this paper, we present Generation-Augmented Retrieval (GAR), a query expansion method that augments a query with relevant contexts through text generation. We demonstrate on open-domain question answering (QA) that the generated contexts significantly enrich the semantics of the queries and thus GAR with sparse representations (BM25) achieves comparable or better performance than the current state-of-the-art dense method DPR \cite{karpukhin2020dense}. We show that generating various contexts of a query is beneficial as fusing their results consistently yields a better retrieval accuracy. Moreover, GAR achieves the state-of-the-art performance of extractive QA on the Natural Questions and TriviaQA datasets when equipped with an extractive reader.

View paper on

Share this with someone who'll enjoy it:

Title:Generation-Augmented Retrieval for Open-domain Question Answering

Paper and Code