Reference-based Weak Supervision for Answer Sentence Selection using Web Data
Krishnamurthy, Vivek, Vu, Thuy, Moschitti, Alessandro
–arXiv.org Artificial Intelligence
Answer sentence selection (AS2) modeling requires annotated data, i.e., hand-labeled question-answer pairs. We present a strategy to collect weakly supervised answers for a question based on its reference to improve AS2 modeling. Specifically, we introduce Reference-based Weak Supervision (RWS), a fully automatic large-scale data pipeline that harvests high-quality weakly-supervised answers from abundant Web data requiring only a question-reference pair as input. We study the efficacy and robustness of RWS in the setting of TANDA, a recent state-of-the-art fine-tuning approach specialized for AS2. Our experiments indicate that the produced data consistently bolsters TANDA. We achieve the state of the art in terms of P@1, 90.1%, and MAP, 92.9%, on WikiQA.
arXiv.org Artificial Intelligence
Apr-18-2021
- Country:
- Asia
- Japan > Kyūshū & Okinawa
- Kyūshū > Miyazaki Prefecture > Miyazaki (0.04)
- Singapore (0.05)
- Japan > Kyūshū & Okinawa
- Europe
- Czechia > Prague (0.04)
- Denmark > Capital Region
- Copenhagen (0.04)
- North America
- Oceania > Australia
- Asia
- Genre:
- Research Report (1.00)
- Technology: