Share Single-dataset Experts for Multi-dataset Question Answering

Copy link

October 04, 2021

Single-dataset Experts for Multi-dataset Question Answering

20 minutes

Many datasets have been created for training reading comprehension models, and a natural question is whether we can combine them to build models that (1) perform better on all of the training datasets and (2) generalize and transfer better to new datasets. Our approach is to model multi-dataset question answering with a collection of single-dataset experts, by training a collection of lightweight, dataset-specific adapter modules (Houlsby et al., 2019) that share an underlying Transformer model. We find that these Multi-Adapter Dataset Experts (MADE) outperform all our baselines in terms of in-distribution accuracy, and simple methods based on parameter-averaging lead to better zero-shot generalization and few-shot transfer performance, offering a strong and versatile starting point for building new reading comprehension systems.

2021: Dan Friedman, Ben Dodge, Danqi Chen

https://arxiv.org/pdf/2109.13880v1.pdf

...more