objectiveIntroduced by BART · 2019
Denoising seq2seq pre-training
Corrupt text in many ways and train an encoder–decoder to reconstruct it.
Drafted by AI · not yet reviewed
How this idea evolved
Suggested from citations. No curator has recorded what this idea builds on yet. These ideas from the same theme come from papers that BART cites, directly or one step removed. Citation is a fact; the connection between the ideas is not verified.
Autoregressive training over all factorisation orders, capturing bidirectional context without masks.
Hide some tokens and predict them from context on both sides.
Train BERT longer, on more data, with bigger batches and dynamic masking.
Papers using this
- 2021Prefix-Tuning
- 2021VL-T5
- 2021Prompt tuning
- 2021Natural Instructions