dataIntroduced by OPT · 2022
Open-weights LLMs
Release large LM weights so researchers can study and build on them.
Drafted by AI · not yet reviewed
How this idea evolved
Suggested from citations. No curator has recorded what this idea builds on yet. These ideas from the same theme come from papers that OPT cites, directly or one step removed. Citation is a fact; the connection between the ideas is not verified.
A large model solves a new task from a few examples in its prompt, with no gradient updates.
Fine-tune dialogue models to consult external tools and knowledge for factual answers.
For a fixed budget, grow parameters and training tokens together; most big LMs were undertrained.
Papers using this
- 2022GPT-NeoX-20B
- 2023BLIP-2