Back to library

AI / Technology

Language Models are Few-Shot Learners

Three key questions about this paper

What problem does Language Models are Few-Shot Learners address?

Tom B. Brown et al. · OpenAI · 2020 · arXiv:2005.14165

What evidence supports the main claim in Language Models are Few-Shot Learners?

Across 42 benchmarks reported with accuracy, aggregate zero-shot performance rises with model size and few-shot performance rises faster. Concrete results show both the promise and the spread: on closed-book TriviaQA, GPT-3 scores 64.3% zero-shot, 68.0% one-shot, and 71.2% few-shot; on CoQA reading comprehension it scores 81.5, 84.0, and 85.0 F1. The latter remains below the fine-tuned state of the art at 90.7 F1.

What limitation should readers know about Language Models are Few-Shot Learners?

The qualitative probes are also revealing. Given demonstrations, GPT-3 can unscramble words, perform some arithmetic, use a newly defined nonsense word in a sentence, and correct grammar. These examples show fast pattern use, but they are not proof of a general reasoning engine. Source: §3.9, pp. 22–29

2 new free reports left todaySubscribe to Pro for unlimited reading and 10 new paper explanations each month.Upgrade to Pro