Skip to content

GPT-3: Language Models are Few-Shot Learners

Authors: Brown et al. Year: 2020 ArXiv/Link: https://arxiv.org/abs/2005.14165

Summary

A 175B parameter language model demonstrating few-shot learning capabilities across diverse tasks without task-specific fine-tuning.

Key Concepts

  • Few-shot learning
  • In-context learning
  • Emergent abilities
  • Task agnostic
  • Large-scale pretraining

Impact

Demonstrated that scale alone enables powerful few-shot learning

Category

GPT Series