Researcher
Tom B. Brown
Papers
-
Language Models are Few-Shot Learners
Showed that scaling a Transformer language model to 175 billion parameters yields few-shot learning: the model performs new tasks from a prompt with no gradient updates. The result that made scale itself the research agenda.
arxiv.org ↗