Researcher

Tom B. Brown

Papers

  • Language Models are Few-Shot Learners

    Showed that scaling a Transformer language model to 175 billion parameters yields few-shot learning: the model performs new tasks from a prompt with no gradient updates. The result that made scale itself the research agenda.

    arxiv.org ↗