Better language models and their implications
OpenAI Blog
OpenAI trained a large-scale unsupervised language model that generates coherent text and achieves state-of-the-art performance on multiple language modeling benchmarks. The model performs reading comprehension, machine translation, question answering, and summarization tasks without any task-specific training. This demonstrates that a single general-purpose model can handle diverse language tasks, reducing the need for separate models trained for individual applications.
Why it matters
We’ve trained a large-scale unsupervised language model which generates coherent paragraphs of text, achieves state-of-the-art performance on many language modeling benchmarks, and performs rudimentary reading comprehension, machine translation, question answering, and summarization—all without task-specific training.