LLM Reference

GPT-1 Models by OpenAI

OpenAIProprietary
This model family is considered obsolete. Consider newer alternatives in Related Model Families below.
1 model2018Up to 512 ctx

Last refreshed 2026-04-15. Next refresh: weekly.

Details

ResearcherOpenAI
LicenseProprietary
Commercial useCommercial use: conditional
Models1
Released2018
Max context512

About

The GPT-1 large language model, introduced by OpenAI in 2018, represented a major leap in natural language processing. As one of the first models to leverage the transformer architecture, GPT-1 employed a decoder-only version, enabling it to generate text that closely mimicked human language based on input prompts. Its pre-training involved a large corpus of text, notably the BooksCorpus, which equipped it with the ability to grasp intricate language patterns and relationships autonomously. However, GPT-1 was also defined by its limitations, such as a modest parameter count of 117 million and a constrained context window, which curtailed its capacity to process long-range dependencies and complex tasks as effectively as its successors. Despite these constraints, GPT-1 set the stage for the evolution of more advanced GPT models that followed, making it a foundational achievement in the field of language models 357.

Decision facts

Best fit
coding
Capability starting point
GPT-1 with 512 context
Lowest tracked input
Not tracked
Closest related family
GPT Realtime 2

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

1 in view
GPT-1Current

Use when the workload needs 512 context and 120M parameters.

2018-06512 context120M parameters

Release Timeline

1 release group
2018-06
1 current
GPT-1
512 context120M parameters
Current

Specifications(1 models)

GPT-1 model specifications comparison
ModelReleasedContextParameters
GPT-12018-06512120M

Popular comparisons in this family