Sharing the latest Model Spec
Related stories
OpenAI GPT-4.5 System Card
We’re releasing a research preview of OpenAI GPT‑4. 5, our largest and most knowledgeable model yet.
Introducing improvements to the fine-tuning API and expanding our custom models program
We’re adding new features to help developers have more control over fine-tuning and announcing new ways to build custom models with OpenAI.
Introducing OpenAI o3 and o4-mini
Our smartest and most capable models to date with full tool access
gpt-oss-120b & gpt-oss-20b Model Card
We introduce gpt-oss-120b and gpt-oss-20b, two open-weight reasoning models available under the Apache 2. 0 license and our gpt-oss usage policy.
GPT-2: 1.5B release
As the final model release of GPT-2’s staged release, we’re releasing the largest version (1. 5B parameters) of GPT-2 along with code and model weights to facilitate detection of outputs of GPT-2 models.
New models and developer products announced at DevDay
GPT-4 Turbo with 128K context and lower prices, the new Assistants API, GPT-4 Turbo with Vision, DALL·E 3 API, and more.
Introducing the Gemini 2.5 Computer Use model
Available in preview via the API, our Computer Use model is a specialized model built on Gemini 2. 5 Pro’s capabilities to power agents that can interact with user interfaces.
GPT-4 API general availability and deprecation of older models in the Completions API
GPT-3. 5 Turbo, DALL·E and Whisper APIs are also generally available, and we are releasing a deprecation plan for older models of the Completions API, which will retire at the beginning of 2024.
On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters
arXiv:2606. 02437v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) is usually treated as a cheaper alternative to full fine-tuning.
LoKA: Low-precision Kernel Applications for Recommendation Models At Scale
arXiv:2605. 10886v3 Announce Type: replace-cross Abstract: Recent GPU generations deliver significantly higher FLOPs using lower-precision arithmetic, such as FP8.
GPT-2: 6-month follow-up
We’re releasing the 774 million parameter GPT-2 language model after the release of our small 124M model in February, staged release of our medium 355M model in May, and subsequent research with partners and the AI community into the model’s potential for misuse and societal benefit. We’re also releasing an open-source legal agreement to make it easier for organizations to initiate model-sharing partnerships with each other, and are publishing a technical report about our experience in coordinating with the wider AI research community on publication norms.