Chronicle

Chronicle

Every dated event on the line, newest first. Filter by branch or milestone.

Model release archive

From GPT and BERT to modern text, multimodal, image, video and audio models. The archive covers public releases and major research previews; every link points to the team's primary announcement, documentation or paper.

2026-09-03
Multimodal · OpenAI
GPT-6 Astra

Announced for complex work, coding and computer use, with access expanding in phases.

2026-09-02
Multimodal · Google
Gemini 3.8 Flash

GA Flash release for coding, agents and complex workflows, with text output and multimodal input.

2026-09-02
Multimodal · Meta
Muse Spark 1.3

Meta updated its long-horizon agentic and coding model, available in Muse Code and Meta Model API.

NEWverifiedMeta · release
2026-09-01
Multimodal · Anthropic
Claude Fable 5.1

An update for coding and long projects in paid Claude products and the API. Account for safeguards and fallback in evaluations.

2026-08-14
Text / LLM · Z.ai
GLM-5.3

Z.ai released a post-training update to the GLM-5.2 base with major gains in complex coding, terminal tasks and long-horizon agentic work.

NEWverifiedZ.ai · release
2026-08-13
Multimodal · Google
Gemini 3.7 Flash

Google released the GA version of a fast reasoning model with a one-million-token context and improvements for coding, web development and agentic workflows.

2026-08-13
Text / LLM · DeepSeek
DeepSeek-V4-Pro-0813

The final V4-Pro checkpoint superseded the preview, substantially improved agentic tasks and shipped with open weights under the MIT license.

2026-08-12
Multimodal · xAI
Grok 4.6

Grok 4.6 improved long-running agents, coding and interactive product work, launching in Cursor, Grok Build, the API and partner platforms.

NEWverifiedxAI · release
2026-08-10
Multimodal · Meta
Muse Glimmer 30B

Meta released a local 30B agentic model under Apache 2.0 with image perception, tool use, failure recovery and official GGUF builds.

NEWverifiedMeta · release
2026-08-05
Multimodal · Meta
Muse Spark 1.2

Meta updated its proprietary coding and agentic model Muse Spark, with version 1.2 improving multi-file development, computer use and long-running workflows.

2026-07-27
Text / LLM · Moonshot AI
Kimi K3

Moonshot AI published Kimi K3 weights with instructions for Transformers, vLLM and SGLang. It is an infrastructure-scale open-weights release: 2.8T parameters does not imply an ordinary local run.

2026-07-23
Multimodal · Black Forest Labs
FLUX 3

Black Forest Labs opened Early Access to a multimodal foundation for images, video, audio and action prediction, with individual capabilities rolling out in stages.

2026-07-09
Multimodal · OpenAI
GPT-5.6 Sol

OpenAI's flagship strengthened agentic work in coding, biology and cybersecurity.

2026-06-01
Text / LLM · MiniMax
MiniMax M3

A million-context model focused on lower compute cost and faster prefill and decoding.

2026-03-18
Text / LLM · OpenAI
GPT-5.4 mini

A compact reasoning model became the fast fallback for GPT-5.4 Thinking in ChatGPT.

2026-02-05
Text / LLM · OpenAI
GPT-5.3-Codex

A specialized agentic model combined Codex and GPT-5 training for long-running coding work.

2026-01-13
Video · Google
Veo 3.1

The update improved subject consistency, vertical video and output up to 4K.

2025-08-07
Multimodal · OpenAI
GPT-5

OpenAI unified fast answers, deeper reasoning and routing between modes in one product system.

2025-08-04
Image · Alibaba
Qwen-Image

A 20B MMDiT model focused on complex typography and precise text editing inside images.

2025-05-22
Multimodal · Anthropic
Claude 4

Opus 4 and Sonnet 4 improved long-running coding, tool use and agentic workflows.

2025-05-20
Video · Google
Veo 3

The video model added native sound, speech and ambience generation alongside visuals.

2025-05-20
Audio · Google
Lyria 2

Google expanded access to its music model and integrated it into creator tools.

2025-04-29
Text / LLM · Alibaba
Qwen3

An open family combined normal and thinking modes across dense and MoE sizes.

2025-04-05
Multimodal · Meta
Llama 4

The natively multimodal Scout and Maverick MoE family expanded context and image understanding.

2025-03-25
Multimodal · Google
Gemini 2.5 Pro

Google's thinking model combined reasoning, code and multimodal understanding with long context.

2025-02-27
Text / LLM · OpenAI
GPT-4.5

OpenAI's largest chat model became a research preview of scaled pretraining without a separate reasoning mode.

2025-01-26
Multimodal · Alibaba
Qwen2.5-VL

A vision-language family learned documents, interfaces, charts and long video sequences.

2025-01-20
Text / LLM · DeepSeek
DeepSeek-R1

An open reasoning model and its distillations showed reinforcement learning competing with closed reasoning systems.

2024-12-26
Text / LLM · DeepSeek
DeepSeek-V3

An open MoE model with 671B total and 37B active parameters sharply reduced frontier-model cost.

2024-12-11
Multimodal · Google
Gemini 2.0

Google emphasized native tool use, streaming multimodal interaction and future agents.

2024-12-09
Video · OpenAI
Sora Turbo

Sora moved from research preview into a product for generating and editing video.

2024-09-25
Multimodal · Meta
Llama 3.2

The family added 11B/90B vision models and small 1B/3B text models for edge devices.

2024-09-19
Text / LLM · Alibaba
Qwen2.5

An open 0.5B-to-72B family improved multilingual, coding, math and structured-output capabilities.

2024-09-12
Text / LLM · OpenAI
OpenAI o1

The model spent dedicated compute before answering, improving math, code and complex reasoning.

2024-07-23
Text / LLM · Meta
Llama 3.1 405B

Meta released a 405B model with 128K context and allowed its outputs to improve other models.

2024-05-14
Video · Google
Veo

Google introduced a 1080p video model with cinematic styles and longer scenes.

2024-05-14
Image · Google
Imagen 3

A new Imagen generation improved photorealism, detail and visual consistency.

2024-05-13
Multimodal · OpenAI
GPT-4o

A single model handled text, vision and speech with low latency.

2024-04-18
Text / LLM · Meta
Llama 3

Open 8B and 70B models raised the quality of dialogue, reasoning and coding.

2024-03-04
Multimodal · Anthropic
Claude 3

Haiku, Sonnet and Opus added vision and separated the family by speed, price and peak capability.

2024-02-15
Multimodal · Google
Gemini 1.5

The model reached a million-token context across long documents, code, audio and video.

2023-12-11
Text / LLM · Mistral AI
Mixtral 8×7B

An open mixture-of-experts model activated only part of its parameters per token, reducing inference cost.

2023-12-06
Multimodal · Google
Gemini 1.0

Google introduced a family trained from the start across text, images, audio and video.

2023-09-27
Text / LLM · Mistral AI
Mistral 7B

A compact open model showed that architecture and data quality could compete with larger systems.

2023-09-20
Image · OpenAI
DALL·E 3

Image generation moved into ChatGPT with much stronger adherence to complex prompts.

2023-07-18
Text / LLM · Meta
Llama 2

Open weights became available for research and most commercial uses.

2023-07-11
Text / LLM · Anthropic
Claude 2

Claude gained a 100K context window, stronger coding and broad public availability.

2023-03-14
Text / LLM · Anthropic
Claude

The first public Claude offered long-form dialogue and a distinct approach to safety training.

2023-03-14
Multimodal · OpenAI
GPT-4

OpenAI's flagship added image input, stronger reasoning and more reliable instruction following.

2023-02-24
Text / LLM · Meta
LLaMA

Meta released a 7B-to-65B research family and accelerated the open-weight movement.

2022-11-30
Text / LLM · OpenAI
ChatGPT

A conversational interface and human-feedback training turned large language models into a mass-market product.

2022-09-21
Audio · OpenAI
Whisper

An open speech-recognition model combined multilingual transcription, translation and noise robustness.

2022-05-23
Image · Google
Imagen

A large text encoder and cascaded diffusion models improved prompt fidelity.

2021-08-10
Text / LLM · OpenAI
Codex

A model trained on code translated natural language into programs and powered the first GitHub Copilot.

2021-01-05
Image · OpenAI
DALL·E

A large GPT-family model turned text instructions into generated images.

2021-01-05
Multimodal · OpenAI
CLIP

A shared image-text space enabled recognition of new visual categories without task-specific training.

2020-05-28
Text / LLM · OpenAI
GPT-3

At 175B parameters, in-context learning became a visible capability of a large language model.

2019-10-23
Text / LLM · Google
T5

Google reframed many language tasks as a single text-to-text problem.

2019-02-14
Text / LLM · OpenAI
GPT-2

Scaling a language model produced coherent long-form text and early convincing zero-shot transfer.

2018-11-02
Text / LLM · Google
BERT

Bidirectional pretraining exposed both left and right context and reshaped search, classification and question answering.

2018-06-11
Text / LLM · OpenAI
GPT

The first GPT showed that one pretrained Transformer could be adapted to several language tasks.