DeepSeek-V3 brings an efficiency-first open model to the frontier race
DeepSeek releases a 671B MoE model with 37B active parameters, open weights, low API prices, and unusually detailed efficiency disclosures.
EVENT ARCHIVE
Browse published English reports with original sources and content provenance.
DeepSeek releases a 671B MoE model with 37B active parameters, open weights, low API prices, and unusually detailed efficiency disclosures.
The Technology Innovation Institute releases a new foundation-model family covering several smaller parameter scales.
Microsoft releases its fourth-generation small-parameter model, using substantial synthetic data to strengthen basic reasoning.
OpenAI moves Sora from research preview to a standalone video-generation product for ChatGPT Plus and Pro users.
xAI released Aurora, an in-house autoregressive mixture-of-experts image model trained on interleaved text and image data, through Grok on X.
LG AI Research releases a medium-scale model optimized for Korean-English bilingual processing.
Meta releases a 70-billion-parameter model that consolidates long-context and multilingual capabilities.
Tencent released HunyuanVideo, a text-to-video model with more than 13 billion parameters, together with model weights and inference code.
R1-Lite-Preview exposes long-form reasoning on the web and promises an open model and API, showing that R1 was already underway before V3's release.
The Allen Institute for AI releases an open model while keeping its training data and underlying code fully public.
Mistral AI releases a large multimodal model with stronger understanding of visual features.
Hugging Face releases a small-parameter model family designed to run locally on hardware such as smartphones.
Mistral AI releases a low-parameter model family for on-device computing and edge workloads.
OpenAI opens the Realtime API public beta for low-latency, native speech-to-speech interaction.
Meta released Llama 3.2 with 1B and 3B lightweight text models and 11B and 90B vision-language models.
Qwen2.5 spans small local models through a 72B flagship, with stronger instruction following and specialist code and math editions.
Mistral AI releases an architecture update and fine-tuned version for low-cost deployment and lightweight workloads.
Mistral AI releases a multimodal vision-language model with native image input.
OpenAI releases o1-preview, shifting frontier-model competition toward reinforcement learning and additional computation at inference time.
AI21 Labs releases an upgraded hybrid-architecture model family with larger scale and context windows of up to 250,000 tokens.