DeepSeek releases the open-weight reasoning model R1
DeepSeek publishes R1, R1-Zero, six distilled models, and its reinforcement-learning recipe, pushing open reasoning models into the frontier conversation.
SEARCH
English reports are listed here. AI-generated drafts are clearly labeled until an editor reviews them.
DeepSeek publishes R1, R1-Zero, six distilled models, and its reinforcement-learning recipe, pushing open reasoning models into the frontier conversation.
ByteDance Seed launches the Doubao Realtime Voice Model and makes it available in the Doubao app.
DeepSeek puts V3, web search, a reasoning mode, and file analysis into a free mobile app five days before the R1 release.
Mistral AI releases a third-generation lightweight foundation model optimized for performance on limited hardware.
DeepSeek releases a 671B MoE model with 37B active parameters, open weights, low API prices, and unusually detailed efficiency disclosures.
Google DeepMind releases a derivative of the 2.0 Flash architecture with explicit reasoning for logic-heavy tasks.
The Technology Innovation Institute releases a new foundation-model family covering several smaller parameter scales.
Microsoft releases its fourth-generation small-parameter model, using substantial synthetic data to strengthen basic reasoning.
Google DeepMind releases a second-generation lightweight foundation model with faster responses and improved concurrent multitask handling.
OpenAI moves Sora from research preview to a standalone video-generation product for ChatGPT Plus and Pro users.
xAI released Aurora, an in-house autoregressive mixture-of-experts image model trained on interleaved text and image data, through Grok on X.
LG AI Research releases a medium-scale model optimized for Korean-English bilingual processing.
Meta releases a 70-billion-parameter model that consolidates long-context and multilingual capabilities.
OpenAI releases an advanced o1 variant that uses more compute for longer reasoning to improve accuracy.
Amazon AWS releases a new family of multimodal foundation models.
Tencent released HunyuanVideo, a text-to-video model with more than 13 billion parameters, together with model weights and inference code.
Upstage releases a 22-billion-parameter model optimized for document understanding and long-context tasks.
R1-Lite-Preview exposes long-form reasoning on the web and promises an open model and API, showing that R1 was already underway before V3's release.
The Allen Institute for AI releases an open model while keeping its training data and underlying code fully public.
Mistral AI releases a large multimodal model with stronger understanding of visual features.