Meta releases Llama 4
Meta releases its fourth-generation open model family with larger parameter scales and an optimized mixture-of-experts architecture.
EVENT ARCHIVE
Browse published English reports with original sources and content provenance.
Meta releases its fourth-generation open model family with larger parameter scales and an optimized mixture-of-experts architecture.
Google releases an intermediate upgrade to Gemini's native multimodal architecture with stronger benchmark performance.
OpenAI releases a new generation of GPT-4o-based speech-to-text models.
NVIDIA releases an open mixture-of-experts model optimized for high-throughput large-scale data processing.
Cohere releases a new model for enterprise automation workflows and deep API integration.
Google DeepMind releases a new Gemma family generation that adds native multimodal support to the lightweight open-model tier.
Mistral AI releases a minor Mistral Small update that fixes known issues and refines selected features.
Manus asks users for an outcome, then plans, researches, codes, operates websites, and returns finished artifacts from a cloud virtual machine.
Alibaba Cloud releases an open medium-scale model that uses reinforcement learning to improve logical computation and reasoning.
xAI began rolling out Grok-3 beta with Think and DeepSearch; the initial release was limited to Grok products, with API access still pending.
Deepgram releases Nova-3 for real-time and batch transcription, with stronger handling of noise, overlapping speech and word-level timestamps.
OpenAI publicly released o3-mini on January 31, 2025, through ChatGPT and the API.
Hugging Face launches an open community project to reproduce R1 and fill in DeepSeek's unpublished training data and complete training pipeline.
DeepSeek passes ChatGPT on the U.S. App Store's free chart as Nvidia loses roughly $593 billion in market value during an AI-sector selloff.
Alibaba Cloud’s Qwen team released the Qwen2.5-VL vision-language model family in 3B, 7B, and 72B sizes, publishing both base and instruction-tuned models.
Operator uses a visual computer-use model to click, type, and scroll through web tasks, handing control back for logins, payments, and sensitive actions.
DeepSeek publishes R1, R1-Zero, six distilled models, and its reinforcement-learning recipe, pushing open reasoning models into the frontier conversation.
ByteDance Seed launches the Doubao Realtime Voice Model and makes it available in the Doubao app.
DeepSeek puts V3, web search, a reasoning mode, and file analysis into a free mobile app five days before the R1 release.
Mistral AI releases a third-generation lightweight foundation model optimized for performance on limited hardware.