NVIDIA releases Nemotron Ultra
NVIDIA releases an open mixture-of-experts model optimized for high-throughput large-scale data processing.
EVENT ARCHIVE
20 English reports are shown on this page. Refine the keywords, date range, company, or event type below.
NVIDIA releases an open mixture-of-experts model optimized for high-throughput large-scale data processing.
Cohere releases a new model for enterprise automation workflows and deep API integration.
Google DeepMind releases a new Gemma family generation that adds native multimodal support to the lightweight open-model tier.
Mistral AI releases a minor Mistral Small update that fixes known issues and refines selected features.
Manus asks users for an outcome, then plans, researches, codes, operates websites, and returns finished artifacts from a cloud virtual machine.
Alibaba Cloud releases an open medium-scale model that uses reinforcement learning to improve logical computation and reasoning.
OpenAI releases an improved GPT-4 model with stronger knowledge, generalization and safety performance.
Inception Labs releases a foundation model built around an alternative architecture and network design.
Alibaba's Wan team releases inference code and model weights for Wan2.1, covering text-to-video, image-to-video and other video tasks.
Anthropic releases an upgrade that can switch between standard responses and extended reasoning mode.
xAI began rolling out Grok-3 beta with Think and DeepSearch; the initial release was limited to Grok products, with API access still pending.
Perplexity fine-tunes its internal core model for its own AI search and precise question-answering workflows.
OpenAI publicly released o3-mini on January 31, 2025, through ChatGPT and the API.
Hugging Face launches an open community project to reproduce R1 and fill in DeepSeek's unpublished training data and complete training pipeline.
DeepSeek passes ChatGPT on the U.S. App Store's free chart as Nvidia loses roughly $593 billion in market value during an AI-sector selloff.
Alibaba Cloud’s Qwen team released the Qwen2.5-VL vision-language model family in 3B, 7B, and 72B sizes, publishing both base and instruction-tuned models.
Operator uses a visual computer-use model to click, type, and scroll through web tasks, handing control back for logins, payments, and sensitive actions.
DeepSeek publishes R1, R1-Zero, six distilled models, and its reinforcement-learning recipe, pushing open reasoning models into the frontier conversation.
ByteDance Seed launches the Doubao Realtime Voice Model and makes it available in the Doubao app.
DeepSeek puts V3, web search, a reasoning mode, and file analysis into a free mobile app five days before the R1 release.