Mistral AI releases Mistral Small
Mistral AI releases a closed commercial model designed for lower latency and deployment cost.
SEARCH
Browse published English reports with original sources and content provenance.
Mistral AI releases a closed commercial model designed for lower latency and deployment cost.
Mistral AI releases a larger, higher-performing closed commercial model.
Google DeepMind releases a lightweight open model derived from the Gemini technology architecture.
Google DeepMind releases a model with a new architecture and native support for context windows of up to 1 million tokens.
Google makes Gemini 1.0 Ultra available through Gemini Advanced, moving the largest Gemini 1.0 model from announcement to a consumer product.
DeepSeek releases math models and Group Relative Policy Optimization, creating a direct algorithmic predecessor to R1's reinforcement-learning recipe.
DeepSeekMoE uses fine-grained experts and shared expert isolation to improve sparse-model efficiency, with 16B weights and training code released.
Upstage releases a 10.7-billion-parameter model built with depth up-scaling.
Microsoft releases a small-parameter model demonstrating the value of high-quality, textbook-style training data.
Mistral AI releases an open model using an 8x7B mixture-of-experts architecture that balances capability and inference efficiency.
Google introduces the natively multimodal Gemini 1.0 family—Ultra, Pro and Nano—with Pro and Nano entering products immediately and developer APIs following a week later.
Inflection AI releases its second foundation model with stronger aggregate benchmark performance.
Stability AI releases Stable Video Diffusion as a research preview and opens its code and model weights.
Microsoft releases the second Orca generation to improve reasoning in small-parameter models.
Anthropic released Claude 2.1 with a 200,000-token context window, system prompts, and tool use in beta.
Nous Research releases a conversational model fine-tuned from an open base model by the open-source community.
OpenAI releases TTS-1 HD for higher-quality speech synthesis at its first DevDay.
At its first DevDay, OpenAI opens DALL·E 3 through the Images API, allowing developers to integrate text-to-image generation under the dall-e-3 model name.
OpenAI announced GPT-4 Turbo with a 128K context window and opened the preview to paying API developers, alongside a separate vision preview model.
OpenAI releases the real-time text-to-speech model TTS-1 at its first DevDay.