Hugging Face releases SmolLM2
Hugging Face releases a small-parameter model family designed to run locally on hardware such as smartphones.
SEARCH
English reports are listed here. AI-generated drafts are clearly labeled until an editor reviews them.
Hugging Face releases a small-parameter model family designed to run locally on hardware such as smartphones.
Mistral AI releases a low-parameter model family for on-device computing and edge workloads.
Z.ai releases an open model with native end-to-end voice interaction and basic emotion recognition.
Inflection AI releases a language model for enterprise customization that balances companionship and productivity use cases.
Writer releases a foundation model focused on enterprise workflow automation and coordinated tool use.
01.AI releases a high-performance closed mixture-of-experts model with strong public benchmark results.
Meta released Llama 3.2 with 1B and 3B lightweight text models and 11B and 90B vision-language models.
Qwen2.5 spans small local models through a 72B flagship, with stronger instruction following and specialist code and math editions.
Mistral AI releases an architecture update and fine-tuned version for low-cost deployment and lightweight workloads.
Mistral AI releases a multimodal vision-language model with native image input.
OpenAI releases o1-preview, shifting frontier-model competition toward reinforcement learning and additional computation at inference time.
AI21 Labs releases an upgraded hybrid-architecture model family with larger scale and context windows of up to 250,000 tokens.
xAI opened an early Grok-2 preview on X with image understanding and planned enterprise API access; image generation came from the external FLUX.1 model.
Black Forest Labs releases the FLUX.1 text-to-image model family with an API version and open-weight versions under different licenses.
Mistral AI releases its second-generation flagship model with stronger coding, mathematics and multilingual reasoning.
Meta releases a 405-billion-parameter open model with a much longer context window and broader multilingual use.
OpenAI released GPT-4o mini on July 18, 2024, making the small model available through its API and ChatGPT.
Mistral AI and NVIDIA release a model with a new tokenizer designed to improve multilingual processing efficiency.
Shanghai AI Laboratory releases an update that further strengthens long-context reasoning and agent tool use.
Google DeepMind releases the second Gemma generation with architectural improvements that raise performance at the same parameter scale.