Meta releases Llama 3.2
Meta released Llama 3.2 with 1B and 3B lightweight text models and 11B and 90B vision-language models.
SEARCH
Browse published English reports with original sources and content provenance.
Meta released Llama 3.2 with 1B and 3B lightweight text models and 11B and 90B vision-language models.
Qwen2.5 spans small local models through a 72B flagship, with stronger instruction following and specialist code and math editions.
Mistral AI releases an architecture update and fine-tuned version for low-cost deployment and lightweight workloads.
Mistral AI releases a multimodal vision-language model with native image input.
OpenAI releases o1-preview, shifting frontier-model competition toward reinforcement learning and additional computation at inference time.
AI21 Labs releases an upgraded hybrid-architecture model family with larger scale and context windows of up to 250,000 tokens.
xAI opened an early Grok-2 preview on X with image understanding and planned enterprise API access; image generation came from the external FLUX.1 model.
Black Forest Labs releases the FLUX.1 text-to-image model family with an API version and open-weight versions under different licenses.
Mistral AI releases its second-generation flagship model with stronger coding, mathematics and multilingual reasoning.
Meta releases a 405-billion-parameter open model with a much longer context window and broader multilingual use.
OpenAI released GPT-4o mini on July 18, 2024, making the small model available through its API and ChatGPT.
Mistral AI and NVIDIA release a model with a new tokenizer designed to improve multilingual processing efficiency.
Shanghai AI Laboratory releases an update that further strengthens long-context reasoning and agent tool use.
Google DeepMind releases the second Gemma generation with architectural improvements that raise performance at the same parameter scale.
ByteDance Seed publishes the Seed-TTS speech-generation foundation model with voice cloning and controllable synthesis capabilities.
Claude 3.5 Sonnet improves reasoning, coding and vision, and introduces Artifacts.
Runway releases Gen-3 Alpha with improvements in video-generation fidelity, temporal consistency, motion and camera control.
DeepSeek releases a V2 architecture model optimized for programming and mathematical reasoning.
Mistral AI releases a code-generation model built with a non-Transformer Mamba architecture.
NVIDIA releases a 340-billion-parameter open model and supporting materials for synthetic-data generation.