xAI launches Grok-2 beta
xAI opened an early Grok-2 preview on X with image understanding and planned enterprise API access; image generation came from the external FLUX.1 model.
EVENT ARCHIVE
Browse published English reports with original sources and content provenance.
xAI opened an early Grok-2 preview on X with image understanding and planned enterprise API access; image generation came from the external FLUX.1 model.
Black Forest Labs releases the FLUX.1 text-to-image model family with an API version and open-weight versions under different licenses.
Mistral AI releases its second-generation flagship model with stronger coding, mathematics and multilingual reasoning.
Meta releases a 405-billion-parameter open model with a much longer context window and broader multilingual use.
OpenAI released GPT-4o mini on July 18, 2024, making the small model available through its API and ChatGPT.
Mistral AI and NVIDIA release a model with a new tokenizer designed to improve multilingual processing efficiency.
Shanghai AI Laboratory releases an update that further strengthens long-context reasoning and agent tool use.
Google DeepMind releases the second Gemma generation with architectural improvements that raise performance at the same parameter scale.
Claude 3.5 Sonnet improves reasoning, coding and vision, and introduces Artifacts.
DeepSeek releases a V2 architecture model optimized for programming and mathematical reasoning.
Mistral AI releases a code-generation model built with a non-Transformer Mamba architecture.
NVIDIA releases a 340-billion-parameter open model and supporting materials for synthetic-data generation.
Stability AI released the 2-billion-parameter Stable Diffusion 3 Medium with downloadable model weights and API access.
Alibaba Cloud releases the second Qwen generation with stronger multilingual support and long-text handling.
Z.ai releases the fourth GLM generation with substantially stronger overall benchmark results.
At Google I/O 2024, Google announces Imagen 3 with improvements in photorealism, complex prompt understanding, detail and text rendered in images.
OpenAI releases the native multimodal GPT-4o model for real-time interaction with text, images and audio.
DeepSeek releases a 236B mixture-of-experts model that combines sparse activation with latent attention to cut training and inference costs.
Meta FAIR researchers propose training language models to predict several future tokens at once, work that later informs DeepSeek-V3's MTP design.
Snowflake releases a mixture-of-experts model optimized for SQL queries and enterprise code generation.