Anthropic releases Claude 2.1
Anthropic released Claude 2.1 with a 200,000-token context window, system prompts, and tool use in beta.
EVENT ARCHIVE
English reports are listed here. AI-generated drafts are clearly labeled until an editor reviews them.
Anthropic released Claude 2.1 with a 200,000-token context window, system prompts, and tool use in beta.
Nous Research releases a conversational model fine-tuned from an open base model by the open-source community.
OpenAI releases TTS-1 HD for higher-quality speech synthesis at its first DevDay.
At its first DevDay, OpenAI opens DALL·E 3 through the Images API, allowing developers to integrate text-to-image generation under the dall-e-3 model name.
OpenAI announced GPT-4 Turbo with a 128K context window and opened the preview to paying API developers, alongside a separate vision preview model.
OpenAI releases the real-time text-to-speech model TTS-1 at its first DevDay.
DeepSeek releases code models from 1.3B to 33B parameters, making open weights and code part of its product strategy from the outset.
01.AI releases an open foundation model supporting context windows of up to 200,000 tokens.
ChatGLM3 adds native function calling and code execution to a compact 6B model, shifting the ChatGLM line from conversation toward agent-like workflows.
Hugging Face releases a dialogue model fine-tuned with direct preference optimization to improve conversational alignment.
OpenAI makes DALL·E 3 available to ChatGPT Plus and Enterprise users, who can generate an image in conversation and request further changes.
Baidu launched ERNIE 4.0 on October 17, 2023, with testing available to invited users and enterprise API applicants.
Mistral AI releases a 7-billion-parameter model that uses sliding-window attention and performs strongly on same-scale benchmarks.
Alibaba Cloud releases Qwen as a multilingual open-weight foundation-model family, giving developers a major China-built alternative across model sizes.
OpenAI announces DALL·E 3; the accompanying paper attributes much of its stronger prompt following to recaptioning training images with detailed synthetic descriptions.
Baichuan AI releases its second open model generation with stronger performance on vertical-domain tasks.
WizardLM releases an open code-generation model fine-tuned with the Evol-Instruct method.
Meta releases Code Llama 7B, 13B, and 34B checkpoints based on Llama 2, with base, Python, instruction-tuned, and selected fill-in-the-middle variants.
Naver releases a new Korean-language large model optimized for enterprise applications.
Alibaba Cloud released Qwen-VL and Qwen-VL-Chat and published model weights on ModelScope and Hugging Face.