DeepSeek releases the efficiency-focused DeepSeek-V2
DeepSeek releases a 236B mixture-of-experts model that combines sparse activation with latent attention to cut training and inference costs.
MODEL
Official links | Chat | Code Plan | Agent tools | API |
|---|---|---|---|---|
| Global | — | — | ||
| Mainland China | — | — |
English reports linked to this entity will appear here as their translations are published.