deepseek

Skip to main content

Tag: deepseek

A laptop on a desk displays the DeepSeek interface with model selection and input field, hands resting on the keyboard.

DeepSeek-V4 Preview Now Live and Open-Sourced

DeepSeek has officially launched DeepSeek-V4, featuring two model variants with 1M context length support. Both are available now at chat.deepseek.com in Expert Mode and Instant Mode, with APIs updated and available today. Model Specifications Model Total Parameters Active Parameters DeepSeek-V4-Pro 1.6 trillion 49 billion DeepSeek-V4-Flash 284 billion 13 billion DeepSeek-V4-Pro contains 1.6 trillion total parameters with 49 billion active parameters. Its performance rivals the world's top closed-source models. DeepSeek-V4-Flash contains 284 billion total parameters with 13...

Continue reading

Sparkling blue particles form the text "V4.1" against a dark, smoky background.

DeepSeek Introduces V4.1-Flash: Smaller Model with Native Visual Understanding

DeepSeek has released V4.1-Flash, the smallest model in its new architecture family, featuring native visual understanding and designed for greater capability, faster inference, and higher throughput. Asymmetric Architecture Delivers Efficiency V4.1-Flash is a 552-parameter mixture-of-experts (MoE) model built on a new Causal Encoder–Decoder architecture that uses just 8 billion active parameters for input processing and 16 billion for output generation. The model combines new pretraining methods with larger-scale reinforcement learning post-training to achieve benchmark results that exceed...

Continue reading