Hugging Face model card
Fields: tags, gated, license, summary, library name, pipeline tag
We present a preview version of **DeepSeek-V4** series, including two strong Mixture-of-Experts (MoE) language models — **DeepSeek-V4-Pro** with 1.6T parameters (49B activated) and **DeepSeek-V4-Flash** with 284B parameters (13B activated) — both supporting a context length of **one million tokens**.
Provider endpoint: deepseek/deepseek-v4-pro
Fields: tags, gated, license, summary, library name, pipeline tag
Fields: model identity
Fields: identity, description, modalities, context window, maximum output, pricing, supported parameters
Answers are generated from the same sourced model and provider facts shown above.
We present a preview version of **DeepSeek-V4** series, including two strong Mixture-of-Experts (MoE) language models — **DeepSeek-V4-Pro** with 1.6T parameters (49B activated) and **DeepSeek-V4-Flash** with 284B parameters (13B activated) — both supporting a context length of **one million tokens**.
Model specifications and prices may vary by provider and change over time. AIToolly displays sources and verification dates so users can confirm critical details before production use.