DeepSeek
DeepSeek V4 Pro 0813
A text model listed as a large mixture-of-experts release with a 1,048,576-token provider context window.
- 컨텍스트
- 1.05M
- 입력
- US$1.32
- 출력
- US$3.96
DeepSeek의 선별 AI 모델 9개를 컨텍스트, API 가격, 모달리티, 기능으로 비교합니다.
현재 9개 모델
DeepSeek
A text model listed as a large mixture-of-experts release with a 1,048,576-token provider context window.
DeepSeek
DeepSeek-V3.1 is a hybrid model that supports both thinking mode and non-thinking mode. Compared to the previous version, this upgrade brings improvements in multiple aspects:
DeepSeek
This update maintains the model's original capabilities while addressing issues reported by users, including:
DeepSeek
We introduce **DeepSeek-V3.2**, a model that harmonizes high computational efficiency with superior reasoning and agent performance. Our approach is built upon three key technical breakthroughs:
DeepSeek
We are excited to announce the official release of DeepSeek-V3.2-Exp, an experimental version of our model. As an intermediate step toward our next-generation architecture, V3.2-Exp builds upon V3.1-Terminus by introducing DeepSeek Sparse Attention—a sparse attention mechanism designed to explore and validate optimizations for training and inference efficiency in long-context scenarios.
DeepSeek
We present a preview version of **DeepSeek-V4** series, including two strong Mixture-of-Experts (MoE) language models — **DeepSeek-V4-Pro** with 1.6T parameters (49B activated) and **DeepSeek-V4-Flash** with 284B parameters (13B activated) — both supporting a context length of **one million tokens**.
DeepSeek
**DeepSeek-V4-Flash-0731** is the official release of **DeepSeek-V4-Flash**, superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as DeepSeek-V4-Flash-DSpark, i.e. it comes with a speculative decoding module attached.
deepseek
This model always redirects to the latest model in the DeepSeek V4 Flash family.
DeepSeek
We present a preview version of **DeepSeek-V4** series, including two strong Mixture-of-Experts (MoE) language models — **DeepSeek-V4-Pro** with 1.6T parameters (49B activated) and **DeepSeek-V4-Flash** with 284B parameters (13B activated) — both supporting a context length of **one million tokens**.
모델 고유 정보와 제공사 엔드포인트 정보를 분리합니다. 서드파티 카탈로그는 탐색과 스냅샷에 사용하며 모델 정보는 공식 문서와 검증된 모델 카드를 우선합니다.