DeepSeek V4 Pro Dspark is a large-scale Mixture-of-Experts (MoE) chat model built by DeepSeek, supporting a context length of one million tokens. It is designed for highly efficient long-context processing, utilizing a hybrid attention architecture to reduce inference FLOPs and KV cache requirements.
Input
Output
Context
1049K
Max Output
384K
Parameters
1650.5B
Input Modalities
Output Modalities
Data sourced from official provider APIs and documentation
Last updated: Aug 8, 2026
Every month new models become cheaper, faster, and more capable. Inferbase ensures your application automatically benefits without changing a single API call.