DeepSeek
Latest models
Newest verified model in each family. Spec lines appear only when the model card already stores them.
DeepSeek-V4.1-Flash is DeepSeek's September 10, 2026 multimodal Mixture-of-Experts model and the smallest member of its new architecture family. The live API ID is deepseek-flash. Official materials describe a 552B-parameter Causal Encoder–Decoder MoE that activates 8B parameters on input and 16B on output, native image-and-text understanding, and a 1-million-token context. DeepSeek says V4.1-Flash outperforms DeepSeek-V4-Pro on performance, cost, speed, and total runtime. Weights are on Hugging Face under MIT.
API ID deepseek-flash
- Context tokens
- 1,000,000
Catalog entry for this named release; see the provider’s official documentation for modalities, pricing, and context limits.
DeepSeek-R1 is a reasoning-focused model family emphasizing chain-of-thought style behavior for math, code, and structured problem solving. Deployment options include API and open-weight variants—verify licensing and hosting constraints for your region.
Also current
Other verified current models from DeepSeek. Older rows stay in the models directory.