DE
DeepSeek-V4-Flash-0731
deepseek-ai
ID: deepseek-v4-flash-0731
open weights·Active·chat reasoning·Updated Jul 31, 2026
DeepSeek-V4-Flash-0731 is a 304.2B-parameter Mixture-of-Experts (MoE) text generation model developed by deepseek-ai. Built on the DeepseekV4ForCausalLM architecture using transformers. Released on 2026-07-31 with 301 likes and 0 downloads on Hugging Face.
DeepSeek-V4-Flash-0731
Model Overview
DeepSeek-V4-Flash-0731 is a 304.2B-parameter Mixture-of-Experts text generation model developed by deepseek-ai. Built on the DeepseekV4ForCausalLM architecture. Released on 2026-07-31.
📊 Quick Specs
Specification Table
| Specification | Value |
|---|---|
| Parameters | 304.2B |
| Architecture | DeepseekV4ForCausalLM |
| Task | text generation |
| Modality | text |
| License | MIT |
| Framework | transformers |
| MoE | Yes |
| Languages | — |
✨ Key Features
- 304.2B parameters with sparse MoE architecture for efficient inference
- Built on DeepseekV4ForCausalLM architecture (transformers)
- Primary task: text generation (text modality)
- Open-weights under MIT license — self-hostable and fine-tunable
📈 Community Adoption
- 301 likes on Hugging Face
- 0 downloads on Hugging Face
🔗 Resources
- Hugging Face Hub: DeepSeek-V4-Flash-0731 on Hugging Face
- Paper: arXiv
📜 License & Access
MIT — Open-weights model available for download, fine-tuning, and self-hosted deployment.
