Models & pricingModelsDeepSeek-V4-Flash-0731
DE

DeepSeek-V4-Flash-0731

deepseek-ai

open weights·Active·chat reasoning·Updated Jul 31, 2026

DeepSeek-V4-Flash-0731 is a 304.2B-parameter Mixture-of-Experts (MoE) text generation model developed by deepseek-ai. Built on the DeepseekV4ForCausalLM architecture using transformers. Released on 2026-07-31 with 301 likes and 0 downloads on Hugging Face.

DeepSeek-V4-Flash-0731

Model Overview

DeepSeek-V4-Flash-0731 is a 304.2B-parameter Mixture-of-Experts text generation model developed by deepseek-ai. Built on the DeepseekV4ForCausalLM architecture. Released on 2026-07-31.


📊 Quick Specs

Specification Table
SpecificationValue
Parameters304.2B
ArchitectureDeepseekV4ForCausalLM
Tasktext generation
Modalitytext
LicenseMIT
Frameworktransformers
MoEYes
Languages

✨ Key Features

  • 304.2B parameters with sparse MoE architecture for efficient inference
  • Built on DeepseekV4ForCausalLM architecture (transformers)
  • Primary task: text generation (text modality)
  • Open-weights under MIT license — self-hostable and fine-tunable

📈 Community Adoption

  • 301 likes on Hugging Face
  • 0 downloads on Hugging Face

🔗 Resources


📜 License & Access

MIT — Open-weights model available for download, fine-tuning, and self-hosted deployment.

Related models comparison