Sign inSign up

amdenterpriseai/aim-deepseek-ai-deepseek-v4-flash-0731

By amdenterpriseai

•Updated 12 days ago

DeepSeek-V4-Flash-0731 is the official release of the DeepSeek-V4-Flash family from DeepSeek AI, ...

Image
0

292

amdenterpriseai/aim-deepseek-ai-deepseek-v4-flash-0731 repository overview

⁠DeepSeek V4 Flash 0731

DeepSeek-V4-Flash-0731 is the official release of the DeepSeek-V4-Flash family from DeepSeek AI, a 284 billion total parameter mixture-of-experts (MoE) language model with 13 billion active parameters. It supports a 1 million token context window and thinking and non-thinking modes with configurable reasoning effort (low, high, max). Post-trained for agentic and coding workloads, it supports tool calling. Architecture uses hybrid Compressed Sparse Attention and Heavily Compressed Attention with Manifold-Constrained Hyper-Connections. The checkpoint is mixed precision, with MoE expert weights in FP4 and attention, norm, and router weights in FP8, and it ships a fused DSpark draft module for speculative decoding. Licensed under MIT.

VendorAMD
LicenseMIT
Sourcehttps://github.com/amd-enterprise-ai/aim-build⁠
Revision1346fcb66c4ab4802fd777abb8d253f639ed80f4
Built2026-09-28T10:50:51Z

Tag summary

Content type

Image

Digest

sha256:bb4e2c8b9…

Size

11 GB

Last updated

12 days ago

docker pull amdenterpriseai/aim-deepseek-ai-deepseek-v4-flash-0731:2026.11.0