Calibration
The expert quantization used 1,426 prompts / 1,081,027 tokens spanning code
review and rewriting, English and Chinese general text, multilingual text,
math and reasoning, and structured tool calls. Quantization was propagated
layer by layer, so each quantized layer generated the activations used for the
next layer. The dSpark blocks used a deterministic 327,680-anchor sample of
the jointly issued five-proposal target frontier.
The calibration corpus was text-only. The visual modules and visual routing
biases are preserved exactly, but multimodal quality for this quantized expert
set has not yet been measured.
Model
DeepSeek-V4-Flash-Vision-Exp is DeepSeek's experimental multimodal V4 Flash
model. It combines the V4 Flash language architecture and built-in dSpark draft
path with a vision encoder and aligner. See the
DeepSeek-V4-Flash-Vision-Exp model card
for architecture details, prompt encoding, evaluation results, and license.
License
The model is provided under the MIT license from the source repository.