Calibration
The expert quantization used 1,426 prompts / 1,081,027 tokens spanning code
review and rewriting, English and Chinese general text, multilingual text,
math and reasoning, and structured tool calls. Quantization was propagated
layer by layer: each layer's final K2/K3 mixture generated the activations for
the next layer. The dSpark blocks used a deterministic 327,680-anchor sample of
the jointly issued five-proposal target frontier.
The calibration corpus was text-only. The visual modules and visual routing
biases are preserved exactly, but multimodal quality for this quantized expert
set has not yet been measured.
Model
DeepSeek-V4-Flash-Vision-Exp is DeepSeek's experimental multimodal V4 Flash
model. It combines the V4 Flash language architecture and built-in dSpark draft
path with a vision encoder and aligner. See the
DeepSeek-V4-Flash-Vision-Exp model card
for architecture details, prompt encoding, evaluation results, and license.
License
The model is provided under the MIT license from the source repository.