axiomofmind
GLM-5.3-Flash-W4A16-NVFP4
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
axiomofmind
Model Tree
Base
zai-org/GLM-5.3-Flash-BF16
Quantized
this modelInput Modalities
TextImageVideo
Output Modalities
Text
Supported Functionality
Dedicated Endpoints