soyrsoyr
GLM-5.3-Flash-NVFP4-MTP
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
soyrsoyr
Model Tree
Base
RedHatAI/GLM-5.3-Flash-NVFP4
Quantized
this modelInput Modalities
TextImageVideo
Output Modalities
Text
Supported Functionality
Dedicated Endpoints