soyrsoyr
GLM5.3-Flash-Tiny-Aligned-FP8MLP-MTP-MXFP4-pr3118-validation
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
soyrsoyr
Model Tree
Base
inference-optimization/GLM-5.3-Flash-0.1B-A0.1B-MTP
Quantized
this modelInput Modalities
TextImageVideo
Output Modalities
Text
Supported Functionality
Dedicated Endpoints