apolloparty
GLM-4.1V-9B-Thinking-NVFP4A16
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
apolloparty
Model Tree
Base
THUDM/GLM-4-9B-0414
Quantized
this modelInput Modalities
Text
Output Modalities
Text
Supported Functionality
Dedicated Endpoints