aliRafik
glm_flash_finetuned_16bit
Available on FriendliAI
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Run this model inference with full control and performance in your environment.
Model Details
Model Provider
aliRafik
Model Tree
unsloth/GLM-4.7-Flash
Input Modalities
Output Modalities
Supported Functionality
