inference-optimization

GLM-5.3-Flash-0.1B-A0.1B

Available on FriendliAI

Dedicated Endpoints

Run this model inference on single tenant GPU with unmatched speed and reliability at scale.

Learn more

Model Details

Model Provider

inference-optimization

Model Tree

Base

zai-org/GLM-5.3-Flash

Fine-tuned
this model

Input Modalities

TextImageVideo

Output Modalities

Text

Supported Functionality

Dedicated Endpoints

Explore FriendliAI today