useful-quants

GLM-4.6V-Flash-W4A16-BF16Vision

Available on FriendliAI

Dedicated Endpoints

Run this model inference on single tenant GPU with unmatched speed and reliability at scale.

Learn more

Model Details

Model Provider

useful-quants

Model Tree

Base

zai-org/GLM-4.6V-Flash

Quantized
this model

Input Modalities

Text

Output Modalities

Text

Supported Functionality

Dedicated Endpoints

Explore FriendliAI today