saital
Qwen3.5-4B-W4A16-g128-lmheadW8
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
saital
Model Tree
Base
Qwen/Qwen3.5-4B
Quantized
this modelInput Modalities
TextImageVideo
Output Modalities
Text
Supported Functionality
Dedicated Endpoints