EZCon
Qwen2.5-VL-7B-Instruct-4bit-mlx
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
EZCon
Model Tree
Base
Qwen/Qwen2.5-VL-7B-Instruct
Quantized
this modelInput Modalities
TextImageVideo
Output Modalities
Text
Supported Functionality
Dedicated Endpoints