Open4bits
Olmo-3.1-32B-Think-mlx-2Bit
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
Open4bits
Model Tree
Base
allenai/Olmo-3.1-32B-Think
Quantized
this modelInput Modalities
Text
Output Modalities
Text
Supported Functionality
Dedicated Endpoints