amd
gemma-4-12B-it-w8a8-llmcompressor-v0.12.0
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
amd
Model Tree
Base
RedHatAI/gemma-4-12B-it
Quantized
this modelInput Modalities
TextAudioImageVideo
Output Modalities
Text
Supported Functionality
Dedicated Endpoints