Loron200
NVIDIA-Nemotron-3-Nano-30B-A3B-NVFP4
Available on FriendliAI
Dedicated Endpoints
Run this model inference on single tenant GPU with unmatched speed and reliability at scale.
Model Details
Model Provider
Loron200
Model Tree
Base
nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16
Quantized
this modelInput Modalities
Text
Output Modalities
Text
Supported Functionality
Dedicated Endpoints