inference-optimization

Nemotron-3.5-Lightning-1.4B-A0.1B-MTP

Available on FriendliAI

Dedicated Endpoints

Run this model inference on single tenant GPU with unmatched speed and reliability at scale.

Learn more

Model Details

Model Provider

inference-optimization

Model Tree

Base

nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4

Fine-tuned
this model

Input Modalities

Text

Output Modalities

Text

Supported Functionality

Dedicated Endpoints

Explore FriendliAI today

Nemotron-3.5-Lightning-1.4B-A0.1B-MTP API & Inference Endpoint | FriendliAI