ryan0712

ryan0712

llama-3-8b-slow-DUS-max-layer-method2

Available on FriendliAI

Dedicated Endpoints

Run this model inference on single tenant GPU with unmatched speed and reliability at scale.

Learn more
Container

Run this model inference with full control and performance in your environment.

Learn more

Model Details

Model Provider

ryan0712 provider

ryan0712

Model Tree

Base

ryan0712/llama-3-8b-slow-DUS-max-layer2-method2

Base

ryan0712/llama-3-8b-slow-DUS-max-layer1-method2

Merged
this model

Input Modalities

Text

Output Modalities

Text

Supported Functionality

Dedicated EndpointsContainer

Explore FriendliAI today