AmberYifan
capsdnum-marin-8b-base-code_random_b8000_s0
Available on FriendliAI
Dedicated EndpointsRun this model inference on single tenant GPU with unmatched speed and reliability at scale.
Learn moreContainerRun this model inference with full control and performance in your environment.
Learn moreModel Details
Model Provider

AmberYifan
Model Tree
Base
marin-community/marin-8b-base
Supported Functionality
Dedicated EndpointsContainer
Model description
More information needed
Intended uses & limitations
More information needed
Training and evaluation data
More information needed
Training procedure
Training hyperparameters
The following hyperparameters were used during training:
- learning_rate: 1e-05
- train_batch_size: 2
- eval_batch_size: 8
- seed: 0
- distributed_type: multi-GPU
- num_devices: 4
- gradient_accumulation_steps: 8
- total_train_batch_size: 64
- total_eval_batch_size: 32
- optimizer: Use OptimizerNames.ADAMW_TORCH with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
lr_scheduler_type: cosinelr_scheduler_warmup_steps: 0.03num_epochs: 1Training results
Framework versions
- Transformers 5.7.0
- Pytorch 2.13.0+cu130
- Datasets 4.0.0
- Tokenizers 0.22.2