ruslandev/llama-3-8b-samantha - Каталог нейросетей
Генерация текста

ruslandev/llama-3-8b-samantha

Добавлено:
ruslandev/llama-3-8b-samantha

— Developed by: ruslandev — License: apache-2.0 — Finetuned from model : unsloth/llama-3-8b-bnb-4bit This model is finetuned on the data of Samantha. Prompt format is Alpaca. I used the same system prompt as the original Samantha. — learningrate: 2e-4 — seed: 3407 — gradientaccumulationsteps: 4 — perdevicetrainbatchsize: 2 — optimizer: adamw8bit — lrschedulertype: linear — warmupsteps: 5 — numepochs: 2 — weight_decay: 0.01 2 epoch finetuning from llama-3-8b took 1 hour on a single A100 with Unsloth and Huggingface’s TRL library.

Модальности:
Генерация текста


Задача: Генерация текста
Автор: ruslandev
Теги: llama, text-generation-inference, unsloth, trl, en, endpoints_compatible
Лайков: 3  |  Загрузок: 30

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.