MKLLM-7B-Translate
MKLLM-7B is an open-source Large Language Model for the Macedonian language. The model is built on top of...
MKLLM-7B is an open-source Large Language Model for the Macedonian language. The model is built on top of...
This model is a fine-tuned version of meta-llama/Meta-Llama-3-8B on the None dataset. It achieves the following results on...
2121-8/TinySlime-1.1B-v1.0 は、TinyLlama/TinyLlama-1.1B-intermediate-step-1431k-3T をベースモデルとし、augmxnt/shisa-pretrain-en-ja-v1 の学習データを使用してトレーニングされたモデルです。本モデルは、5.5B のトークンで学習されました。 このモデルは、スマートフォン、NVIDIA Jetson などの組み込みで動かすことを想定し作成されました。 — ベースモデル: TinyLlama/TinyLlama-1.1B-intermediate-step-1431k-3T — 学習データセット: augmxnt/shisa-pretrain-en-ja-v1 — 学習トークン: 55 億...
ролевая тонкая настройка kalo-team/qwen-4b-10k-WSD-CEdiff (которая, в свою очередь, представляет собой перегонку qwen 1.5 32b на qwen 1.5 4b,...
Original model: https://huggingface.co/cognitivecomputations/dolphin-2.9.1-llama-3-8b All quants made using imatrix option with dataset provided by Kalomaze here If the model...
Curated and trained by Eric Hartford, Lucas Atkins, and Fernando Fernandes, and Cognitive Computations We have retrained our...
This model is a fine-tune (DPO) of microsoft/Phi-3-mini-4k-instruct model. All GGUF models are available here: MaziyarPanahi/calme-2.2-phi3-4b-GGUF You can...
Original model: https://huggingface.co/H-D-T/Buzz-8b-Large-v0.5 All quants made using imatrix option with dataset provided by Kalomaze here No chat template...
Оригинальная модель: https://huggingface.co/cognitivecomputations/dolphin-2.9-llama3-8b-1m Все кванты сделаны с использованием опции imatrix с набором данных, предоставленным Kalomaze здесь Если модель...
— Homepage: https://lklab.kaist.ac.kr/Janus/ — Repository: https://github.com/kaistAI/Janus — Paper: https://arxiv.org/abs/2405.17977 — Point of Contact: seongyun@kaist.ac.kr Janus is a model...