calme-3.3-qwenloi-3b
> [!TIP] > This is avery small model, so it might not perform well for some prompts and...
> [!TIP] > This is avery small model, so it might not perform well for some prompts and...
> [!TIP] > This is avery small model, so it might not perform well for some prompts and...
 Это квантованная версия jpacifico/Chocolatine-3B-Instruct-DPO-Revised, созданная с использованием llama.cpp DPO, точно настроенная на microsoft/Phi-3-mini-4k-instruct (3.82B params) с использованием...
This is a French DPO fine-tune of Microsoft’s Phi-3-mini-4k-instruct, improving its global understanding performances, even in English. Fine-tuned...
French-Alpaca based on microsoft/Phi-3-mini-128k-instruct 128k is the context length (in tokens) fine-tuned from the original French-Alpaca-dataset entirely generated...
The MoE architecture of Magiq 3 combines the specialized capabilities of MAGIQ Core-0, MAGIQ Translator-0, and MAGIQ Logic-0...
The growing need of artificial intelligence tools around the world has created a run for GPU power. We...
This adapter was created with the PEFT library and allowed the base model BigScience/BLOOMz 7B1 to be fine-tuned...
> ⚠️ Предварительные результаты — 1 семя, дисперсия не контролируется. Каждое число ниже — это один тренировочный прогон,...
Эта модель была преобразована в формат GGUF из jpacifico/Chocolatine-14B-Instruct-DPO-v1.2 с использованием llama.cpp через пространство GGUF-my-repo ggml.ai. Более подробную...