Qwen3-1.7B-f16
Это GGUF-количественная версия языковой модели Qwen/Qwen3-1.7B — сбалансированный LLM с 1,7 миллиардами параметров, предназначенный для эффективного локального вывода...
Это GGUF-количественная версия языковой модели Qwen/Qwen3-1.7B — сбалансированный LLM с 1,7 миллиардами параметров, предназначенный для эффективного локального вывода...
Это демонстрационная версия модели иерархического мышления, новой рекуррентной архитектуры, вдохновленной иерархической и многовременной обработкой в человеческом мозге. Модель...
Hermes 4 14B is a frontier, hybrid-mode reasoning model based on Qwen 3 14B by Nous Research that...
DES Reasoning is an experimental specialist reasoning AI with custom output format; for general reasoning and chat, try...
> [!NOTE] > Includes Unsloth chat template fixes! For llama.cpp, use —jinja > Unsloth Dynamic 2.0 achieves superior...
This model was converted to GGUF format from AmanPriyanshu/gpt-oss-4.2b-specialized-all-pruned-moe-only-4-experts using llama.cpp via the ggml.ai’s GGUF-my-repo space. Refer to...
Project: https://amanpriyanshu.github.io/GPT-OSS-MoE-ExpertFingerprinting/ This is a pruned variant of OpenAI’s GPT-OSS-20B model, reduced to 4 experts per layer based...
InfiAlign — это масштабируемая и эффективная в отношении данных структура пост-обучения, которая сочетает в себе контролируемую точную настройку...
This LoRA adapter enhances google/gemma-3-1b-it with structured reasoning capabilities using tags. Trained with GRPO (Group Relative Policy Optimization)...
This is a merge of pre-trained language models created using mergekit. This model aims to combine the code...