Qwen3.5-14B-A3B-Claude-4.6-Opus-Reasoning-Distilled-reap
can i fit moe qwen3.5 in 10gb vram? since thats already risky, lets yolo and use claude distil...
can i fit moe qwen3.5 in 10gb vram? since thats already risky, lets yolo and use claude distil...
> 📢 Release Note > Build Environment Upgrades: > — Fine-tuning Framework: Unsloth 2026.3.3 > — Core Dependencies:...
This model was introduced in the paper CHIMERA: Compact Synthetic Data for Generalizable LLM Reasoning. Authors: Xinyu Zhu,...
— Разработано: Madras1 — Лицензия: apache-2.0 — Тонкая настройка из модели : unsloth/qwen2.5-1.5b-instruct Модальности:Генерация текста Области применения:Математика Логика...
Это модель Nanbeige4.1-3B, преобразованная в формат MLX с 4-разрядным квантованием (аффин, размер группы=64) для эффективного вывода на Apple...
Full Production unrestricted/unchained PRISM-PRO version of StepFun’s Step 3.5 Flash intended particularly for full over-refusal and propaganda mechanisms...
Today we’re releasing DeepBrainz-R1, a family of reasoning-first Small Language Models (SLMs) designed for agentic AI systems in...
CURE-MED-14B is a 14 billion parameter large language model specialized for multilingual medical reasoning, fine-tuned from Qwen/Qwen2.5-14B using...
CURE-MED-7B is a 7 billion parameter large language model specialized for multilingual medical reasoning, fine-tuned from Qwen/Qwen2.5-7B using...
This repository is not: — A single fine-tuned model — A benchmark-optimized demo — A plug-and-play chatbot framework...