VibeThinker-3B-heretic
> [!TIP] > This model is reproducible! > > See the README in the reproduce directory for more...
> [!TIP] > This model is reproducible! > > See the README in the reproduce directory for more...
> TL;DR — A local Python-coding assistant that thinks before it codes. 8.25 GB, runs on one 16...
Glint-Trace is a QLoRA adapter that teaches a tiny language model (Qwen 3.5 0.8B Base) to think before it answers. The...
Этот репозиторий содержит квантования формата GGUF модели DavidAU/granite-4.1-8b-Claude-Opus-4.6-Thinking-MAX. Квантование проводили локально с использованием llama.cpp (сборка b9556). Модальности:Генерация текста...
A fully uncensored version of openbmb/MiniCPM5-1B, produced with a single training-free stage: single-direction abliteration (Arditi et al., 2024)....
Cosmos-T-80M is the first model in the Cosmos-T series — small, from-scratch, decoder-only Transformers pretrained on chain-of-thought data...
We argue that efficient agentic reasoning benefits from decomposing deliberation into three interacting systems: reactive execution (System I)...
Original Model: OrionLLM/GRM-2.6-Opus Architecture: Qwen3.6-27B License: Apache 2.0 MTP Support: Yes Prompt: Create a complete SVG loading animation...
Что произойдет, если вы возьмете аргументацию Клода 4.7 Opus и поместите ее в Qwen 3 0.6B? Вы получите...
> Calling for independent benchmarks. I’ve only run internal smoke tests against this model — full eval numbers...