QwQ-32B-preview-AWQ-AIMO-earlysharing
This model is slightly modified from the QwQ-32B-preview-AWQ model, which is the AWQ Quantized version of the QwQ-32B-preview...
This model is slightly modified from the QwQ-32B-preview-AWQ model, which is the AWQ Quantized version of the QwQ-32B-preview...
> Update February 23 2025: 🔥 BATCHING MODE SUPPORT. See 🌌 Flan-T5 provider for bulk-chain project. Test is...
🌐[Website] 📝[Paper] 🤗[Data] 🤗[Model] 🤗[Demo] We introduce 🪄Lumos, Language Agents with Unified Formats,...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
ViAble-1.7-9B — это ориентированная на рассуждения тонкая настройка Qwen 3.5 9B (архитектура qwen35): контролируемая тонкая настройка на ~98...
> 9B агентская модель, обученная на основе взаимодействий, сгенерированных в средах выполнения, развернутых в реальных рабочих процессах офиса...
Original model: https://huggingface.co/ibm-granite/granite-4.2-3b Model details: — Parameter count: 4B — Input support: text — Speculative decoding: no —...
Ulam-1-Small is a 3.086-billion-parameter mathematical reasoning model developed by Ulam AI. It is a standalone BF16 Transformers model....
This model was converted to GGUF format from rohit267/Qwen3.8-9B-heretic-uncensored using llama.cpp via the ggml.ai’s GGUF-my-repo space. Refer to...
This model was converted to GGUF format from rohit267/Qwen3.8-9B-heretic-uncensored using llama.cpp via the ggml.ai’s GGUF-my-repo space. Refer to...