granite-4.2-3b-GGUF
Original model: https://huggingface.co/ibm-granite/granite-4.2-3b Model details: — Parameter count: 4B — Input support: text — Speculative decoding: no —...
Original model: https://huggingface.co/ibm-granite/granite-4.2-3b Model details: — Parameter count: 4B — Input support: text — Speculative decoding: no —...
weighted/imatrix quants of https://huggingface.co/hotdogs/Qwen35B-Agent-R2 For a convenient overview and download list, visit our model page for this model....
Ориентированный на кодирование SFT Qwen 3.6 27B, объединенный с полноточным BF16. Обучался на смешанном корпусе данных кодирования, вызова...
!MMLU-Pro !GSM8K !GPQA !Uncensored All the capability, none of the refusals. > Uncensored at no cost to quality....
> Calling for independent benchmarks. I’ve only run internal smoke tests against this model — full eval numbers...
Trinity-Large-Thinking is a reasoning-optimized variant of Arcee AI’s Trinity-Large family — a 398B-parameter sparse Mixture-of-Experts (MoE) model with...
World’s first sub-1B parameter model with functional tool calling capability. Generates structured JSON execution plans for tool/plugin orchestration....
Это модель Nanbeige4.1-3B, преобразованная в формат MLX с 4-разрядным квантованием (аффин, размер группы=64) для эффективного вывода на Apple...
Sara is a fine-tuned variant of Google’s MedGemma-1.5-4B-it that excels at medical tool calling and agentic tasks in...
A specialized fine-tuned version of the meta-llama/Llama-3.2-1B-Instruct model enhanced with function/tool calling capabilities. The model leverages the hiyouga/glaive-function-calling-v2-sharegpt...