nanoBeard-frigate-360M-GGUF
GGUF quantizations of a nanoBeard Frigate pirate chat model (358.3M params), for on-device inference with llama.cpp and the...
GGUF quantizations of a nanoBeard Frigate pirate chat model (358.3M params), for on-device inference with llama.cpp and the...
World’s first sub-1B parameter model with functional tool calling capability. Generates structured JSON execution plans for tool/plugin orchestration....
> 4-битный квантованный GGUF Qwen/Qwen3.5-9B, оптимизированный для вывода iOS на устройстве через llama.cpp. Самая мощная модель, которую можно...
Модель на основе T5 с параметрами 7,3 млн, которая отвечает на вопросы общего характера. Меньше, чем мозг комара,...