financial-gpt-oss-20b-q8
This is a quantized Q8_0 GGUF version of a fine-tuned financial analysis model based on GPT-OSS 20B. The...
This is a quantized Q8_0 GGUF version of a fine-tuned financial analysis model based on GPT-OSS 20B. The...
— Архитектура модели: gemma-3n-E4B-it — Вход: Audio-Vision-Text — Выход: Текст — Оптимизация модели: — Квантование веса: FP8 —...
We introduce the updated version of the Qwen3-235B-A22B non-thinking mode, named Qwen3-235B-A22B-Instruct-2507, featuring the following key enhancements: —...
A tiny yet powerful instruction-tuned language model optimized for CPU inference. With only 135 million parameters and a...
— Model Architecture: Qwen3ForCausalLM — Input: Text — Output: Text — Model Optimizations: — Activation quantization: FP8 —...
— Архитектура модели: Qwen3ForCausalLM — Вход: Текст — Выход: Текст — Оптимизация модели: — Квантование веса: INT4 —...
MaziyarPanahi/QwQ-32B-GGUF contains GGUF format model files for Qwen/QwQ-32B. GGUF is a new format introduced by the llama.cpp team...
— Model creator: Nitral-AI — Original model: Nitral-AI/Captain-ErisVioletToxic-Magnum-12B MaziyarPanahi/Captain-ErisVioletToxic-Magnum-12B-GGUF contains GGUF format model files for Nitral-AI/Captain-ErisVioletToxic-Magnum-12B. GGUF is...
— Model creator: sail — Original model: sail/Sailor2-1B-Chat MaziyarPanahi/Sailor2-1B-Chat-GGUF contains GGUF format model files for sail/Sailor2-1B-Chat. GGUF is...
AstroSage-Llama-3.1-8B-GGUF — это квантованная версия AstroSage-Llama-3.1-8B, оптимизированная для эффективного развертывания при сохранении специализированных возможностей модели в астрономии, астрофизике...