Qwen3.5-122B-A10B-Opus-Reasoning-GGUF
Первая тонкая настройка рассуждений Клода Опуса Qwen3.5-122B в полном масштабе. Улучшенное многоэтапное рассуждение, аналитическая глубина и выход без...
Первая тонкая настройка рассуждений Клода Опуса Qwen3.5-122B в полном масштабе. Улучшенное многоэтапное рассуждение, аналитическая глубина и выход без...
Nemotron-Cascade-2-30B-A3B ships with custom model code (modelingnemotronh.py) that has two bugs exposed by causalconv1d 1.6.x. These are bugs...
This model is a fine-tuned version of Qwen/Qwen3-1.7B optimized for mathematical reasoning on the GSM8K benchmark. For deployment...
Qwen3-1.7B-SFT is a supervised fine-tuned model based on Qwen3-1.7B-Base, trained on the OpenThought3-Qwen3-4B dataset for mathematical reasoning and...
Expected output: Something about roundhouse kicks and bug-free code. Results may vary. Confidence will not. Preliminary Observations: Document...
AWQ quantization: done by stelterlab in INT4 with llm-compressor (https://github.com/vllm-project/llm-compressor — v0.10.0.1) from the vllm-project. See recipe.yaml for...
We’re excited to introduce Nemotron-Cascade-2-30B-A3B, an open 30B MoE model with 3B activated parameters that delivers strong reasoning...
Qwen3-4B-Kimi2.5-Reasoning-Distilled is a fine-tuned language model optimized for structured, long-form reasoning. It is derived from the Qwen3-4b-Thinking-2507 base...
Advanced reasoning. Competition-level mathematics. 96.6% TruthfulQA. 8B parameters. DeepSeek-R1 base. State of the art across every evaluated dimension....
A reasoning-focused fine-tune of Qwen/Qwen3.5-9B by Empero AI, trained to produce detailed chain-of-thought reasoning inside tags before providing...