Qwen3.5-2B-Turkish-SFT
Qwen3.5-2B базовая модель üzerine Türkçe, следующая инструкциям verisiyle, тонкая настройка edilmiş bir dil modelidir. AlicanKiraz0/Turkish-SFT-Dataset-v1.0 — genel amaçlı...
Qwen3.5-2B базовая модель üzerine Türkçe, следующая инструкциям verisiyle, тонкая настройка edilmiş bir dil modelidir. AlicanKiraz0/Turkish-SFT-Dataset-v1.0 — genel amaçlı...
Это LoRA для gemma-3-270m-it, позволяющий упростить процесс мышления с помощью элементарного инструмента, требующего инструмента калькулятора, хотя, похоже, он...
Модель Qwen3.5-122B-A10B, точно настроенная на Claude-Distills, высококачественный набор данных рассуждений, полученный от Клода. — Базовая модель: Qwen3.5-122B-A10B (MoE,...
> Calling for independent benchmarks. I’ve only run internal smoke tests against this model — full eval numbers...
An elite offensive security AI fine-tuned for bug bounty hunting Security topics covered: — XSS, SQLi, SSRF, RCE,...
Nemotron-Cascade-2-30B-A3B ships with custom model code (modelingnemotronh.py) that has two bugs exposed by causalconv1d 1.6.x. These are bugs...
Qwen3-1.7B-SFT is a supervised fine-tuned model based on Qwen3-1.7B-Base, trained on the OpenThought3-Qwen3-4B dataset for mathematical reasoning and...
AWQ quantization: done by stelterlab in INT4 with llm-compressor (https://github.com/vllm-project/llm-compressor — v0.10.0.1) from the vllm-project. See recipe.yaml for...
We’re excited to introduce Nemotron-Cascade-2-30B-A3B, an open 30B MoE model with 3B activated parameters that delivers strong reasoning...
Qwen3-4B-Kimi2.5-Reasoning-Distilled is a fine-tuned language model optimized for structured, long-form reasoning. It is derived from the Qwen3-4b-Thinking-2507 base...