Ulam-1-Small
Ulam-1-Small is a 3.086-billion-parameter mathematical reasoning model developed by Ulam AI. It is a standalone BF16 Transformers model....
Ulam-1-Small is a 3.086-billion-parameter mathematical reasoning model developed by Ulam AI. It is a standalone BF16 Transformers model....
MiniBananaMind-v3-9M is a small causal language model trained from scratch on FineWeb-Edu and FineMath. The model has about...
> VibeThinker-3B-heretic_decensored is a reasoning-focused language model built on top of WeiboAI/VibeThinker-3B and modified using the Heretic abliteration...
VibeThinker-3B-hereticdecensored Reasoning-focused language model modified using the Heretic abliteration toolkit Abliteration 3B Parameters STEM Reasoning Uncensored VibeThinker-3B-hereticdecensored is...
> [!TIP] > This model is reproducible! > > See the README in the reproduce directory for more...
Argon-0.5B 一个自研的基线模型,复刻 DeepSeek-V4 模型训练典型优化器,并加入 Engram 模块。 本仓库计划上传模型权重、切分后的训练数据、训练代码、tokenizer 资产和完整配置,使 Argon-0.5B 成为一个可审计、可复现、可继续训练的研究型预训练样例。 这个项目的初衷是复刻 DeepSeek 技术栈中的关键训练流程,并尝试在 500M 参数规模上实现一个完整的预训练闭环。 — 复刻 、数据...
Supertron1-8B is an instruction-tuned language model built on top of Qwen3-8B-Base. Designed to be a reliable, efficient daily...
This model was converted to GGUF format from Surpem/Supertron1-4B using llama.cpp. Refer to the original model card for...
This model is a fine-tuned version of Qwen/Qwen3-1.7B optimized for mathematical reasoning on the GSM8K benchmark. For deployment...
Qwen3-1.7B-SFT is a supervised fine-tuned model based on Qwen3-1.7B-Base, trained on the OpenThought3-Qwen3-4B dataset for mathematical reasoning and...