STEMerald-2b
Model description: STEMerald-2b is a fine-tuned version of the Gemma-2b model, designed specifically for answering university-level STEM multiple-choice...
Model description: STEMerald-2b is a fine-tuned version of the Gemma-2b model, designed specifically for answering university-level STEM multiple-choice...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Chat & support: TheBloke’s Discord server Want to contribute? TheBloke’s Patreon page TheBloke’s LLM work is generously supported...
Argon-0.5B 一个自研的基线模型,复刻 DeepSeek-V4 模型训练典型优化器,并加入 Engram 模块。 本仓库计划上传模型权重、切分后的训练数据、训练代码、tokenizer 资产和完整配置,使 Argon-0.5B 成为一个可审计、可复现、可继续训练的研究型预训练样例。 这个项目的初衷是复刻 DeepSeek 技术栈中的关键训练流程,并尝试在 500M 参数规模上实现一个完整的预训练闭环。 — 复刻 、数据...
This repository is not: — A single fine-tuned model — A benchmark-optimized demo — A plug-and-play chatbot framework...
weighted/imatrix quants of https://huggingface.co/vanta-research/atom-astronomy-7b For a convenient overview and download list, visit our model page for this model....
static quants of https://huggingface.co/vanta-research/atom-astronomy-7b For a convenient overview and download list, visit our model page for this model....
weighted/imatrix quants of https://huggingface.co/vanta-research/atom-v1-preview-12b For a convenient overview and download list, visit our model page for this model....
Чат и поддержка: сервер Discord TheBloke Хотите внести свой вклад? Страница TheBloke на Patreon Работа TheBloke в области...
merlyn-education-corpus-qa-v2 — это модель преобразователя в стиле декодера с параметрами 13b для сферы образования. Он представляет собой доработанную...