This is a progressive (mostly dare-ties, but also slerp i.a.) merge with the intention of suitable compromise for English and German local tasks. Spaetzle-v60-7b is a merge of the following models using LazyMergekit: abideen/AlphaMonarch-dora cstr/Spaetzle-v58-7b The performance looks ok so far: e.g. we get in EQ-Bench: Score (v2_de): 65.08 (Parseable: 171.0). And for the int4-inc quantized version, from Low-bit Quantized Open LLM Leaderboard: Contamination check results (reference model: Mistral instruct 7b v0.1): — MMLU: result < 0.1, %: 0.19 — TruthfulQA: result < 0.1, %: 0.34 — GSM8k: result < 0.1, %: 0.39 This is a model merge, not a format conversion. Most cstr/ repositories are GGUF conversions, where the upstream research team remains the provider of the model and the conversion changes only the numeric representation of the weights. A merge produces a model that did not previously exist, so under Regulation (EU) 2024/1689 the maintainer of this repository is plausibly the provider of it, and the duties that survive the Art. 53(2) free-and-open-source exemption — Art. 53(1)(c) and 53(1)(d) — attach here rather than upstream. Art. 53(1)(c) — copyright policy. No…
Модальности:
Генерация текста
Области применения:
Диалог / чат
Задача: Генерация текста
Автор: cstr
Теги: mistral, merge, mergekit, lazymergekit, abideen/AlphaMonarch-dora, conversational, de, en
Лайков: 3 | Загрузок: 712
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.