This repository contains a bfloat16 MLX conversion of empero-ai/Qwythos-9B-Claude-Mythos-5-1M for Apple Silicon inference with MLX, MLX-LM, MLX-VLM, and local apps that use MLX backends such as LM Studio. No additional fine-tuning was performed for this repository. The weights were converted from the upstream checkpoint to MLX-compatible safetensors while preserving the upstream Apache-2.0 license and model behavior. This MLX conversion has been refreshed with the upstream v3 files. If you downloaded this model before the v3 refresh, please redownload or update this repository. v3 is a hotfix for the embedded chat template. The updated files: — update the embedded chat template for preserved reasoning and adaptive thinking; — fix looping during long generation traces; — fix agentic use in harnesses such as OpenCode, Abacus, Hermes, and Claude Code. Users with older local copies should update this MLX model before using it in LM Studio, MLX-LM, MLX-VLM, OpenCode, Abacus, Hermes, Claude Code, or other agentic harnesses. — Format: MLX safetensors — Precision: bfloat16 — Parameters: about 9B — Context length: 1,048,576 tokens in the model config — Architecture: Qwen3.5-style hybrid…
Модальности:
Генерация текста
Области применения:
Логика и рассуждение Диалог / чат Вызов функций (Tool use)
Задача: Генерация текста
Автор: xunkutech-ai
Теги: mlx, qwen3_5, mlx-lm, mlx-vlm, bfloat16, qwen3.5, reasoning, long-context
Лайков: 4 | Загрузок: 418
Описание основано на материалах HuggingFace. Перевод выполнен автоматически.