Nex-N2.5-mini-MLX-4bit
MLX 4-bit conversion of Nex-N2.5-mini, a sparse MoE language model for local inference, coding, reasoning, and long-context work....
MLX 4-bit conversion of Nex-N2.5-mini, a sparse MoE language model for local inference, coding, reasoning, and long-context work....
Четырехстороннее слияние MoE архитектуры Qwen 35B-A3B, объединяющее векторы задач из трех специализированных тонких настроек в базовый якорь через...
> Built with mlx-optiq, the MLX-native toolkit to quantize, prune, fine-tune, and serve LLMs locally on Apple Silicon....
MLX conversion of apodex/Apodex-1.1-mini, a 35.95B-parameter Qwen3.5 MoE model for research, data, files, code, and tool-driven work. This...
Qwen3.5-35B-A3B trained with online RL inside the unmodified OpenHands SDK harness, at 200K context, on 2,699 real repository...
0 жестких отказов в 842 внутренних тестовых подсказках, 0 жестких отказов в отдельном 126-подсказке и 23/24 проверки согласованности...
Unsloth Dynamic 2.0 achieves superior accuracy & outperforms other leading quants. Aloha! 🌺 Today, we are releasing Ornith-1.0,...
FP8 prompt-expansion model based on Qwen3.6-35B-A3B for Ideogram v4. Takes a short image prompt and expands it into...
Калиброванное квантование NVFP4 InternScience/Agents-A1 (агент Qwen3.5-35B-A3B hybrid MoE) для vLLM. 21,8 ГБ — и оно соответствует или превосходит...
Версия 26.05.01 Calibration STEM and Agentic Languages EN ZH HI AR RU JA KO NL FR ES Model...