Youssofal/Gemma4-MTPLX-Optimized-Speed - Каталог нейросетей
Генерация текста

Youssofal/Gemma4-MTPLX-Optimized-Speed

Добавлено:
Youssofal/Gemma4-MTPLX-Optimized-Speed

This is an MTPLX pair bundle for Gemma 4 31B speculative decoding on Apple Silicon. It is not a single vanilla Transformers model directory. The repository contains two MLX-format artifacts: — target/ — Gemma 4 31B IT target, MLX Q4 affine group-size 64 — assistant/ — official Gemma 4 31B assistant drafter, MLX Q6 affine group-size 64 — Target source: google/gemma-4-31B-it — Target revision: 145dc2508c480a64b47242f160d286cff94a2343 — Assistant source: google/gemma-4-31B-it-assistant — Assistant revision: cffbbd2cea41ea56a0fa5b0487e0d445121fd204 After downloading this repository, point MTPLX at the two subdirectori MTPLX uses exact speculative sampling with target verification and residual correction. Prompt: single-file HTML5 Canvas Flappy Bird game, capped at 1000 generated tokens. This release is optimized for MTPLX speed experiments. For a higher-precision target, use Youssofal/Gemma4-MTPLX-Optimized-Quality. Gemma 4 is released by Google under the Gemma 4 license terms linked above.

Модальности:
Генерация текста


Задача: Генерация текста
Автор: Youssofal
Теги: mlx, gemma4, mtplx, speculative-decoding, apple-silicon
Лайков: 4  |  Загрузок: 0

Открыть на HuggingFace →

Описание основано на материалах HuggingFace. Перевод выполнен автоматически.