Qwen3.5-2B
LiteRT is Google’s on-device runtime, the new name for TensorFlow Lite (Android: com.google.ai.edge.litert:litert), and litert-torch, the renamed ai-edge-torch,...
LiteRT is Google’s on-device runtime, the new name for TensorFlow Lite (Android: com.google.ai.edge.litert:litert), and litert-torch, the renamed ai-edge-torch,...
LiteRT is Google’s on-device runtime, the new name for TensorFlow Lite (Android: com.google.ai.edge.litert:litert), and litert-torch, the renamed ai-edge-torch,...
LiteRT is Google’s on-device runtime, the new name for TensorFlow Lite (Android: com.google.ai.edge.litert:litert), and litert-torch, the renamed ai-edge-torch,...
LiteRT is Google’s on-device runtime, the new name for TensorFlow Lite (Android: com.google.ai.edge.litert:litert), and litert-torch, the renamed ai-edge-torch,...
MobileMoE — это семейство языковых моделей Mixture-of-Experts (MoE) на устройстве с субмиллиардными активными параметрами, предназначенных для продвижения границы...
GGUF quantizations of a nanoBeard Frigate pirate chat model (358.3M params), for on-device inference with llama.cpp and the...
This model provides a variant of WeiboAI/VibeThinker-1.5B that is ready for deployment on Android using the LiteRT-LM. To...
Lightning-1.7B is a high-efficiency utility model designed for edge computing and low-latency workflows. Finetuned from the powerful Qwen3-1.7B...
Попросите что-нибудь на простом английском языке, и Macaw сделает это — отправит электронное письмо, найдет файл, прочитает PDF-файл...
Крошечный, полностью автономный сервер с расширенной генерацией данных. Он включает в себя небольшую языковую модель GGUF (Pleias Redline,...