Llama-3.3-70B-Instruct-HIGGS-4bit
This is the model card of a 🤗 transformers model that has been pushed on the Hub. This...
This is the model card of a 🤗 transformers model that has been pushed on the Hub. This...
Data-driven mixed-precision native TurboQuant checkpoint of Qwen/Qwen3.6-35B-A3B. Extends -TQ-apex2 by skipping the shared-expert down-projection — the tensor family...