models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Meta-Llama-3.1-8B-Instruct-int8Phi-4-mini-instruct-int8-ovLlama-3.1-8B-Instruct-GPTQ-Int8LFM2.5-350M-int8-ovLFM2.5-2.6B-heretic-int8-ovark-asr-0.6b-int8-onnxLFM2.5-1.2B-Thinking-ToMoE-INT8LFM2.5-350M-ToMoE-INT8Llama-3.2-1B-Instruct-GPTQ-Int8Llama-3.2-3B-Instruct-ct2-int8meta-llama_Llama-3.2-3B-Instruct-auto_round-int8-gs128-asymmeta-llama_Llama-3.2-3B-Instruct-auto_round-int8-gs128-symct2-int8-bloomz-7b1-mtmeta-llama_Llama-3.1-8B-Instruct-auto_gptq-int8-gs128-symmeta-llama_Llama-3.2-1B-Instruct-auto_gptq-int8-gs128-asymmeta-llama_Llama-3.2-3B-Instruct-auto_gptq-int8-gs64-asymt5-small-int8-dynamicmeta-llama_Llama-3.2-3B-Instruct-auto_gptq-int8-gs128-asymtiiuae_Falcon3-10B-Base-autogptq-int8-gs128-symmeta-llama_Llama-3.1-8B-Instruct-auto_round-int8-gs128-asymmeta-llama_Llama-3.1-8B-auto_gptq-int8-gs128-syminternlm_internlm3-8b-instruct-autoround-int8-gs64-asymmeta-llama_Llama-3.2-3B-Instruct-auto_gptq-int8-gs64-symmeta-llama_Llama-3.2-1B-auto_gptq-int8-gs128-symtiiuae_Falcon3-3B-Base-autoround-int8-gs128-asymtiiuae_Falcon3-3B-Instruct-autoround-int8-gs128-symLFM2.5-2.6B-openvino-int8-npumeta-llama_Llama-3.2-3B-Instruct-auto_round-int8-gs64-asymmeta-llama_Llama-3.2-1B-Instruct-auto_gptq-int8-gs128-symmistralai_Mistral-7B-Instruct-v0.3-autogptq-int8-gs128-sym
