models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Meta-Llama-3.1-70B-Instruct-AWQ-INT4Meta-Llama-3.1-8B-Instruct-AWQ-INT4Llama-3.2-1B-Instruct-Q4_K_M-GGUFMeta-Llama-3.1-8B-Instruct-quantized.w4a16NVIDIA-Nemotron-3.5-Lightning-30B-A3B-W4A16Llama-3.2-3B-Instruct-Q4_K_M-GGUFLTX-2.3-GGUFMeta-Llama-3.1-8B-Instruct-quantized.w8a8Llama-4-Scout-17B-16E-Instruct-quantized.w4a16Meta-Llama-3.1-8B-Instruct-GPTQ-INT4DarkIdol-Llama-3.1-8B-Instruct-1.2-Uncensored-GGUFMeta-Llama-3.1-70B-Instruct-quantized.w4a16Llama-3.2-1B-Instruct-quantized.w8a8Llama-3.3-70B-Instruct-quantized.w4a16Llama-3.2-1B-Instruct-Q8_0-GGUFMeta-Llama-3.1-8B-quantized.w8a8Mistral-Small-24B-Instruct-2501-quantized.w8a8Mistral-Nemo-Instruct-2407-GGUFLlama-3.1-8B-Instruct-Fei-v1-Uncensored-GGUFMistral-Nemo-Instruct-2407-abliterated-GGUFLlama-3.2-3B-Instruct-GGUFMistral-Nemo-Base-2407-GGUFMixtral-8x7B-Instruct-v0.1-AWQ-INT4Meta-Llama-3.1-8B-Instruct-GGUFLlama-3.2-3B-Instruct-Q8_0-GGUFphi4-multimodal-quantisized-ggufLlama-3.2-1B-Instruct-GGUFLlama-3.2-1B-GGUFMistral-Small-3.1-24B-Instruct-2503-quantized.w4a16LTX-2-GGUF
