models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
GSA-FT-Qwen2-7B-Instruct-chunk4-chunk4GSA-FT-Qwen2-7B-Instruct-chunk8-chunk4GSA-PT-Qwen2-7B-Instruct-chunk16ann-sparseattentionGSA-PT-Llama-3.2-1B-chunk4-chunk4GSA-link-FT-Llama-3.2-1B-chunk16GSA-PT-Qwen2-7B-Instruct-chunk8GSA-PT-Qwen2-7B-Instruct-chunk8-chunk4GSA-PT-Llama-3.2-1B-chunk16GSA-FT-Qwen2-7B-Instruct-chunk16GSA-FT-Llama-3.2-1B-chunk16GSA-PT-Qwen2-7B-Instruct-chunk32GSA-FT-Llama-3.2-1B-chunk4-chunk4GSA-PT-Qwen2-7B-Instruct-chunk4-chunk4GSA-FT-Qwen2-7B-Instruct-chunk8GSA-FT-Qwen2-7B-Instruct-chunk32bert-hybrid-sparse-sliding-window-attentionGSA-FT-Llama-3.2-1B-chunk8GSA-PT-Llama-3.2-1B-chunk8kkhugface-sparse-attention-transformerGSA-link-FT-Llama-3.2-1B-chunk8GSA-link-FT-Llama-3.2-1B-chunk4-chunk4bert-sparse-sliding-window-attentionsparse-attention-transformerblock-sparse-attentionLLama-Deepseek-Sparse-Attentionflash-sparse-attentiondvlt-sparse-attentionb200_sparse_attention
