Team Ai
Modelpublic

prithivMLmods/Qwen3-Reranker-0.6B-seq-cls-GGUF

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
2likes755downloads
Model Card

Qwen3-Reranker-0.6B-seq-cls-GGUF

Qwen3-Reranker-0.6B-seq-cls is a sequence classification adaptation of the Qwen3-Reranker-0.6B model, designed for advanced text reranking and classification tasks across over 100 languages, including code and multilingual retrieval scenarios. Built on the Qwen3 series, this 0.6B parameter model offers a 32k context window and supports instruction-aware customization for specific tasks, typically improving performance by 1% to 5% when using tailored English instructions. It inherits the robust reasoning, long-text understanding, and versatility of its parent models, excelling in text retrieval, code retrieval, clustering, classification, and bitext mining, and ranks among the top models in benchmarks such as the MTEB multilingual leaderboard.

Model Files

File nameSizeQuant Type
Qwen3-4B-abliterated.F32.gguf16.1 GBF32
Qwen3-4B-abliterated.BF16.gguf8.05 GBBF16
Qwen3-4B-abliterated.F16.gguf8.05 GBF16
Qwen3-4B-abliterated.Q8_0.gguf4.28 GBQ8_0
Qwen3-4B-abliterated.Q6_K.gguf3.31 GBQ6_K
Qwen3-4B-abliterated.Q5KM.gguf2.89 GBQ5KM
Qwen3-4B-abliterated.Q5KS.gguf2.82 GBQ5KS
Qwen3-4B-abliterated.Q4KM.gguf2.5 GBQ4KM
Qwen3-4B-abliterated.Q4KS.gguf2.38 GBQ4KS
Qwen3-4B-abliterated.Q3KL.gguf2.24 GBQ3KL
Qwen3-4B-abliterated.Q3KM.gguf2.08 GBQ3KM
Qwen3-4B-abliterated.Q3KS.gguf1.89 GBQ3KS
Qwen3-4B-abliterated.Q2_K.gguf1.67 GBQ2_K

Quants Usage

(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)

Here is a handy graph by ikawrakow comparing some lower-quality quant types (lower is better):

image.png