Models / ggml-org / Qwen3 Reranker 0.6B Q8_0 GGUF

Qwen3 Reranker 0.6B Q8_0 GGUF

JSON →
reranking

A quantized GGUF version of the Qwen3 Reranker 0.6B model for efficient reranking with 8-bit quantization.

texttext-classification
Specs
Context & limits
32K
context window
32K
max output
Lifecycle
Dates

No lifecycle dates recorded.

Resources