libregistry
Registry
Models
Docker
IAM
MCP
Docs
Sign in with Google
Models
/
alibaba
/ Qwen3 Reranker 0.6B GGUF
Qwen3 Reranker 0.6B GGUF
JSON →
reranking
A GGUF quantized 0.6-billion parameter Qwen3 reranker for lightweight local inference via llama.cpp.
text
streaming
homepage ↗
Specs
Context & limits
32K
context window
32K
max output
Lifecycle
Dates
released
2025-04-01
Resources
homepage
https://huggingface.co/Voodisss/Qwen3-Reranker-0.6B-GGUF-llama_cpp ↗