libregistry
Registry
Models
Docker
IAM
MCP
Docs
Sign in with Google
Models
/
alibaba
/ Qwen3 Reranker 4B GGUF
Qwen3 Reranker 4B GGUF
JSON →
reranking
A GGUF quantized version of the Qwen3 4B reranker model for efficient local inference via llama.cpp.
text
streaming
homepage ↗
Specs
Context & limits
41K
context window
41K
max output
Lifecycle
Dates
released
2025-04-01
Resources
homepage
https://huggingface.co/Voodisss/Qwen3-Reranker-4B-GGUF-llama_cpp ↗