Models / boboliu / Qwen3 Embedding 4B W4A16 G128

Qwen3 Embedding 4B W4A16 G128

JSON →
embedding

A 4-bit quantized embedding model based on Qwen3 with 4B parameters, using W4A16 and group size 128 for efficient retrieval.

textfine-tunable
Specs
Context & limits
41K
context window
41K
max output
Lifecycle
Dates
released2025-04-01
Resources