A quantized 4-bit AWQ version of the Qwen3 vision-language embedding model with 2B parameters for multimodal retrieval.