Qwen3 Embedding 4B is the 4B size in the Qwen3 embedding and reranking family, between the 0.6B and 8B cuts. It embeds text over a 32,768-token window at 2,560 dimensions, supports 100+ languages, is instruction-aware for task prefixes, and supports Matryoshka truncation from 32 to 2,560 dimensions. It ranks below the 8B cut on MTEB multilingual but well above the 0.6B cut, at a much lower cost.
Share cards2 images

