Qwen3 offers multilingual embedding and reranking models in 0.6B, 4B, and 8B sizes. These are inference components, not a replacement RAG platform. I would compare the smaller sizes first in a cloud deployment; no winner on this Markdown corpus is established.