Running this model locally is fastest when deployed through Docker.
Refer to the instructions below to proceed.
The setup auto-streams the model assets (expect a multi-GB download).
To guarantee smooth performance, the installation process auto-selects the best possible options for your PC.
| đź’ľ File hash: df2fd63b3aeb3d3cc2eeca775d409fba (Update date: 2026-06-24)
|
| Model | Qwen3-VL-Reranker-8B |
| Parameters | 8 B |
| Input Modalities | Text, Images |
| Output | Ranked list of candidates |
| Training Data | Large‑scale vision‑language corpora |
| Inference Speed | ~200 tokens/s on GPU |