Instructions to use mlx-community/Qwen3-VL-Reranker-2B-mxfp8 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use mlx-community/Qwen3-VL-Reranker-2B-mxfp8 with Transformers:
# pip install -U transformers accelerate # Load model directly from transformers import AutoProcessor, AutoModelForMultimodalLM processor = AutoProcessor.from_pretrained("mlx-community/Qwen3-VL-Reranker-2B-mxfp8") model = AutoModelForMultimodalLM.from_pretrained("mlx-community/Qwen3-VL-Reranker-2B-mxfp8", device_map="auto") - MLX
How to use mlx-community/Qwen3-VL-Reranker-2B-mxfp8 with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] hf download mlx-community/Qwen3-VL-Reranker-2B-mxfp8 --local-dir Qwen3-VL-Reranker-2B-mxfp8
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
Download tokenizer_config.json from mlx-community/Qwen3-VL-Reranker-2B-mxfp8: direct link, hf CLI and curl.
- Browser
- Download file 389 Bytes
-
https://huggingface.co/mlx-community/Qwen3-VL-Reranker-2B-mxfp8/resolve/main/tokenizer_config.json
- Command line
-
hf download hf://mlx-community/Qwen3-VL-Reranker-2B-mxfp8/tokenizer_config.json
-
curl -L -o tokenizer_config.json https://huggingface.co/mlx-community/Qwen3-VL-Reranker-2B-mxfp8/resolve/main/tokenizer_config.json
389 Bytes
| { | |
| "add_prefix_space": false, | |
| "backend": "tokenizers", | |
| "bos_token": null, | |
| "clean_up_tokenization_spaces": false, | |
| "eos_token": "<|im_end|>", | |
| "errors": "replace", | |
| "is_local": true, | |
| "model_max_length": 262144, | |
| "pad_token": "<|endoftext|>", | |
| "processor_class": "Qwen3VLProcessor", | |
| "split_special_tokens": false, | |
| "tokenizer_class": "Qwen2Tokenizer", | |
| "unk_token": null | |
| } | |