Visual Document Retrieval
Transformers
Safetensors
ColPali
multilingual
qwen2_5_vl
image-text-to-text
vidore
multimodal-embedding
multilingual-embedding
Text-to-Visual Document (T→VD) retrieval
feature-extraction
sentence-similarity
mteb
text-generation-inference
🇪🇺 Region: EU
Instructions to use jinaai/jina-embeddings-v4-vllm-retrieval with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use jinaai/jina-embeddings-v4-vllm-retrieval with Transformers:
# Load model directly from transformers import AutoProcessor, AutoModelForMultimodalLM processor = AutoProcessor.from_pretrained("jinaai/jina-embeddings-v4-vllm-retrieval") model = AutoModelForMultimodalLM.from_pretrained("jinaai/jina-embeddings-v4-vllm-retrieval", device_map="auto") - ColPali
How to use jinaai/jina-embeddings-v4-vllm-retrieval with ColPali:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Notebooks
- Google Colab
- Kaggle
| { | |
| "min_pixels": 3136, | |
| "max_pixels": 12845056, | |
| "patch_size": 14, | |
| "temporal_patch_size": 2, | |
| "merge_size": 2, | |
| "image_mean": [ | |
| 0.48145466, | |
| 0.4578275, | |
| 0.40821073 | |
| ], | |
| "image_std": [ | |
| 0.26862954, | |
| 0.26130258, | |
| 0.27577711 | |
| ], | |
| "image_processor_type": "Qwen2VLImageProcessor", | |
| "processor_class": "Qwen2_5_VLProcessor" | |
| } |