Visual Document Retrieval
Transformers
Safetensors
sentence-transformers
ColPali
multilingual
colqwen3
feature-extraction
multi-vector
text
image
video
multimodal-embedding
vidore
multilingual-embedding
custom_code
Instructions to use TomoroAI/tomoro-colqwen3-embed-4b with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use TomoroAI/tomoro-colqwen3-embed-4b with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("TomoroAI/tomoro-colqwen3-embed-4b", trust_remote_code=True, device_map="auto") - sentence-transformers
How to use TomoroAI/tomoro-colqwen3-embed-4b with sentence-transformers:
from sentence_transformers import SentenceTransformer model = SentenceTransformer("TomoroAI/tomoro-colqwen3-embed-4b", trust_remote_code=True) sentences = [ "The weather is lovely today.", "It's so sunny outside!", "He drove to the stadium." ] embeddings = model.encode(sentences) similarities = model.similarity(embeddings, embeddings) print(similarities.shape) # [3, 3] - ColPali
How to use TomoroAI/tomoro-colqwen3-embed-4b with ColPali:
# No code snippets available yet for this library. # To use this model, check the repository files and the library's documentation. # Want to help? PRs adding snippets are welcome at: # https://github.com/huggingface/huggingface.js
- Notebooks
- Google Colab
- Kaggle
Update configuration logic for supporting any to any batch retrieval
Browse files- processor_config.json +4 -2
processor_config.json
CHANGED
|
@@ -4,6 +4,8 @@
|
|
| 4 |
},
|
| 5 |
"processor_class": "ColQwen3Processor",
|
| 6 |
"query_prefix": "",
|
| 7 |
-
"video_prompt_prefix": "<|im_start|>user\n<|vision_start|><|video_pad|><|vision_end|>Describe the video.
|
| 8 |
-
"
|
|
|
|
|
|
|
| 9 |
}
|
|
|
|
| 4 |
},
|
| 5 |
"processor_class": "ColQwen3Processor",
|
| 6 |
"query_prefix": "",
|
| 7 |
+
"video_prompt_prefix": "<|im_start|>user\n<|vision_start|><|video_pad|><|vision_end|>Describe the video.",
|
| 8 |
+
"video_prompt_suffix": "<|im_end|><|endoftext|>",
|
| 9 |
+
"visual_prompt_prefix": "<|im_start|>user\n<|vision_start|><|image_pad|><|vision_end|>Describe the image.",
|
| 10 |
+
"visual_prompt_suffix": "<|im_end|><|endoftext|>"
|
| 11 |
}
|