Image-Text-to-Text
PEFT
Safetensors
graph-learning
multimodal-graphs
vision-language-model
qwen-vl
lora
node-classification
link-prediction
Instructions to use oofwite/OMG-VLM with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- PEFT
How to use oofwite/OMG-VLM with PEFT:
from peft import PeftModel from transformers import AutoModelForCausalLM base_model = AutoModelForCausalLM.from_pretrained("Qwen/Qwen-VL-Chat") model = PeftModel.from_pretrained(base_model, "oofwite/OMG-VLM") - Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -71,7 +71,7 @@ These values must match at inference time:
|
|
| 71 |
| Visual adapter layers / queries / heads | 1 / 32 / 32 |
|
| 72 |
| Visual compressor | on |
|
| 73 |
| Textual aggregation heads / pool layers / MLP ratio | 16 / 1 / 4.0 |
|
| 74 |
-
|
|
| 75 |
| Max neighbors | 10 |
|
| 76 |
|
| 77 |
## Usage
|
|
|
|
| 71 |
| Visual adapter layers / queries / heads | 1 / 32 / 32 |
|
| 72 |
| Visual compressor | on |
|
| 73 |
| Textual aggregation heads / pool layers / MLP ratio | 16 / 1 / 4.0 |
|
| 74 |
+
| Textual aggregation context tokens | 8 |
|
| 75 |
| Max neighbors | 10 |
|
| 76 |
|
| 77 |
## Usage
|