Any-to-Any
Transformers
Safetensors
PEFT
PyTorch
English
molmo
text-generation
molmo-audio
multimodal
audio-text-to-text
image-text-to-text
audio
speech
diarization
vllm
blaster
cc-by-nc-sa-4.0
custom_code
Instructions to use 0x8badbeef/molmo-audio-serving-diar-d with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use 0x8badbeef/molmo-audio-serving-diar-d with Transformers:
# Load model directly from transformers import AutoModelForCausalLM model = AutoModelForCausalLM.from_pretrained("0x8badbeef/molmo-audio-serving-diar-d", trust_remote_code=True, device_map="auto") - PEFT
How to use 0x8badbeef/molmo-audio-serving-diar-d with PEFT:
Task type is invalid.
- Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -26,6 +26,8 @@ pretty_name: "Blaster"
|
|
| 26 |
|
| 27 |
# Blaster
|
| 28 |
|
|
|
|
|
|
|
| 29 |
**Blaster** is a multimodal companion model: it can look at images, listen to microphone speech, and reply in text (with optional spoken replies via a codec). It is a fine-tune of [AllenAI Molmo-7B-D](https://huggingface.co/allenai/Molmo-7B-D-0924) with continuous audio understanding (SLAP), speaker markers, and discrete speech codes for generation.
|
| 30 |
|
| 31 |
This repository is the **merged serving checkpoint** — the weights you load for inference (vLLM / Transformers). Adapters and companion modules live in sibling repos under the same `molmo-audio-*` ids (marketing name: **Blaster**).
|
|
|
|
| 26 |
|
| 27 |
# Blaster
|
| 28 |
|
| 29 |
+

|
| 30 |
+
|
| 31 |
**Blaster** is a multimodal companion model: it can look at images, listen to microphone speech, and reply in text (with optional spoken replies via a codec). It is a fine-tune of [AllenAI Molmo-7B-D](https://huggingface.co/allenai/Molmo-7B-D-0924) with continuous audio understanding (SLAP), speaker markers, and discrete speech codes for generation.
|
| 32 |
|
| 33 |
This repository is the **merged serving checkpoint** — the weights you load for inference (vLLM / Transformers). Adapters and companion modules live in sibling repos under the same `molmo-audio-*` ids (marketing name: **Blaster**).
|