How to use from
Lemonade
Pull the model
# Download Lemonade from https://lemonade-server.ai/
lemonade pull DevQuasar/mistralai.Mistral-Medium-3.5-128B-GGUF:
Run and chat with the model
lemonade run user.mistralai.Mistral-Medium-3.5-128B-GGUF-
List all available models
lemonade list
Quick Links

UPDATED Using the fixed Transformers config from the original model!

'Make knowledge free for everyone'

Quantized version of: mistralai/Mistral-Medium-3.5-128B Buy Me a Coffee at ko-fi.com

Downloads last month
100
GGUF
Model size
125B params
Architecture
mistral3
Hardware compatibility
Log In to add your hardware

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for DevQuasar/mistralai.Mistral-Medium-3.5-128B-GGUF

Quantized
(24)
this model

Collection including DevQuasar/mistralai.Mistral-Medium-3.5-128B-GGUF