How to use from
Unsloth Studio
Install Unsloth Studio (macOS, Linux, WSL)
curl -fsSL https://unsloth.ai/install.sh | sh
# Run unsloth studio
unsloth studio -H 0.0.0.0 -p 8888
# Then open http://localhost:8888 in your browser
# Search for UnluckyOrangutan/thomas-zhu-lean-premise-gguf to start chatting
Install Unsloth Studio (Windows)
irm https://unsloth.ai/install.ps1 | iex
# Run unsloth studio
unsloth studio -H 0.0.0.0 -p 8888
# Then open http://localhost:8888 in your browser
# Search for UnluckyOrangutan/thomas-zhu-lean-premise-gguf to start chatting
Using HuggingFace Spaces for Unsloth
# No setup required
# Open https://huggingface.co/spaces/unsloth/studio in your browser
# Search for UnluckyOrangutan/thomas-zhu-lean-premise-gguf to start chatting
Quick Links

Thomas Zhu Lean Premise GGUF

GGUF conversion of the Lean premise embedding model used by hanwenzhu/lean-premise-server.

  • Source model: l3lab/all-distilroberta-v1-lr2e-4-bs256-nneg3-ml-ne2
  • Source revision: v4.30.0
  • Architecture: RobertaModel
  • Converted file: thomas-zhu-lean-premise.f16.gguf

Example with joint-server in non-joint premise mode:

server/build/joint-server.exe --host 127.0.0.1 --port 8081 --model D:/hparam_outputs/thomas-zhu-lean-premise.f16.gguf --no-joint --pooling mean --ctx-size 512

Use /version, /cache, and /select for premise retrieval. Autoregressive chat generation is intentionally rejected in --no-joint mode.

Downloads last month
7
GGUF
Model size
81.5M params
Architecture
bert
Hardware compatibility
Log In to add your hardware

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for UnluckyOrangutan/thomas-zhu-lean-premise-gguf

Quantized
(2)
this model