stevelionheart commited on
Commit
f37820f
·
verified ·
1 Parent(s): a4b57b9

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +3 -16
README.md CHANGED
@@ -42,7 +42,7 @@ model-index:
42
  # 📝 Claude-OSS (Fable 5)
43
 
44
  > [!NOTE]
45
- > **Hugging Face Metadata:** The automated metadata tag at the top of this page miscalculates this model as a 12B BF16/U8 model because the Hub parser cannot natively calculate the 32-expert MoE layout of `gpt_oss`. The actual model size is **21B parameters** running natively on **MXFP4 microquantization**.
46
 
47
  <Gallery />
48
 
@@ -124,24 +124,11 @@ To learn about how to use this model with PyTorch and Triton, check out our [ref
124
 
125
  ## Ollama
126
 
127
- If you are trying to run gpt-oss on consumer hardware, you can use Ollama by running the following commands after [installing Ollama](https://ollama.com/download).
128
-
129
- ```bash
130
- # Tesleum/claude-oss
131
- ollama pull Tesleum/claude-oss
132
- ollama run Tesleum/claude-oss
133
- ```
134
-
135
- [Learn more about how to use gpt-oss with Ollama.](https://cookbook.openai.com/articles/gpt-oss/run-locally-ollama)
136
 
137
  #### LM Studio
138
 
139
- If you are using [LM Studio](https://lmstudio.ai/) you can use the following commands to download.
140
-
141
- ```bash
142
- # claude-oss
143
- lms get Tesleum/claude-oss
144
- ```
145
 
146
  ---
147
 
 
42
  # 📝 Claude-OSS (Fable 5)
43
 
44
  > [!NOTE]
45
+ > **Hugging Face Metadata:** The automated metadata tag at the top of this page miscalculates this model as a 12B BF16/U8 model because the Hub parser cannot natively calculate the 32-expert MoE layout of `Claude-OSS`. The actual model size is **21B parameters** running natively on **MXFP4 microquantization**.
46
 
47
  <Gallery />
48
 
 
124
 
125
  ## Ollama
126
 
127
+ To achieve better performance and quality, use **vLLM** instead
 
 
 
 
 
 
 
 
128
 
129
  #### LM Studio
130
 
131
+ To achieve better performance and quality, use **vLLM** instead
 
 
 
 
 
132
 
133
  ---
134