JunHowie commited on
Commit
f29f15d
·
verified ·
1 Parent(s): 7914f84

Upload README.md

Browse files
Files changed (1) hide show
  1. README.md +3 -1
README.md CHANGED
@@ -51,6 +51,8 @@ designed to challenge conventional integer and floating-point formats at
51
  ultra-low precision while relaxing the usual need for small group sizes; most
52
  quantized layers in this model use group size 512.
53
 
 
 
54
  - **57 output tok/s** — single-request short-context decode
55
  - **800–900+ output tok/s** — 64-way concurrency, without dSpark or other
56
  speculative decoding
@@ -991,4 +993,4 @@ Both the code repository and the model weights are released under the [Kimi K3 L
991
 
992
  ## 8. Contact Us
993
 
994
- If you have any questions, please reach out at [support@moonshot.ai](mailto:support@moonshot.ai).
 
51
  ultra-low precision while relaxing the usual need for small group sizes; most
52
  quantized layers in this model use group size 512.
53
 
54
+ Technical report: [arXiv:2608.06763](https://arxiv.org/abs/2608.06763).
55
+
56
  - **57 output tok/s** — single-request short-context decode
57
  - **800–900+ output tok/s** — 64-way concurrency, without dSpark or other
58
  speculative decoding
 
993
 
994
  ## 8. Contact Us
995
 
996
+ If you have any questions, please reach out at [support@moonshot.ai](mailto:support@moonshot.ai).