kobkrit commited on
Commit
e7ea45e
·
verified ·
1 Parent(s): 63f93cb

Thai LLM Leaderboard v0 — single-protocol scores from arXiv:2504.01789

Browse files
Files changed (1) hide show
  1. README.md +22 -6
README.md CHANGED
@@ -1,10 +1,26 @@
1
  ---
2
- title: Thai Llm Leaderboard
3
- emoji: 🏃
4
- colorFrom: gray
5
- colorTo: green
6
  sdk: static
7
- pinned: false
 
8
  ---
9
 
10
- Check out the configuration reference at https://huggingface.co/docs/hub/spaces-config-reference
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
  ---
2
+ title: Thai LLM Leaderboard
3
+ emoji: 📊
4
+ colorFrom: blue
5
+ colorTo: indigo
6
  sdk: static
7
+ pinned: true
8
+ short_description: Thai LLMs benchmarked under one identical protocol
9
  ---
10
 
11
+ # Thai LLM Leaderboard
12
+
13
+ Scores for Thai large language models evaluated under **one identical protocol** — same
14
+ prompts, same harness, every model.
15
+
16
+ - **v0 (current):** all scores from the single SkyThought evaluation run published in the
17
+ [OpenThaiGPT 1.6 & R1 technical report](https://arxiv.org/abs/2504.01789) (April 2025).
18
+ - **v1 (planned):** fresh re-runs including ThaiLLM consortium models, newer Typhoon
19
+ releases, Pathumma updates, THaLLE, and closed APIs (GPT, Gemini).
20
+
21
+ Reproduce or submit a model: [openthaigpt_eval](https://github.com/OpenThaiGPT/openthaigpt_eval)
22
+ · [Discord](https://discord.gg/7KDdKkBGUs)
23
+
24
+ Maintained by [OpenThai](https://openthai.aieat.or.th) (AIEAT · iApp Technology).
25
+ The best score in each column is computed, not chosen — our models do not win every
26
+ column, and that is the point.