163 GB
24 files
Updated 3 months ago
Name
Size
.gitattributes3.13 kB
xet
README.md8.53 kB
xet
mistral-nemo-bophades-12B.IQ3_M.gguf5.72 GB
xet
mistral-nemo-bophades-12B.IQ3_S.gguf5.56 GB
xet
mistral-nemo-bophades-12B.IQ3_XS.gguf5.31 GB
xet
mistral-nemo-bophades-12B.IQ4_NL.gguf7.14 GB
xet
mistral-nemo-bophades-12B.IQ4_XS.gguf6.8 GB
xet
mistral-nemo-bophades-12B.Q2_K.gguf4.79 GB
xet
mistral-nemo-bophades-12B.Q3_K.gguf6.08 GB
xet
mistral-nemo-bophades-12B.Q3_K_L.gguf6.56 GB
xet
mistral-nemo-bophades-12B.Q3_K_M.gguf6.08 GB
xet
mistral-nemo-bophades-12B.Q3_K_S.gguf5.53 GB
xet
mistral-nemo-bophades-12B.Q4_0.gguf7.07 GB
xet
mistral-nemo-bophades-12B.Q4_1.gguf7.8 GB
xet
mistral-nemo-bophades-12B.Q4_K.gguf7.48 GB
xet
mistral-nemo-bophades-12B.Q4_K_M.gguf7.48 GB
xet
mistral-nemo-bophades-12B.Q4_K_S.gguf7.12 GB
xet
mistral-nemo-bophades-12B.Q5_0.gguf8.52 GB
xet
mistral-nemo-bophades-12B.Q5_1.gguf9.24 GB
xet
mistral-nemo-bophades-12B.Q5_K.gguf8.73 GB
xet
mistral-nemo-bophades-12B.Q5_K_M.gguf8.73 GB
xet
mistral-nemo-bophades-12B.Q5_K_S.gguf8.52 GB
xet
mistral-nemo-bophades-12B.Q6_K.gguf10.1 GB
xet
mistral-nemo-bophades-12B.Q8_0.gguf13 GB
xet
README.md

Quantization made by Richard Erkhov.

Github

Discord

Request more models

mistral-nemo-bophades-12B - GGUF

Original model description:

license: apache-2.0 library_name: transformers base_model: - mistralai/Mistral-Nemo-Instruct-2407 datasets: - jondurbin/truthy-dpo-v0.1 - kyujinpy/orca_math_dpo model-index: - name: mistral-nemo-bophades-12B results: - task: type: text-generation name: Text Generation dataset: name: IFEval (0-Shot) type: HuggingFaceH4/ifeval args: num_few_shot: 0 metrics: - type: inst_level_strict_acc and prompt_level_strict_acc value: 67.94 name: strict accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-bophades-12B name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: BBH (3-Shot) type: BBH args: num_few_shot: 3 metrics: - type: acc_norm value: 29.54 name: normalized accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-bophades-12B name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: MATH Lvl 5 (4-Shot) type: hendrycks/competition_math args: num_few_shot: 4 metrics: - type: exact_match value: 6.27 name: exact match source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-bophades-12B name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: GPQA (0-shot) type: Idavidrein/gpqa args: num_few_shot: 0 metrics: - type: acc_norm value: 4.7 name: acc_norm source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-bophades-12B name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: MuSR (0-shot) type: TAUR-Lab/MuSR args: num_few_shot: 0 metrics: - type: acc_norm value: 12.09 name: acc_norm source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-bophades-12B name: Open LLM Leaderboard - task: type: text-generation name: Text Generation dataset: name: MMLU-PRO (5-shot) type: TIGER-Lab/MMLU-Pro config: main split: test args: num_few_shot: 5 metrics: - type: acc value: 27.79 name: accuracy source: url: https://huggingface.co/spaces/open-llm-leaderboard/open_llm_leaderboard?query=nbeerbower/mistral-nemo-bophades-12B name: Open LLM Leaderboard

image/png

mistral-nemo-bophades-12B

mistralai/Mistral-Nemo-Instruct-2407 finetuned on jondurbin/truthy-dpo-v0.1 and kyujinpy/orca_math_dpo.

Method

Finetuned using an A100 on Google Colab for 1 epoch.

Fine-tune Llama 3 with ORPO

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

Metric Value
Avg. 24.72
IFEval (0-Shot) 67.94
BBH (3-Shot) 29.54
MATH Lvl 5 (4-Shot) 6.27
GPQA (0-shot) 4.70
MuSR (0-shot) 12.09
MMLU-PRO (5-shot) 27.79
Total size
163 GB
Files
24
Last updated
Jul 2
Pre-warmed CDN
US EU US EU

Contributors