Commit ·
f9e7364
1
Parent(s): e4b2407
Updated model card (#1)
Browse files- Updated model card (8c550d07a5833a4c5da8d4e65ba69ef7c665674b)
Co-authored-by: Helen Zhao <shanjiaz@users.noreply.huggingface.co>
README.md
CHANGED
|
@@ -17,7 +17,7 @@ This is a DFlash speculator model for [RedHatAI/NVIDIA-Nemotron-3-Ultra-550B-A55
|
|
| 17 |
|
| 18 |
## Training Details
|
| 19 |
|
| 20 |
-
This model was trained using the [Speculators](https://github.com/vllm-project/speculators) library on a subset of [Magpie-Align/Magpie-Llama-3.1-Pro-300K-Filtered](https://huggingface.co/datasets/Magpie-Align/Magpie-Llama-3.1-Pro-300K-Filtered) and the `train_sft` split of [HuggingFaceH4/ultrachat_200k](https://huggingface.co/datasets/HuggingFaceH4/ultrachat_200k). Responses were regenerated by NVIDIA-Nemotron-3-Ultra-550B-A55B.
|
| 21 |
|
| 22 |
<details>
|
| 23 |
<summary> Commands </summary>
|
|
|
|
| 17 |
|
| 18 |
## Training Details
|
| 19 |
|
| 20 |
+
This model was trained using the [Speculators](https://github.com/vllm-project/speculators) library on a subset of [Magpie-Align/Magpie-Llama-3.1-Pro-300K-Filtered](https://huggingface.co/datasets/Magpie-Align/Magpie-Llama-3.1-Pro-300K-Filtered) and the `train_sft` split of [HuggingFaceH4/ultrachat_200k](https://huggingface.co/datasets/HuggingFaceH4/ultrachat_200k). Responses were regenerated by NVIDIA-Nemotron-3-Ultra-550B-A55B. Training compute for this model was generously provided by [Lambda](https://lambda.ai/), a leading cloud platform for AI training and inference.
|
| 21 |
|
| 22 |
<details>
|
| 23 |
<summary> Commands </summary>
|