krishnateja95 shanjiaz commited on
Commit
f9e7364
·
1 Parent(s): e4b2407

Updated model card (#1)

Browse files

- Updated model card (8c550d07a5833a4c5da8d4e65ba69ef7c665674b)


Co-authored-by: Helen Zhao <shanjiaz@users.noreply.huggingface.co>

Files changed (1) hide show
  1. README.md +1 -1
README.md CHANGED
@@ -17,7 +17,7 @@ This is a DFlash speculator model for [RedHatAI/NVIDIA-Nemotron-3-Ultra-550B-A55
17
 
18
  ## Training Details
19
 
20
- This model was trained using the [Speculators](https://github.com/vllm-project/speculators) library on a subset of [Magpie-Align/Magpie-Llama-3.1-Pro-300K-Filtered](https://huggingface.co/datasets/Magpie-Align/Magpie-Llama-3.1-Pro-300K-Filtered) and the `train_sft` split of [HuggingFaceH4/ultrachat_200k](https://huggingface.co/datasets/HuggingFaceH4/ultrachat_200k). Responses were regenerated by NVIDIA-Nemotron-3-Ultra-550B-A55B.
21
 
22
  <details>
23
  <summary> Commands </summary>
 
17
 
18
  ## Training Details
19
 
20
+ This model was trained using the [Speculators](https://github.com/vllm-project/speculators) library on a subset of [Magpie-Align/Magpie-Llama-3.1-Pro-300K-Filtered](https://huggingface.co/datasets/Magpie-Align/Magpie-Llama-3.1-Pro-300K-Filtered) and the `train_sft` split of [HuggingFaceH4/ultrachat_200k](https://huggingface.co/datasets/HuggingFaceH4/ultrachat_200k). Responses were regenerated by NVIDIA-Nemotron-3-Ultra-550B-A55B. Training compute for this model was generously provided by [Lambda](https://lambda.ai/), a leading cloud platform for AI training and inference.
21
 
22
  <details>
23
  <summary> Commands </summary>