Upload README.md with huggingface_hub
Browse files
README.md
CHANGED
|
@@ -202,7 +202,7 @@ The table below compares three locally validated TensorRT-LLM variants built for
|
|
| 202 |
|
| 203 |
This comparison is intentionally local and narrow. It should not be treated as a universal benchmark across all prompts, datasets, GPUs, or TensorRT-LLM versions.
|
| 204 |
|
| 205 |
-
|
| 206 |
|
| 207 |
## Notes
|
| 208 |
|
|
|
|
| 202 |
|
| 203 |
This comparison is intentionally local and narrow. It should not be treated as a universal benchmark across all prompts, datasets, GPUs, or TensorRT-LLM versions.
|
| 204 |
|
| 205 |
+
On that same `40`-question subset, the upstream Hugging Face FP16 model scored `0.725`, while the local TensorRT-LLM FP16 engine scored `0.75`.
|
| 206 |
|
| 207 |
## Notes
|
| 208 |
|