Text Classification
Transformers
PyTorch
Safetensors
English
bert
wnli
glue
kd
torchdistill
text-embeddings-inference
Instructions to use yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("text-classification", model="yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli")# Load model directly from transformers import AutoTokenizer, AutoModelForSequenceClassification tokenizer = AutoTokenizer.from_pretrained("yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli") model = AutoModelForSequenceClassification.from_pretrained("yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli", device_map="auto") - Notebooks
- Google Colab
- Kaggle
|
Download README.md from yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli: direct link, hf CLI and curl.
- Browser
- Download file 1.44 kB
-
https://huggingface.co/yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli/resolve/ff3efccdfc66d1be5fe6d845b7877b8e783791c5/README.md
- Command line
-
hf download hf://yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli@ff3efccdfc66d1be5fe6d845b7877b8e783791c5/README.md
-
curl -L -o README.md https://huggingface.co/yoshitomo-matsubara/bert-base-uncased-wnli_from_bert-large-uncased-wnli/resolve/ff3efccdfc66d1be5fe6d845b7877b8e783791c5/README.md
1.44 kB
| language: en | |
| tags: | |
| - bert | |
| - wnli | |
| - glue | |
| - kd | |
| - torchdistill | |
| license: apache-2.0 | |
| datasets: | |
| - wnli | |
| metrics: | |
| - accuracy | |
| `bert-base-uncased` fine-tuned on WNLI dataset, using fine-tuned `bert-large-uncased` as a teacher model, [***torchdistill***](https://github.com/yoshitomo-matsubara/torchdistill) and [Google Colab](https://colab.research.google.com/github/yoshitomo-matsubara/torchdistill/blob/master/demo/glue_kd_and_submission.ipynb) for knowledge distillation. | |
| The training configuration (including hyperparameters) is available [here](https://github.com/yoshitomo-matsubara/torchdistill/blob/main/configs/sample/glue/wnli/kd/bert_base_uncased_from_bert_large_uncased.yaml). | |
| I submitted prediction files to [the GLUE leaderboard](https://gluebenchmark.com/leaderboard), and the overall GLUE score was **78.9**. | |
| Yoshitomo Matsubara: **"torchdistill Meets Hugging Face Libraries for Reproducible, Coding-Free Deep Learning Studies: A Case Study on NLP"** at *EMNLP 2023 Workshop for Natural Language Processing Open Source Software (NLP-OSS)* | |
| [[OpenReview](https://openreview.net/forum?id=A5Axeeu1Bo)] [[Preprint](https://arxiv.org/abs/2310.17644)] | |
| ```bibtex | |
| @article{matsubara2023torchdistill, | |
| title={{torchdistill Meets Hugging Face Libraries for Reproducible, Coding-Free Deep Learning Studies: A Case Study on NLP}}, | |
| author={Matsubara, Yoshitomo}, | |
| journal={arXiv preprint arXiv:2310.17644}, | |
| year={2023} | |
| } | |
| ``` | |