Instructions to use thanhduc1180/fino1-viqa-sft with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use thanhduc1180/fino1-viqa-sft with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("thanhduc1180/fino1-viqa-sft", device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 1,661 Bytes
b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a 5c8964d b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 b22200a fcc54a4 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 46 47 48 49 50 51 52 53 54 55 56 57 58 | ---
base_model: TheFinAI/Fin-o1-8B
library_name: transformers
model_name: fino1-viQA-sft
tags:
- generated_from_trainer
- sft
- trl
licence: license
---
# Model Card for fino1-viQA-sft
This model is a fine-tuned version of [TheFinAI/Fin-o1-8B](https://huggingface.co/TheFinAI/Fin-o1-8B).
It has been trained using [TRL](https://github.com/huggingface/trl).
## Quick start
```python
from transformers import pipeline
question = "If you had a time machine, but could only go to the past or the future once and never return, which would you choose and why?"
generator = pipeline("text-generation", model="thanhduc1180/fino1-viQA-sft", device="cuda")
output = generator([{"role": "user", "content": question}], max_new_tokens=128, return_full_text=False)[0]
print(output["generated_text"])
```
## Training procedure
[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="150" height="24"/>](https://wandb.ai/duc-11/fino1-vi-financial-sft/runs/celf7b9c)
This model was trained with SFT.
### Framework versions
- TRL: 0.26.1
- Transformers: 4.57.3
- Pytorch: 2.9.1
- Datasets: 4.4.1
- Tokenizers: 0.22.1
## Citations
Cite TRL as:
```bibtex
@misc{vonwerra2022trl,
title = {{TRL: Transformer Reinforcement Learning}},
author = {Leandro von Werra and Younes Belkada and Lewis Tunstall and Edward Beeching and Tristan Thrush and Nathan Lambert and Shengyi Huang and Kashif Rasul and Quentin Gallou{\'e}dec},
year = 2020,
journal = {GitHub repository},
publisher = {GitHub},
howpublished = {\url{https://github.com/huggingface/trl}}
}
``` |