Reinforcement Learning
Transformers
Safetensors
qwen2
text-generation
grpo
combinatorial-optimization
code-generation
sds
icml-2026
text-generation-inference
Instructions to use IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303 with Transformers:
# pip install -U transformers accelerate # Load model directly from transformers import AutoTokenizer, AutoModelForCausalLM tokenizer = AutoTokenizer.from_pretrained("IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303") model = AutoModelForCausalLM.from_pretrained("IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Download latest from IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303: direct link, hf CLI and curl.
- Browser
- Download file 13 Bytes
-
https://huggingface.co/IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303/resolve/main/latest
- Command line
-
hf download hf://IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303/latest
-
curl -L -o latest https://huggingface.co/IDEALLab/Qwen2.5-Coder-14B-Instruct-GRPO-SDS-Hero-seed303/resolve/main/latest
13 Bytes
| global_step90 |