Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
steven0226
/
qwen2.5-1.5b-wordle-grpo
like
0
Reinforcement Learning
Safetensors
English
grpo
agentic-rl
multi-turn
wordle
trl
lora
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
qwen2.5-1.5b-wordle-grpo
/
adapter_config.json
Commit History
Upload folder using huggingface_hub
936d9ea
verified
steven0226
commited on
21 days ago