Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Log In
Sign Up
khazarai
/
Math-RL
like
1
Follow
KhazarAI
7
Text Generation
Transformers
Safetensors
HoangHa/pensez-grpo
English
qwen2
math
trl
unsloth
grpo
conversational
text-generation-inference
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Deploy
Use this model
main
Math-RL
/
.gitattributes
Commit History
Upload tokenizer
8f3cab6
verified
Rustamshry
commited on
3 days ago
initial commit
b256aae
verified
Rustamshry
commited on
3 days ago