Hugging Face
Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up
SultanR
/
SmolTulu-1.7b-Reinforced
like
3
Text Generation
Transformers
Safetensors
allenai/RLVR-GSM-MATH-IF-Mixed-Constraints
English
llama
Tulu3
Smollm
SLMs
Small
Huggingface
Allenai
SFT
DPO
GGUF
RLVR
RL
conversational
text-generation-inference
Inference Endpoints
arxiv:
2411.15124
arxiv:
2412.08347
License:
apache-2.0
Model card
Files
Files and versions
Community
Train
Deploy
Use this model
main
SmolTulu-1.7b-Reinforced
Commit History
Update README.md
2116ae7
verified
SultanR
commited on
1 day ago
Update README.md
f999641
verified
SultanR
commited on
1 day ago
Update README.md
530b6c0
verified
SultanR
commited on
1 day ago
Upload smoltulubanner.png
7247435
verified
SultanR
commited on
1 day ago
Update README.md
82b207a
verified
SultanR
commited on
1 day ago
Upload tokenizer
dc7e573
verified
SultanR
commited on
1 day ago
Upload LlamaForCausalLM
178eebc
verified
SultanR
commited on
1 day ago
initial commit
12e032f
verified
SultanR
commited on
1 day ago