metadata
pipeline_tag: text-generation
inference: true
license: apache-2.0
datasets:
- simplescaling/s1K
Model Summary
s1 is a reasoning model finetuned from Qwen2.5-32B-Instruct on just 1,000 examples. It matches o1-preview & exhibits test-time scaling via budget forcing.
- Repository: simplescaling/s1
- Paper: TODO
Use
The model usage is documented here.
Citation
TODO