Model Summary

Use

This is the token-conditional control model for our paper. You can evaluate using the information here.

Training information

Visualize in Weights & Biases

  • TRL: 0.13.0
  • Transformers: 4.48.0
  • Pytorch: 2.3.1
  • Datasets: 3.0.1
  • Tokenizers: 0.21.0

Citation

@misc{muennighoff2025s1simpletesttimescaling,
      title={s1: Simple test-time scaling}, 
      author={Niklas Muennighoff and Zitong Yang and Weijia Shi and Xiang Lisa Li and Li Fei-Fei and Hannaneh Hajishirzi and Luke Zettlemoyer and Percy Liang and Emmanuel Candès and Tatsunori Hashimoto},
      year={2025},
      eprint={2501.19393},
      archivePrefix={arXiv},
      primaryClass={cs.CL},
      url={https://arxiv.org/abs/2501.19393}, 
}
Downloads last month
87
Safetensors
Model size
32.8B params
Tensor type
F32
·
Inference Providers NEW
This model is not currently available via any of the supported third-party Inference Providers, and the model is not deployed on the HF Inference API.

Model tree for simplescaling/step-conditional-control

Base model

Qwen/Qwen2.5-32B
Finetuned
(112)
this model