TaH
Collection
Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models
•
4 items
•
Updated
•
2
This is the general version of Standard-1.7B, trained on a mixture of math, code, and science data, presented in the paper Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models.
Please visit our GitHub repo for more information.
Please see Github Example for sample usage.