All
ITS Hub
SDG Hub
Training Hub
17 Feb 2025
Understanding the distinction between reasoning and inference-time scaling in LLMs - insights from our R1 reproduction experiments.
07 Feb 2025
Second update on R1 reasoning research - new results on training small LLMs with synthetic reasoning data and particle filtering methods.
06 Feb 2025
First update on R1-like reasoning experiments - Granite models show significant gains with particle filtering and new data quality experiments.
05 Feb 2025
Learn how to reproduce R1-like reasoning in small LLMs using particle filtering, synthetic data, and GRPO - achieving GPT-4o accuracy with only 4 rollouts.