about
Stories from January 25, 2025 (UTC)
Go back a day, month, or year. Go forward a day.
1. DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL (arxiv.org)
1351 points by gradus_ad on Jan 25, 2025 | hide | past | pdf | 1056 comments
2. Kimi K1.5 Technical Report (arxiv.org)
4 points by anticensor on Jan 25, 2025 | hide | past | pdf | discuss
3. Position Information Emerges in Causal Transformers Without Positional Encoding (arxiv.org)
3 points by PaulHoule on Jan 25, 2025 | hide | past | pdf | discuss
4. Advancing Language Model Reasoning Through RL and Inference Scaling (arxiv.org)
2 points by frozenseven on Jan 25, 2025 | hide | past | pdf | discuss