about
Stories from October 4, 2025 (UTC)
Go back a day, month, or year. Go forward a day.
1. How to inject knowledge efficiently? Knowledge infusion scaling law for LLMs (arxiv.org)
105 points by PaulHoule 364 days ago | hide | past | pdf | 35 comments
2. Provable scaling laws of feature emergence from learning dynamics of grokking (arxiv.org)
29 points by sva_ 364 days ago | hide | past | pdf | discuss
3. Physics of Learning: A Lagrangian perspective to different learning paradigms (arxiv.org)
3 points by Anon84 364 days ago | hide | past | pdf | discuss
4. Formally Verified Code Benchmark (arxiv.org)
2 points by yuppiemephisto 364 days ago | hide | past | pdf | discuss
5. Scaling Test Time Compute (arxiv.org)
2 points by math-llm-agi 364 days ago | hide | past | pdf | discuss
6. The Missing Link Between the Transformer and Models of the Brain (arxiv.org)
2 points by dominik-m 364 days ago | hide | past | pdf | discuss
7. Recursive self-aggregation unlocks deep thinking in large language models (arxiv.org)
1 point by ivansavz 364 days ago | hide | past | pdf | 1 comment