about
Stories from December 30, 2025 (UTC)
Go back a day, month, or year. Go forward a day.
1. Stable-Pretraining-v1: Foundation Model Research Made Simple (arxiv.org)
7 points by PaulHoule 277 days ago | hide | past | pdf | discuss
2. Frontier Models are Capable of In-context Scheming (arxiv.org)
2 points by william-evans 277 days ago | hide | past | pdf | 1 comment
3. ReCollab: Retrieval-Augmented LLMs for Cooperative Ad-Hoc Teammate Modeling (arxiv.org)
1 point by StatsAreFun 277 days ago | hide | past | pdf | discuss
4. LLM Efficiency: From Hyperscale Optimizations to Universal Deployability (arxiv.org)
1 point by PaulHoule 277 days ago | hide | past | pdf | discuss
5. The Sparsely-Gated Mixture-of-Experts Layer (2017) [pdf] (arxiv.org)
1 point by swatson741 278 days ago | hide | past | pdf | discuss
6. Large Language Models Struggle to Learn Long-Tail Knowledge (2023) (arxiv.org)
1 point by wslh 278 days ago | hide | past | pdf | discuss