about
Stories from May 15, 2025 (UTC)
Go back a day, month, or year. Go forward a day.
1. LLMs get lost in multi-turn conversation (arxiv.org)
374 points by simonpure on May 15, 2025 | hide | past | pdf | 259 comments
2. Self Rewarding Self Improving: Autonomous LLM Improvement (arxiv.org)
28 points by tamassimond on May 15, 2025 | hide | past | pdf | discuss
3. DeepSeek-V3: Achieving Efficient LLM Scaling with 2,048 GPUs (arxiv.org)
7 points by qtwhat on May 15, 2025 | hide | past | pdf | 1 comment
4. Why do LLMs attend to the first token? (arxiv.org)
2 points by adhi01 on May 15, 2025 | hide | past | pdf | 1 comment
5. A General Theoretical Paradigm to Understand Learning from Human Preferences (arxiv.org)
2 points by yenniejun111 on May 15, 2025 | hide | past | pdf | discuss
6. Online Isolation Forest (arxiv.org)
2 points by badmonster on May 15, 2025 | hide | past | pdf | discuss
7. Understanding Perception and Reasoning Through Model Merging (arxiv.org)
2 points by veryluckyxyz on May 15, 2025 | hide | past | pdf | discuss
8. Language Agents Mirror Human Causal Reasoning Biases. How Can We Help Them Think (arxiv.org)
1 point by badmonster on May 15, 2025 | hide | past | pdf | discuss
9. Insights into DeepSeek-V3: Scaling Challenges and Reflections on Hardware for AI (arxiv.org)
1 point by matt_d on May 15, 2025 | hide | past | pdf | discuss