about
Stories from May 27, 2025 (UTC)
Go back a day, month, or year. Go forward a day.
1. Outcome-Based Reinforcement Learning to Predict the Future (arxiv.org)
99 points by bturtel on May 27, 2025 | hide | past | pdf | 15 comments
2. Grammars of Formal Uncertainty (arxiv.org)
34 points by barthelomew on May 27, 2025 | hide | past | pdf | 5 comments
3. Gradient-Based Program Repair: Fixing Bugs in Continuous Program Spaces (arxiv.org)
16 points by andre15silva on May 27, 2025 | hide | past | pdf | discuss
4. Deep Reinforcement Learning, a Textbook (2023) (arxiv.org)
4 points by Anon84 on May 27, 2025 | hide | past | pdf | discuss
5. Learning to Reason Without External Rewards (arxiv.org)
4 points by epipolar on May 27, 2025 | hide | past | pdf | discuss
6. Optimization by unifying stochastic gradient and quasi-Newton methods (2013) (arxiv.org)
3 points by fzliu on May 27, 2025 | hide | past | pdf | discuss
7. Extracting memorized pieces of books from open-weight language models (arxiv.org)
2 points by Tomte on May 27, 2025 | hide | past | pdf | discuss
8. Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents (arxiv.org)
2 points by pontiacbandit8 on May 27, 2025 | hide | past | pdf | discuss
9. Self-Reflective Uncertainties: Do LLMs Know Their Internal Answer Distribution? (arxiv.org)
1 point by badmonster on May 27, 2025 | hide | past | pdf | discuss
10. Arc-NCA: Towards Developmental Solutions to the Abstraction and Reasoning Corpus (arxiv.org)
1 point by jarmitage on May 27, 2025 | hide | past | pdf | discuss
11. Frontier Models are Capable of In-context Scheming (arxiv.org)
1 point by doener on May 27, 2025 | hide | past | pdf | discuss
12. ARC-NCA: Towards Developmental Solutions to the Abstraction and Reasoning Corpus (arxiv.org)
1 point by jekude on May 27, 2025 | hide | past | pdf | discuss