|
|
Stories from May 27, 2025 (UTC)
|
| 1. |
Outcome-Based Reinforcement Learning to Predict the Future (arxiv.org) |
|
99 points by bturtel on May 27, 2025 | hide | past | pdf | 15 comments
|
| 2. |
Grammars of Formal Uncertainty (arxiv.org) |
|
34 points by barthelomew on May 27, 2025 | hide | past | pdf | 5 comments
|
| 3. |
Gradient-Based Program Repair: Fixing Bugs in Continuous Program Spaces (arxiv.org) |
|
16 points by andre15silva on May 27, 2025 | hide | past | pdf | discuss
|
| 4. |
Deep Reinforcement Learning, a Textbook (2023) (arxiv.org) |
|
4 points by Anon84 on May 27, 2025 | hide | past | pdf | discuss
|
| 5. |
Learning to Reason Without External Rewards (arxiv.org) |
|
4 points by epipolar on May 27, 2025 | hide | past | pdf | discuss
|
| 6. |
Optimization by unifying stochastic gradient and quasi-Newton methods (2013) (arxiv.org) |
|
3 points by fzliu on May 27, 2025 | hide | past | pdf | discuss
|
| 7. |
Extracting memorized pieces of books from open-weight language models (arxiv.org) |
|
2 points by Tomte on May 27, 2025 | hide | past | pdf | discuss
|
| 8. |
Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents (arxiv.org) |
|
2 points by pontiacbandit8 on May 27, 2025 | hide | past | pdf | discuss
|
| 9. |
Self-Reflective Uncertainties: Do LLMs Know Their Internal Answer Distribution? (arxiv.org) |
|
1 point by badmonster on May 27, 2025 | hide | past | pdf | discuss
|
| 10. |
Arc-NCA: Towards Developmental Solutions to the Abstraction and Reasoning Corpus (arxiv.org) |
|
1 point by jarmitage on May 27, 2025 | hide | past | pdf | discuss
|
| 11. |
Frontier Models are Capable of In-context Scheming (arxiv.org) |
|
1 point by doener on May 27, 2025 | hide | past | pdf | discuss
|
| 12. |
ARC-NCA: Towards Developmental Solutions to the Abstraction and Reasoning Corpus (arxiv.org) |
|
1 point by jekude on May 27, 2025 | hide | past | pdf | discuss
|
|