about
Stories from June 4, 2026 (UTC)
Go back a day, month, or year. Go forward a day.
1. Do transformers need three projections? Systematic study of QKV variants (arxiv.org)
225 points by Anon84 121 days ago | hide | past | pdf | 47 comments
2. Latent Agents: A Post-Training Procedure for Internalized Multi-Agent Debate (arxiv.org)
29 points by PaulHoule 121 days ago | hide | past | pdf | 1 comment
3. Consciousness in AI: Insights from the Science of Consciousness (2023) (arxiv.org)
4 points by i5heu 121 days ago | hide | past | pdf | 1 comment
4. LLM memory systems benchmark: high recall near-zero precision for tested systems (arxiv.org)
4 points by decorner 121 days ago | hide | past | pdf | discuss
5. Constrained Adaptive Rejection Sampling (arxiv.org)
3 points by matt_d 121 days ago | hide | past | pdf | discuss
6. Simulation Theology: A Testable Framework for AI Alignment (arxiv.org)
3 points by allangrant 121 days ago | hide | past | pdf | discuss
7. Your AI Text is not Mine (arxiv.org)
3 points by berlianta 122 days ago | hide | past | pdf | discuss
8. AutoLab: Can Frontier Models Solve Long-Horizon Auto Research Engineering Tasks? (arxiv.org)
2 points by Anon84 121 days ago | hide | past | pdf | discuss
9. Mellum2 Technical Report (arxiv.org)
2 points by gmays 121 days ago | hide | past | pdf | discuss
10. A Primer in Post-Training Reasoning Data: What We Know About How It Works (arxiv.org)
2 points by Anon84 122 days ago | hide | past | pdf | discuss
11. Arithmetic Pedagogy for Language Models (arxiv.org)
1 point by berlianta 122 days ago | hide | past | pdf | discuss
12. Same Weights, Different Robot: A Deployment Safety View of VLA Policies (arxiv.org)
1 point by sbulaev 122 days ago | hide | past | pdf | discuss
13. Forge: Multi-Agent Graduated Exploitation and Detection Engineering (arxiv.org)
1 point by sbulaev 122 days ago | hide | past | pdf | discuss
14. Large AI Models in Dental Healthcare (arxiv.org)
1 point by berlianta 122 days ago | hide | past | pdf | discuss