about
Stories from June 11, 2026 (UTC)
Go back a day, month, or year. Go forward a day.
1. Superficial Beliefs in LLM Decision-Making (arxiv.org)
3 points by MediaSquirrel 115 days ago | hide | past | pdf | discuss
2. Cheap Reward Hacking Detection (arxiv.org)
3 points by steven_pareto 115 days ago | hide | past | pdf | discuss
3. Self-Harness: Harnesses That Improve Themselves (arxiv.org)
3 points by 0xkvyb 115 days ago | hide | past | pdf | 1 comment
4. Agents' Last Exam (arxiv.org)
2 points by matt_d 115 days ago | hide | past | pdf | discuss
5. ECO: An LLM-Driven Efficient Code Optimizer for Warehouse Scale Computers (arxiv.org)
2 points by bone_tag 115 days ago | hide | past | pdf | discuss
6. Demystifying NVSHMEM: System-Level: Symmetric Memory, Device-Initiated Ops (arxiv.org)
1 point by matt_d 115 days ago | hide | past | pdf | discuss
7. Harness-Bench: Measuring Harness Effects Across Models (arxiv.org)
1 point by ZeljkoS 115 days ago | hide | past | pdf | discuss