|
|
Stories from May 13, 2025 (UTC)
|
| 1. |
Type-constrained code generation with language models (arxiv.org) |
|
257 points by tough on May 13, 2025 | hide | past | pdf | 127 comments
|
| 2. |
TransMLA: Multi-head latent attention is all you need (arxiv.org) |
|
123 points by ocean_moist on May 13, 2025 | hide | past | pdf | 32 comments
|
| 3. |
Rethinking Memory in AI: Taxonomy, Operations, Topics, and Future Directions (arxiv.org) |
|
5 points by wjSgoWPm5bWAhXB on May 13, 2025 | hide | past | pdf | discuss
|
| 4. |
LithOS: An Operating System for Efficient Machine Learning on GPUs (arxiv.org) |
|
3 points by PaulHoule on May 13, 2025 | hide | past | pdf | discuss
|
| 5. |
Backslash: Rate Constrained Optimized Training of Large Language Models (arxiv.org) |
|
3 points by PaulHoule on May 13, 2025 | hide | past | pdf | discuss
|
| 6. |
AWRS SMC: Fast new algorithm for guiding LLMs as Bayesian inference (arxiv.org) |
|
2 points by benlipkin on May 13, 2025 | hide | past | pdf | discuss
|
| 7. |
Can Third-Parties Read Our Emotions? (arxiv.org) |
|
2 points by PaulHoule on May 13, 2025 | hide | past | pdf | discuss
|
| 8. |
CRANE: Reasoning with Constrained LLM Generation (arxiv.org) |
|
1 point by tough on May 13, 2025 | hide | past | pdf | discuss
|
| 9. |
Intellect-2: A Reasoning Model Trained Through Globally Decentralized RL (arxiv.org) |
|
1 point by nkko on May 13, 2025 | hide | past | pdf | discuss
|
| 10. |
In-Context Learning can distort the relationship between likelihoods and fitness (arxiv.org) |
|
1 point by PaulHoule on May 13, 2025 | hide | past | pdf | discuss
|
| 11. |
Base Models Beat Aligned Models at Randomness and Creativity (arxiv.org) |
|
1 point by todsacerdoti on May 13, 2025 | hide | past | pdf | discuss
|
|