about
Stories from May 21, 2025 (UTC)
Go back a day, month, or year. Go forward a day.
1. Harnessing the Universal Geometry of Embeddings (arxiv.org)
123 points by jxmorris12 on May 21, 2025 | hide | past | pdf | 40 comments
2. Sugar-Coated Poison: Benign Generation Unlocks LLM Jailbreaking (arxiv.org)
45 points by favoboa on May 21, 2025 | hide | past | pdf | 46 comments
3. µPC: Scaling Predictive Coding to 100 Layer Networks (arxiv.org)
32 points by frozenseven on May 21, 2025 | hide | past | pdf | discuss
4. Reinforcement Learning for Symbolic Mathematics (arxiv.org)
7 points by MarcoDewey on May 21, 2025 | hide | past | pdf | discuss
5. The Unreasonable Effectiveness of Reasonless Intermediate Tokens (arxiv.org)
4 points by YeGoblynQueenne on May 21, 2025 | hide | past | pdf | 1 comment
6. Show HN: Phare: A Safety Probe for Large Language Models (arxiv.org)
4 points by dberenstein1957 on May 21, 2025 | hide | past | pdf | discuss
7. Training-Free Acceleration for Diffusion Transformers (arxiv.org)
3 points by badmonster on May 21, 2025 | hide | past | pdf | 1 comment
8. Alfred: Ask a Large-Language Model for Reliable ECG Diagnosis (arxiv.org)
3 points by PaulHoule on May 21, 2025 | hide | past | pdf | discuss
9. Reinforcement Learning Finetunes Small Subnetworks in Large Language Models (arxiv.org)
3 points by s-macke on May 21, 2025 | hide | past | pdf | discuss
10. Griffin: Towards a Graph-Centric Relational Database Foundation Model (arxiv.org)
3 points by PaulHoule on May 21, 2025 | hide | past | pdf | discuss
11. Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization (arxiv.org)
2 points by badmonster on May 21, 2025 | hide | past | pdf | discuss
12. Insights into DeepSeek-V3: Scaling Challenges on Hardware for AI Architectures (arxiv.org)
2 points by tanelpoder on May 21, 2025 | hide | past | pdf | discuss
13. Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents (arxiv.org)
2 points by Fake4d on May 21, 2025 | hide | past | pdf | discuss
14. The Dangers of Browsing AI Agents (arxiv.org)
2 points by walterbell on May 21, 2025 | hide | past | pdf | discuss
15. Model Merging in Pre-Training of Large Language Models (arxiv.org)
2 points by veryluckyxyz on May 21, 2025 | hide | past | pdf | discuss
16. Emerging Properties in Unified Multimodal Pretraining (arxiv.org)
1 point by buildbot on May 21, 2025 | hide | past | pdf | 1 comment
17. Measuring General Intelligence with Generated Games (arxiv.org)
1 point by jonbaer on May 21, 2025 | hide | past | pdf | discuss
18. A Method for the Architecture of a Medical Vertical LLM Based on Deepseek R1 (arxiv.org)
1 point by PaulHoule on May 21, 2025 | hide | past | pdf | discuss
19. Scalable Quantification of User Attention in Multi-Slot Sponsored Search (arxiv.org)
1 point by PaulHoule on May 21, 2025 | hide | past | pdf | discuss
20. Pperformance of a large language model on the reasoning tasks of a physician (arxiv.org)
1 point by ibobev on May 21, 2025 | hide | past | pdf | discuss
21. Dark LLMs: The Growing Threat of Unaligned AI Models (arxiv.org)
1 point by uxhacker on May 21, 2025 | hide | past | pdf | 1 comment