|
|
Stories from July 8, 2025 (UTC)
|
| 1. |
DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Gen (arxiv.org) |
|
7 points by jlaneve on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2. |
Energy-Based Transformers Are Scalable Learners and Thinkers (arxiv.org) |
|
5 points by jonbaer on Jul 8, 2025 | hide | past | pdf | discuss
|
| 3. |
Why Do Some Language Models Fake Alignment While Others Don't? (arxiv.org) |
|
3 points by mfiguiere on Jul 8, 2025 | hide | past | pdf | discuss
|
| 4. |
Design Patterns for Securing LLM Agents Against Prompt Injections (arxiv.org) |
|
3 points by rbanffy on Jul 8, 2025 | hide | past | pdf | discuss
|
| 5. |
Measuring AI Ability to Complete Long Tasks (arxiv.org) |
|
3 points by sonabinu on Jul 8, 2025 | hide | past | pdf | discuss
|
| 6. |
Paper: Disambiguation-Centric Finetuning Makes Tool-Calling LLMs More Realistic (arxiv.org) |
|
2 points by ashutosh1919 on Jul 8, 2025 | hide | past | pdf | discuss
|
| 7. |
InfoFlood: Jailbreaking Large Language Models with Information Overload (arxiv.org) |
|
2 points by rswerve on Jul 8, 2025 | hide | past | pdf | 1 comment
|
| 8. |
Interpreting Large Language Model's Personality Through Critical Event Analysis (arxiv.org) |
|
2 points by PaulHoule on Jul 8, 2025 | hide | past | pdf | discuss
|
| 9. |
Strategic Intelligence in Large Language Models: Evidence from Evolutionary GT (arxiv.org) |
|
2 points by psychoslave on Jul 8, 2025 | hide | past | pdf | discuss
|
| 10. |
Reading Smiles: Proxy Bias in Foundation Models for Facial Emotion Recognition (arxiv.org) |
|
2 points by PaulHoule on Jul 8, 2025 | hide | past | pdf | discuss
|
| 11. |
X-Master as Foundation: Can We Lead on Humanity's Last Exam? (arxiv.org) |
|
1 point by Leary on Jul 8, 2025 | hide | past | pdf | discuss
|
| 12. |
Cats Confuse Reasoning LLM – Adversarial Triggers for Reasoning Models (arxiv.org) |
|
1 point by tfpgh on Jul 8, 2025 | hide | past | pdf | discuss
|
|