about
Stories from July 8, 2025 (UTC)
Go back a day, month, or year. Go forward a day.
1. DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Gen (arxiv.org)
7 points by jlaneve on Jul 8, 2025 | hide | past | pdf | discuss
2. Energy-Based Transformers Are Scalable Learners and Thinkers (arxiv.org)
5 points by jonbaer on Jul 8, 2025 | hide | past | pdf | discuss
3. Why Do Some Language Models Fake Alignment While Others Don't? (arxiv.org)
3 points by mfiguiere on Jul 8, 2025 | hide | past | pdf | discuss
4. Design Patterns for Securing LLM Agents Against Prompt Injections (arxiv.org)
3 points by rbanffy on Jul 8, 2025 | hide | past | pdf | discuss
5. Measuring AI Ability to Complete Long Tasks (arxiv.org)
3 points by sonabinu on Jul 8, 2025 | hide | past | pdf | discuss
6. Paper: Disambiguation-Centric Finetuning Makes Tool-Calling LLMs More Realistic (arxiv.org)
2 points by ashutosh1919 on Jul 8, 2025 | hide | past | pdf | discuss
7. InfoFlood: Jailbreaking Large Language Models with Information Overload (arxiv.org)
2 points by rswerve on Jul 8, 2025 | hide | past | pdf | 1 comment
8. Interpreting Large Language Model's Personality Through Critical Event Analysis (arxiv.org)
2 points by PaulHoule on Jul 8, 2025 | hide | past | pdf | discuss
9. Strategic Intelligence in Large Language Models: Evidence from Evolutionary GT (arxiv.org)
2 points by psychoslave on Jul 8, 2025 | hide | past | pdf | discuss
10. Reading Smiles: Proxy Bias in Foundation Models for Facial Emotion Recognition (arxiv.org)
2 points by PaulHoule on Jul 8, 2025 | hide | past | pdf | discuss
11. X-Master as Foundation: Can We Lead on Humanity's Last Exam? (arxiv.org)
1 point by Leary on Jul 8, 2025 | hide | past | pdf | discuss
12. Cats Confuse Reasoning LLM – Adversarial Triggers for Reasoning Models (arxiv.org)
1 point by tfpgh on Jul 8, 2025 | hide | past | pdf | discuss