about
2731. Strategic Intelligence in LLMs: Evidence from Evolutionary Game Theory (arxiv.org)
2 points by ArmageddonIt on Jul 14, 2025 | hide | past | pdf | discuss
2732. One Token to Fool LLM-as-a-Judge (arxiv.org)
2 points by simonpure on Jul 14, 2025 | hide | past | pdf | discuss
2733. MemOS is a breakthrough "memory operating system" for AI (arxiv.org)
2 points by amirkabbara on Jul 14, 2025 | hide | past | pdf | discuss
2734. Persona Features Control Emergent Misalignment (arxiv.org)
6 points by Bluestein on Jul 14, 2025 | hide | past | pdf | discuss
2735. Cats Confuse LLM: Query Agnostic Adversarial Triggers for Reasoning Models (arxiv.org)
1 point by DyslexicAtheist on Jul 14, 2025 | hide | past | pdf | discuss
2736. ASK HN: Why Google's Gemini 2.5 paper has 3295 authors? (arxiv.org)
2 points by tzury on Jul 14, 2025 | hide | past | pdf | 4 comments
2737. Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs (arxiv.org)
181 points by martythemaniak on Jul 13, 2025 | hide | past | pdf | 48 comments
2738. Defending Against Prompt Injection with a Few DefensiveTokens (arxiv.org)
2 points by belter on Jul 13, 2025 | hide | past | pdf | discuss
2739. LLMs for Drug-Drug Interaction Prediction: A Comprehensive Comparison (arxiv.org)
2 points by stacktrust on Jul 13, 2025 | hide | past | pdf | discuss
2740. Large Language Models Are Not Stable Recommender Systems (2023) (arxiv.org)
4 points by wslh on Jul 13, 2025 | hide | past | pdf | discuss
2741. Dynamic Chunking for End-to-End Hierarchical Sequence Modeling (arxiv.org)
10 points by fzliu on Jul 13, 2025 | hide | past | pdf | discuss
2742. ZipNN: Lossless Compression for AI Models (2024) (arxiv.org)
4 points by sandwichsphinx on Jul 13, 2025 | hide | past | pdf | discuss
2743. Dynamic Chunking for End-to-End Hierarchical Sequence Modeling (arxiv.org)
4 points by simonpure on Jul 12, 2025 | hide | past | pdf | discuss
2744. Humans overrely on overconfident language models, across languages (arxiv.org)
7 points by softwaredoug on Jul 11, 2025 | hide | past | pdf | discuss
2745. Can Performant LLMs Be Ethical? Quantifying the Impact of Web Crawling Opt-Outs (arxiv.org)
1 point by layer8 on Jul 11, 2025 | hide | past | pdf | discuss
2746. Dynamic Chunking for End-to-End Hierarchical Sequence Modeling (arxiv.org)
5 points by Anon84 on Jul 11, 2025 | hide | past | pdf | discuss
2747. Impact of Pretraining Word Co-Occurrence on Compositional Generalization In (arxiv.org)
2 points by badmonster on Jul 11, 2025 | hide | past | pdf | 1 comment
2748. Human-Like Forgetting Curves in Deep Neural Networks (arxiv.org)
2 points by PaulHoule on Jul 11, 2025 | hide | past | pdf | discuss
2749. Strategic Intelligence in Large Language Models (arxiv.org)
1 point by RansomStark on Jul 11, 2025 | hide | past | pdf | discuss
2750. Adaptive Two Sided Laplace Transforms as a Replacement for Self-Attention (arxiv.org)
2 points by PaulHoule on Jul 11, 2025 | hide | past | pdf | 1 comment
2751. Foundation Models of Behavioral Data from Wearables Improve Health Predictions (arxiv.org)
1 point by tosh on Jul 11, 2025 | hide | past | pdf | discuss
2752. Extreme Low-Bit Clustering for Large Language Models via Knowledge Distillation (arxiv.org)
2 points by PaulHoule on Jul 11, 2025 | hide | past | pdf | discuss
2753. Violent Tendencies in LLMs: Analysis via Behavioral Vignettes (arxiv.org)
3 points by PaulHoule on Jul 10, 2025 | hide | past | pdf | discuss
2754. HRM: 27M parameters, 1000 training samples, no pre-training required (arxiv.org)
2 points by mountainview on Jul 10, 2025 | hide | past | pdf | discuss
2755. Stochastic Interpolants (arxiv.org)
2 points by gone35 on Jul 10, 2025 | hide | past | pdf | discuss
2756. Amazon gets serious with AI Safety (arxiv.org)
3 points by prizeon on Jul 10, 2025 | hide | past | pdf | discuss
2757. An analytic theory of creativity in convolutional diffusion models (arxiv.org)
7 points by foltik on Jul 9, 2025 | hide | past | pdf | discuss
2758. Generative Blocks World: Moving Things Around in Pictures (arxiv.org)
3 points by PaulHoule on Jul 9, 2025 | hide | past | pdf | discuss
2759. Large Language Models as Autonomous Spacecraft Operators in Kerbal Space Program (arxiv.org)
6 points by Bluestein on Jul 9, 2025 | hide | past | pdf | discuss
2760. Hi-SQL: Optimizing Text-to-SQL Systems Through Dynamic Hint Integration (arxiv.org)
1 point by PaulHoule on Jul 9, 2025 | hide | past | pdf | discuss