| 2761. |
Hi-SQL: Optimizing Text-to-SQL Systems Through Dynamic Hint Integration (arxiv.org) |
|
1 point by PaulHoule on Jul 9, 2025 | hide | past | pdf | discuss
|
| 2762. |
MemOS: A Memory OS for AI System (arxiv.org) |
|
3 points by handfuloflight on Jul 9, 2025 | hide | past | pdf | 2 comments
|
| 2763. |
The Cost of an Image: The Energy Consumption of AI Image Generation (arxiv.org) |
|
6 points by GodelInTheShell on Jul 9, 2025 | hide | past | pdf | discuss
|
| 2764. |
Can LLMs Replace Humans During Code Chunking? (arxiv.org) |
|
1 point by PaulHoule on Jul 9, 2025 | hide | past | pdf | discuss
|
| 2765. |
How to Train a Model on a Cheap Cluster Using Block Coordinate Descent (arxiv.org) |
|
2 points by PaulHoule on Jul 9, 2025 | hide | past | pdf | discuss
|
| 2766. |
Empirical Evaluation of Large Language Models in Automated Program Repair (arxiv.org) |
|
5 points by Bluestein on Jul 9, 2025 | hide | past | pdf | discuss
|
| 2767. |
UCD: Unlearning in LLMs via Contrastive Decoding (arxiv.org) |
|
2 points by PaulHoule on Jul 9, 2025 | hide | past | pdf | discuss
|
| 2768. |
Paper: Disambiguation-Centric Finetuning Makes Tool-Calling LLMs More Realistic (arxiv.org) |
|
2 points by ashutosh1919 on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2769. |
Why Do Some Language Models Fake Alignment While Others Don't? (arxiv.org) |
|
3 points by mfiguiere on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2770. |
Design Patterns for Securing LLM Agents Against Prompt Injections (arxiv.org) |
|
3 points by rbanffy on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2771. |
InfoFlood: Jailbreaking Large Language Models with Information Overload (arxiv.org) |
|
2 points by rswerve on Jul 8, 2025 | hide | past | pdf | 1 comment
|
| 2772. |
X-Master as Foundation: Can We Lead on Humanity's Last Exam? (arxiv.org) |
|
1 point by Leary on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2773. |
Cats Confuse Reasoning LLM – Adversarial Triggers for Reasoning Models (arxiv.org) |
|
1 point by tfpgh on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2774. |
Interpreting Large Language Model's Personality Through Critical Event Analysis (arxiv.org) |
|
2 points by PaulHoule on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2775. |
DiffuCoder: Understanding and Improving Masked Diffusion Models for Code Gen (arxiv.org) |
|
7 points by jlaneve on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2776. |
Strategic Intelligence in Large Language Models: Evidence from Evolutionary GT (arxiv.org) |
|
2 points by psychoslave on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2777. |
Energy-Based Transformers Are Scalable Learners and Thinkers (arxiv.org) |
|
5 points by jonbaer on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2778. |
Measuring AI Ability to Complete Long Tasks (arxiv.org) |
|
3 points by sonabinu on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2779. |
Reading Smiles: Proxy Bias in Foundation Models for Facial Emotion Recognition (arxiv.org) |
|
2 points by PaulHoule on Jul 8, 2025 | hide | past | pdf | discuss
|
| 2780. |
Frustratingly Simple Retrieval for Challenging, Reasoning-Intensive Benchmarks (arxiv.org) |
|
2 points by amirkabbara on Jul 7, 2025 | hide | past | pdf | discuss
|
| 2781. |
Query Agnostic Adversarial Triggers for Reasoning Models (arxiv.org) |
|
5 points by layer8 on Jul 7, 2025 | hide | past | pdf | discuss
|
| 2782. |
Because We Have LLMs, We Can and Should Pursue Agentic Interpretability (arxiv.org) |
|
4 points by favoboa on Jul 7, 2025 | hide | past | pdf | discuss
|
| 2783. |
AsyncFlow: An Asynchronous Streaming RL Framework for LLM Post-Training (arxiv.org) |
|
4 points by robertnishihara on Jul 7, 2025 | hide | past | pdf | discuss
|
| 2784. |
Mercury: Ultra-fast language models based on diffusion (arxiv.org) |
|
576 points by PaulHoule on Jul 7, 2025 | hide | past | pdf | 242 comments
|
| 2785. |
Sequential Diagnosis with Language Models (arxiv.org) |
|
4 points by herrherr on Jul 7, 2025 | hide | past | pdf | discuss
|
| 2786. |
Segmentation and Representation Trade-Offs in Chemistry-Aware RAG (arxiv.org) |
|
2 points by PaulHoule on Jul 7, 2025 | hide | past | pdf | discuss
|
| 2787. |
LLMs should not replace therapists (arxiv.org) |
|
303 points by layer8 on Jul 6, 2025 | hide | past | pdf | 417 comments
|
| 2788. |
Primitive-Based Generation of Controllable and Editable 3D Semantic Scenes (arxiv.org) |
|
3 points by PaulHoule on Jul 6, 2025 | hide | past | pdf | discuss
|
| 2789. |
Establishing Best Practices for Building Rigorous Agentic Benchmarks (arxiv.org) |
|
4 points by frontfor on Jul 6, 2025 | hide | past | pdf | discuss
|
| 2790. |
Fast and Simplex: 2-Simplicial Attention in Triton (arxiv.org) |
|
3 points by anythingworks on Jul 5, 2025 | hide | past | pdf | discuss
|
| More |