about
3301. Visual Language Models show widespread deficits on neuropsychological tests (arxiv.org)
3 points by PaulHoule on Apr 22, 2025 | hide | past | pdf | discuss
3302. Should We Respect LLMs? A Study on Influence of Prompt Politeness on Performance (arxiv.org)
48 points by rbanffy on Apr 22, 2025 | hide | past | pdf | 105 comments
3303. In between myth and reality: AI for math – a case study in category theory (arxiv.org)
1 point by belter on Apr 22, 2025 | hide | past | pdf | discuss
3304. Green Prompting (arxiv.org)
4 points by saikatsg on Apr 22, 2025 | hide | past | pdf | discuss
3305. Machine learning with neural networks (2021) (arxiv.org)
1 point by mackeye on Apr 22, 2025 | hide | past | pdf | discuss
3306. Learnable Multi-Scale Wavelet Transformer: A Novel Alternative to Self-Attention (arxiv.org)
3 points by PaulHoule on Apr 21, 2025 | hide | past | pdf | discuss
3307. Defeating Prompt Injections by Design (arxiv.org)
1 point by redbell on Apr 21, 2025 | hide | past | pdf | discuss
3308. Sleep-Time Compute: Beyond Inference Scaling at Test-Time (arxiv.org)
5 points by wooders on Apr 21, 2025 | hide | past | pdf | discuss
3309. (How) Do reasoning models reason? (arxiv.org)
2 points by nyrikki on Apr 21, 2025 | hide | past | pdf | discuss
3310. Pushing the Limits of LLM Quantization via the Linearity Theorem (arxiv.org)
95 points by felineflock on Apr 20, 2025 | hide | past | pdf | 2 comments
3311. Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents (arxiv.org)
14 points by distalx on Apr 19, 2025 | hide | past | pdf | discuss
3312. Using Small Models to Compare Language Learning and Tokenizer Performance (arxiv.org)
1 point by PaulHoule on Apr 19, 2025 | hide | past | pdf | discuss
3313. Do Reasoning Models Show Better Verbalized Calibration? (arxiv.org)
2 points by veryluckyxyz on Apr 19, 2025 | hide | past | pdf | discuss
3314. Defeating Prompt Injections by Design (arxiv.org)
2 points by theptip on Apr 19, 2025 | hide | past | pdf | discuss
3315. Inferring the Phylogeny of Large Language Models (arxiv.org)
69 points by weinzierl on Apr 19, 2025 | hide | past | pdf | 6 comments
3316. How to evaluate control measures for LLM agents? (arxiv.org)
2 points by handfuloflight on Apr 19, 2025 | hide | past | pdf | discuss
3317. CaMeL: Defeating Prompt Injections by Design (arxiv.org)
71 points by tomrod on Apr 19, 2025 | hide | past | pdf | 16 comments
3318. MageSQL: Enhancing In-Context Learning for Text-to-SQL Applications with LLMs (arxiv.org)
2 points by PaulHoule on Apr 18, 2025 | hide | past | pdf | discuss
3319. Self-Steering Language Models (arxiv.org)
2 points by marojejian on Apr 18, 2025 | hide | past | pdf | 1 comment
3320. Attention is all you need (2017) (arxiv.org)
3 points by simonebrunozzi on Apr 18, 2025 | hide | past | pdf | discuss
3321. SDFs from Unoriented Point Clouds Using Neural Variational Heat Distances (arxiv.org)
38 points by haxiomic on Apr 18, 2025 | hide | past | pdf | 5 comments
3322. Research Paper: Generative Agent Simulations of 1k People (arxiv.org)
2 points by virtual_rf on Apr 18, 2025 | hide | past | pdf | discuss
3323. Parameter-Efficient Fine-Tuning of LLMs for Personality Detection (arxiv.org)
1 point by PaulHoule on Apr 18, 2025 | hide | past | pdf | discuss
3324. PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-Bit LLMs (arxiv.org)
3 points by PaulHoule on Apr 18, 2025 | hide | past | pdf | discuss
3325. Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning (arxiv.org)
3 points by b-man on Apr 18, 2025 | hide | past | pdf | discuss
3326. The Imitation Game According to Turing (arxiv.org)
2 points by ifdefdebug on Apr 17, 2025 | hide | past | pdf | 1 comment
3327. Task-Aware Parameter-Efficient Fine-Tuning of Large Pre-Trained Models (arxiv.org)
3 points by PaulHoule on Apr 17, 2025 | hide | past | pdf | discuss
3328. BitNet b1.58 2B4T Technical Report (arxiv.org)
111 points by galeos on Apr 17, 2025 | hide | past | pdf | 30 comments
3329. LLMs, Syntax, and Semantics: Long-Distance Binding of Chinese Reflexive Ziji (arxiv.org)
2 points by PaulHoule on Apr 16, 2025 | hide | past | pdf | discuss
3330. HybridRAG: Integrating Knowledge Graphs and Vector RAG (arxiv.org)
2 points by GeorgeCurtis on Apr 16, 2025 | hide | past | pdf | discuss