| 3301. |
Visual Language Models show widespread deficits on neuropsychological tests (arxiv.org) |
|
3 points by PaulHoule on Apr 22, 2025 | hide | past | pdf | discuss
|
| 3302. |
Should We Respect LLMs? A Study on Influence of Prompt Politeness on Performance (arxiv.org) |
|
48 points by rbanffy on Apr 22, 2025 | hide | past | pdf | 105 comments
|
| 3303. |
In between myth and reality: AI for math – a case study in category theory (arxiv.org) |
|
1 point by belter on Apr 22, 2025 | hide | past | pdf | discuss
|
| 3304. |
Green Prompting (arxiv.org) |
|
4 points by saikatsg on Apr 22, 2025 | hide | past | pdf | discuss
|
| 3305. |
Machine learning with neural networks (2021) (arxiv.org) |
|
1 point by mackeye on Apr 22, 2025 | hide | past | pdf | discuss
|
| 3306. |
Learnable Multi-Scale Wavelet Transformer: A Novel Alternative to Self-Attention (arxiv.org) |
|
3 points by PaulHoule on Apr 21, 2025 | hide | past | pdf | discuss
|
| 3307. |
Defeating Prompt Injections by Design (arxiv.org) |
|
1 point by redbell on Apr 21, 2025 | hide | past | pdf | discuss
|
| 3308. |
Sleep-Time Compute: Beyond Inference Scaling at Test-Time (arxiv.org) |
|
5 points by wooders on Apr 21, 2025 | hide | past | pdf | discuss
|
| 3309. |
(How) Do reasoning models reason? (arxiv.org) |
|
2 points by nyrikki on Apr 21, 2025 | hide | past | pdf | discuss
|
| 3310. |
Pushing the Limits of LLM Quantization via the Linearity Theorem (arxiv.org) |
|
95 points by felineflock on Apr 20, 2025 | hide | past | pdf | 2 comments
|
| 3311. |
Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents (arxiv.org) |
|
14 points by distalx on Apr 19, 2025 | hide | past | pdf | discuss
|
| 3312. |
Using Small Models to Compare Language Learning and Tokenizer Performance (arxiv.org) |
|
1 point by PaulHoule on Apr 19, 2025 | hide | past | pdf | discuss
|
| 3313. |
Do Reasoning Models Show Better Verbalized Calibration? (arxiv.org) |
|
2 points by veryluckyxyz on Apr 19, 2025 | hide | past | pdf | discuss
|
| 3314. |
Defeating Prompt Injections by Design (arxiv.org) |
|
2 points by theptip on Apr 19, 2025 | hide | past | pdf | discuss
|
| 3315. |
Inferring the Phylogeny of Large Language Models (arxiv.org) |
|
69 points by weinzierl on Apr 19, 2025 | hide | past | pdf | 6 comments
|
| 3316. |
How to evaluate control measures for LLM agents? (arxiv.org) |
|
2 points by handfuloflight on Apr 19, 2025 | hide | past | pdf | discuss
|
| 3317. |
CaMeL: Defeating Prompt Injections by Design (arxiv.org) |
|
71 points by tomrod on Apr 19, 2025 | hide | past | pdf | 16 comments
|
| 3318. |
MageSQL: Enhancing In-Context Learning for Text-to-SQL Applications with LLMs (arxiv.org) |
|
2 points by PaulHoule on Apr 18, 2025 | hide | past | pdf | discuss
|
| 3319. |
Self-Steering Language Models (arxiv.org) |
|
2 points by marojejian on Apr 18, 2025 | hide | past | pdf | 1 comment
|
| 3320. |
Attention is all you need (2017) (arxiv.org) |
|
3 points by simonebrunozzi on Apr 18, 2025 | hide | past | pdf | discuss
|
| 3321. |
SDFs from Unoriented Point Clouds Using Neural Variational Heat Distances (arxiv.org) |
|
38 points by haxiomic on Apr 18, 2025 | hide | past | pdf | 5 comments
|
| 3322. |
Research Paper: Generative Agent Simulations of 1k People (arxiv.org) |
|
2 points by virtual_rf on Apr 18, 2025 | hide | past | pdf | discuss
|
| 3323. |
Parameter-Efficient Fine-Tuning of LLMs for Personality Detection (arxiv.org) |
|
1 point by PaulHoule on Apr 18, 2025 | hide | past | pdf | discuss
|
| 3324. |
PIM-LLM: A High-Throughput Hybrid PIM Architecture for 1-Bit LLMs (arxiv.org) |
|
3 points by PaulHoule on Apr 18, 2025 | hide | past | pdf | discuss
|
| 3325. |
Comparing LLMs with Neuro-Symbolic on Arithmetic Relations in Abstract Reasoning (arxiv.org) |
|
3 points by b-man on Apr 18, 2025 | hide | past | pdf | discuss
|
| 3326. |
The Imitation Game According to Turing (arxiv.org) |
|
2 points by ifdefdebug on Apr 17, 2025 | hide | past | pdf | 1 comment
|
| 3327. |
Task-Aware Parameter-Efficient Fine-Tuning of Large Pre-Trained Models (arxiv.org) |
|
3 points by PaulHoule on Apr 17, 2025 | hide | past | pdf | discuss
|
| 3328. |
BitNet b1.58 2B4T Technical Report (arxiv.org) |
|
111 points by galeos on Apr 17, 2025 | hide | past | pdf | 30 comments
|
| 3329. |
LLMs, Syntax, and Semantics: Long-Distance Binding of Chinese Reflexive Ziji (arxiv.org) |
|
2 points by PaulHoule on Apr 16, 2025 | hide | past | pdf | discuss
|
| 3330. |
HybridRAG: Integrating Knowledge Graphs and Vector RAG (arxiv.org) |
|
2 points by GeorgeCurtis on Apr 16, 2025 | hide | past | pdf | discuss
|
| More |