| 2281. |
An Efficient Vision-Language-Action Model for Combat Tasks in 3D Action RPGs (arxiv.org) |
|
1 point by mikhael 364 days ago | hide | past | pdf | discuss
|
| 2282. |
Expected Attention: KV Cache Compression by Estimating Attention (arxiv.org) |
|
20 points by sonabinu on Oct 6, 2025 | hide | past | pdf | 3 comments
|
| 2283. |
The threat of analytic flexibility in using LLMs to simulate human data (arxiv.org) |
|
1 point by PaulHoule on Oct 6, 2025 | hide | past | pdf | discuss
|
| 2284. |
The Dragon Hatchling: The Missing Link Between the Transformer and the Brain (arxiv.org) |
|
1 point by flux3125 on Oct 6, 2025 | hide | past | pdf | discuss
|
| 2285. |
Code-to-Metric Regression: Predicting Numeric Outcomes of Code Executions (arxiv.org) |
|
2 points by chabad360 on Oct 6, 2025 | hide | past | pdf | discuss
|
| 2286. |
Brain Graph Augmentation via Learnable Edge Masking for Psychiatric Diagnosis (arxiv.org) |
|
1 point by PaulHoule on Oct 5, 2025 | hide | past | pdf | discuss
|
| 2287. |
Hybrid unary-binary design for multiplier-less printed ML classifiers (arxiv.org) |
|
2 points by PaulHoule on Oct 5, 2025 | hide | past | pdf | discuss
|
| 2288. |
Implicit Actor Critic Coupling via a Supervised Learning Framework for RLVR (arxiv.org) |
|
38 points by getnormality on Oct 5, 2025 | hide | past | pdf | 10 comments
|
| 2289. |
Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence (arxiv.org) |
|
4 points by rntn on Oct 5, 2025 | hide | past | pdf | discuss
|
| 2290. |
The Missing Link Between the Transformer and Models of the Brain (arxiv.org) |
|
1 point by SweetSoftPillow on Oct 5, 2025 | hide | past | pdf | discuss
|
| 2291. |
Formally Verified Code Benchmark (arxiv.org) |
|
2 points by yuppiemephisto on Oct 4, 2025 | hide | past | pdf | discuss
|
| 2292. |
Provable scaling laws of feature emergence from learning dynamics of grokking (arxiv.org) |
|
29 points by sva_ on Oct 4, 2025 | hide | past | pdf | discuss
|
| 2293. |
How to inject knowledge efficiently? Knowledge infusion scaling law for LLMs (arxiv.org) |
|
105 points by PaulHoule on Oct 4, 2025 | hide | past | pdf | 35 comments
|
| 2294. |
Recursive self-aggregation unlocks deep thinking in large language models (arxiv.org) |
|
1 point by ivansavz on Oct 4, 2025 | hide | past | pdf | 1 comment
|
| 2295. |
Scaling Test Time Compute (arxiv.org) |
|
2 points by math-llm-agi on Oct 4, 2025 | hide | past | pdf | discuss
|
| 2296. |
Physics of Learning: A Lagrangian perspective to different learning paradigms (arxiv.org) |
|
3 points by Anon84 on Oct 4, 2025 | hide | past | pdf | discuss
|
| 2297. |
The Missing Link Between the Transformer and Models of the Brain (arxiv.org) |
|
2 points by dominik-m on Oct 4, 2025 | hide | past | pdf | discuss
|
| 2298. |
Pretraining Large Language Models with NVFP4 (arxiv.org) |
|
1 point by aportnoy on Oct 3, 2025 | hide | past | pdf | discuss
|
| 2299. |
Pretraining Under Infinite Compute (arxiv.org) |
|
3 points by jedharris on Oct 3, 2025 | hide | past | pdf | 1 comment
|
| 2300. |
The AI Productivity Index (Apex) (arxiv.org) |
|
1 point by paulpauper on Oct 3, 2025 | hide | past | pdf | discuss
|
| 2301. |
Aristotle: IMO-Level Automated Theorem Proving (arxiv.org) |
|
3 points by jasondavies on Oct 3, 2025 | hide | past | pdf | discuss
|
| 2302. |
xLSTM Scaling Laws: Competitive Performance with Linear Time-Complexity (arxiv.org) |
|
1 point by lairv on Oct 3, 2025 | hide | past | pdf | discuss
|
| 2303. |
Thoughtbubbles: An Unsupervised Method for Parallel Thinking in Latent Space (arxiv.org) |
|
4 points by shetaye on Oct 3, 2025 | hide | past | pdf | 1 comment
|
| 2304. |
Delta-Code: How Does RL Unlock and Transfer New Programming Algorithms in LLMs? (arxiv.org) |
|
1 point by sonabinu on Oct 3, 2025 | hide | past | pdf | discuss
|
| 2305. |
Security Degradation in Iterative AI Code Generation (arxiv.org) |
|
1 point by chillax on Oct 3, 2025 | hide | past | pdf | discuss
|
| 2306. |
Dragon Hatchling: The Missing Link B. The Transformer and Models of the Brain (arxiv.org) |
|
6 points by polskibus on Oct 2, 2025 | hide | past | pdf | discuss
|
| 2307. |
The Missing Link Between the Transformer and Models of the Brain (arxiv.org) |
|
8 points by feelingsonice on Oct 2, 2025 | hide | past | pdf | 1 comment
|
| 2308. |
Efficient LLM:Bandwidth, Compute, Synchronization, and Capacity are all you need (arxiv.org) |
|
6 points by matt_d on Oct 1, 2025 | hide | past | pdf | discuss
|
| 2309. |
Introduction to Machine Learning(2024) (arxiv.org) |
|
1 point by runningmike on Oct 1, 2025 | hide | past | pdf | 1 comment
|
| 2310. |
The AI Productivity Index – LLMs by Economic Impact (arxiv.org) |
|
3 points by hereme888 on Oct 1, 2025 | hide | past | pdf | 1 comment
|
| More |