ML News
new
|
past
|
best
|
rss
|
submit
about
1411.
Memex(RL): Scaling Long-Horizon LLM Agents via Indexed Experience Memory
(
arxiv.org
)
2 points
by
simonpure
213 days ago
|
hide
|
past
|
pdf
|
discuss
1412.
Asymmetric Goal Drift in Coding Agents Under Value Conflict
(
arxiv.org
)
1 point
by
lrakster
213 days ago
|
hide
|
past
|
pdf
|
discuss
1413.
Agentic Code Reasoning
(
arxiv.org
)
3 points
by
gmays
213 days ago
|
hide
|
past
|
pdf
|
discuss
1414.
General Agentic Memory via Deep Research
(
arxiv.org
)
2 points
by
gmays
214 days ago
|
hide
|
past
|
pdf
|
discuss
1415.
A Dual-LLM Policy for Reducing Noise in Agentic Program Repair
(
arxiv.org
)
1 point
by
azhenley
214 days ago
|
hide
|
past
|
pdf
|
discuss
1416.
Actor-Curator: Learning the Training Curriculum for RL Post-Training
(
arxiv.org
)
2 points
by
jonathanlight
214 days ago
|
hide
|
past
|
pdf
|
1 comment
1417.
Evaluating Theory of Mind and Internal Beliefs in LLM-Based Multi-Agent Systems
(
arxiv.org
)
1 point
by
Anon84
215 days ago
|
hide
|
past
|
pdf
|
discuss
1418.
DualPath: Breaking the Storage Bandwidth Bottleneck in Agentic LLM Inference
(
arxiv.org
)
2 points
by
nsoonhui
215 days ago
|
hide
|
past
|
pdf
|
1 comment
1419.
CuTe Layout Representation and Algebra
(
arxiv.org
)
4 points
by
matt_d
215 days ago
|
hide
|
past
|
pdf
|
discuss
1420.
Speculative Speculative Decoding (SSD)
(
arxiv.org
)
61 points
by
E-Reverance
215 days ago
|
hide
|
past
|
pdf
|
9 comments
1421.
How Well Does Agent Development Reflect Real-World Work?
(
arxiv.org
)
3 points
by
salkahfi
215 days ago
|
hide
|
past
|
pdf
|
discuss
1422.
130k Lines of Formal Topology: Simple and Cheap Autoformalization for Everyone?
(
arxiv.org
)
34 points
by
PaulHoule
215 days ago
|
hide
|
past
|
pdf
|
11 comments
1423.
Learning-Based Multi-Stage Strategy for Aircraft to Evade Missile
(
arxiv.org
)
1 point
by
rbanffy
215 days ago
|
hide
|
past
|
pdf
|
discuss
1424.
LeRobot: An Open-Source Library for End-to-End Robot Learning
(
arxiv.org
)
2 points
by
nill0
216 days ago
|
hide
|
past
|
pdf
|
discuss
1425.
CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation
(
arxiv.org
)
3 points
by
petethomas
216 days ago
|
hide
|
past
|
pdf
|
discuss
1426.
Run your agent 10 times – you won't get the same answer
(
arxiv.org
)
5 points
by
amanmehta1997
216 days ago
|
hide
|
past
|
pdf
|
1 comment
1427.
Language Model Contains Personality Subnetworks
(
arxiv.org
)
58 points
by
PaulHoule
216 days ago
|
hide
|
past
|
pdf
|
34 comments
1428.
Toward Guarantees for Clinical Reasoning in Vision Language Models
(
arxiv.org
)
2 points
by
tinarchitect
217 days ago
|
hide
|
past
|
pdf
|
discuss
1429.
Toward Guarantees for Clinical Reasoning in Vision Language Models
(
arxiv.org
)
5 points
by
barthelomew
217 days ago
|
hide
|
past
|
pdf
|
3 comments
1430.
FlyTrap: Attract autonomous drones with an adversarial umbrella
(
arxiv.org
)
2 points
by
fainpul
217 days ago
|
hide
|
past
|
pdf
|
1 comment
1431.
GPT detectors are biased against non-native English writers (2023)
(
arxiv.org
)
2 points
by
maxloh
218 days ago
|
hide
|
past
|
pdf
|
discuss
1432.
Latent-Space Communication in Heterogeneous Multi-Agent Systems
(
arxiv.org
)
7 points
by
ekaesmem
218 days ago
|
hide
|
past
|
pdf
|
1 comment
1433.
A Reinforcement Learning Environment for Automatic Code Optimization in MLIR
(
arxiv.org
)
1 point
by
matt_d
218 days ago
|
hide
|
past
|
pdf
|
discuss
1434.
Frontier AI Models Exhibit Sophisticated Reasoning in Simulated Nuclear Crises
(
arxiv.org
)
3 points
by
iamskeole
218 days ago
|
hide
|
past
|
pdf
|
discuss
1435.
Doc-to-LoRA: Learning to Instantly Internalize Contexts
(
arxiv.org
)
1 point
by
rbanffy
218 days ago
|
hide
|
past
|
pdf
|
discuss
1436.
Agents of Chaos
(
arxiv.org
)
4 points
by
ukuina
218 days ago
|
hide
|
past
|
pdf
|
1 comment
1437.
Kimi K2: Open Agentic Intelligence
(
arxiv.org
)
2 points
by
Anon84
219 days ago
|
hide
|
past
|
pdf
|
discuss
1438.
Prompt Repetition Improves Non-Reasoning LLMs
(
arxiv.org
)
1 point
by
tosh
219 days ago
|
hide
|
past
|
pdf
|
discuss
1439.
Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use
(
arxiv.org
)
2 points
by
Anon84
219 days ago
|
hide
|
past
|
pdf
|
discuss
1440.
Deep Learning: Our Year 1990-1991
(
arxiv.org
)
1 point
by
vinhnx
219 days ago
|
hide
|
past
|
pdf
|
discuss
More
About
|
RSS
|
RSS (all)
|
HN arXiv