ML News
new
|
past
|
best
|
rss
|
submit
about
1081.
When innocent tools form dangerous chains to jailbreak LLM agents
(
arxiv.org
)
2 points
by
leecoursey
152 days ago
|
hide
|
past
|
pdf
|
discuss
1082.
A Theory of Generalization in Deep Learning
(
arxiv.org
)
4 points
by
E-Reverance
152 days ago
|
hide
|
past
|
pdf
|
discuss
1083.
GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents
(
arxiv.org
)
163 points
by
gmays
152 days ago
|
hide
|
past
|
pdf
|
32 comments
1084.
The Last Human-Written Paper: Agent-Native Research Artifacts
(
arxiv.org
)
2 points
by
amberjcjj
152 days ago
|
hide
|
past
|
pdf
|
discuss
1085.
Faster RL Post-Training Rollouts via System-Integrated Speculative Decoding
(
arxiv.org
)
1 point
by
gmays
152 days ago
|
hide
|
past
|
pdf
|
discuss
1086.
DeepSeek V4's indexer OOMs at 65K context. We got it to 1M in 6G
(
arxiv.org
)
8 points
by
OsamaJaber
152 days ago
|
hide
|
past
|
pdf
|
discuss
1087.
Process-Level Reward Modeling for Agentic Data Analysis
(
arxiv.org
)
4 points
by
gmays
153 days ago
|
hide
|
past
|
pdf
|
discuss
1088.
Transformers Are Inherently Succinct (2025)
(
arxiv.org
)
62 points
by
bearseascape
153 days ago
|
hide
|
past
|
pdf
|
9 comments
1089.
Exploring LLM biases to manipulate AI search overview
(
arxiv.org
)
1 point
by
Brajeshwar
153 days ago
|
hide
|
past
|
pdf
|
discuss
1090.
Hallucination Is Inevitable: An Innate Limitation of Large Language Models (2025)
(
arxiv.org
)
14 points
by
drob518
153 days ago
|
hide
|
past
|
pdf
|
11 comments
1091.
Story of Two GPUs: Characterizing the Resilience of Hopper H100 and Ampere A100
(
arxiv.org
)
2 points
by
rbanffy
153 days ago
|
hide
|
past
|
pdf
|
discuss
1092.
Iarpa Trojans in Artificial Intelligence
(
arxiv.org
)
2 points
by
hlynurd
153 days ago
|
hide
|
past
|
pdf
|
discuss
1093.
Learning Randomized Reductions
(
arxiv.org
)
2 points
by
matt_d
154 days ago
|
hide
|
past
|
pdf
|
discuss
1094.
New research on analyzing and predicting token consumption of coding agents
(
arxiv.org
)
4 points
by
jiaxinpei
154 days ago
|
hide
|
past
|
pdf
|
1 comment
1095.
Learning Pseudorandom Numbers with Transformers
(
arxiv.org
)
11 points
by
pizza
154 days ago
|
hide
|
past
|
pdf
|
3 comments
1096.
LLMs can hide text in other text of the same length
(
arxiv.org
)
5 points
by
m-hodges
155 days ago
|
hide
|
past
|
pdf
|
discuss
1097.
A Note on TurboQuant and the Earlier Eden Work
(
arxiv.org
)
2 points
by
amitport
155 days ago
|
hide
|
past
|
pdf
|
discuss
1098.
Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework
(
arxiv.org
)
1 point
by
gmays
155 days ago
|
hide
|
past
|
pdf
|
discuss
1099.
AI Self-preferencing in Algorithmic Hiring: Empirical Evidence and Insights
(
arxiv.org
)
335 points
by
laurex
155 days ago
|
hide
|
past
|
pdf
|
178 comments
1100.
Refusal in Language Models Is Mediated by a Single Direction
(
arxiv.org
)
118 points
by
fagnerbrack
155 days ago
|
hide
|
past
|
pdf
|
45 comments
1101.
Tessera: Unlocking Heterogeneous GPUs Through Kernel-Granularity Disaggregation
(
arxiv.org
)
1 point
by
matt_d
156 days ago
|
hide
|
past
|
pdf
|
discuss
1102.
Xmemory: Benchmarking Structured AI Memory Against RAG and Hybrid RAG
(
arxiv.org
)
9 points
by
alex_petrov
156 days ago
|
hide
|
past
|
pdf
|
2 comments
1103.
Agentic Harness Engineering
(
arxiv.org
)
15 points
by
Anon84
157 days ago
|
hide
|
past
|
pdf
|
discuss
1104.
MathDuels: Evaluating LLMs as Problem Posers and Solvers
(
arxiv.org
)
2 points
by
matt_d
157 days ago
|
hide
|
past
|
pdf
|
1 comment
1105.
Performance Analysis of AI Query Approximation Using Lightweight Proxy Models
(
arxiv.org
)
1 point
by
tanelpoder
157 days ago
|
hide
|
past
|
pdf
|
discuss
1106.
Kernel Contracts: A Spec. Language for Correctness Across Heterogeneous Silicon
(
arxiv.org
)
1 point
by
matt_d
157 days ago
|
hide
|
past
|
pdf
|
discuss
1107.
Outcome Rewards Do Not Guarantee Verifiable or Causally Important Reasoning
(
arxiv.org
)
2 points
by
krackers
157 days ago
|
hide
|
past
|
pdf
|
1 comment
1108.
Prism: Demystifying Retention and Interaction in Mid-Training
(
arxiv.org
)
1 point
by
mdp2021
157 days ago
|
hide
|
past
|
pdf
|
discuss
1109.
Fast GPU Linear Algebra via Compile Time Expression Fusion
(
arxiv.org
)
12 points
by
matt_d
157 days ago
|
hide
|
past
|
pdf
|
discuss
1110.
LLMs Corrupt Your Documents When You Delegate
(
arxiv.org
)
4 points
by
uxhacker
157 days ago
|
hide
|
past
|
pdf
|
discuss
More
About
|
RSS
|
RSS (all)
|
HN arXiv