ML News
new
|
past
|
best
|
rss
|
submit
about
151.
Decoupled DiLoCo for Resilient Distributed Pre-Training
(
arxiv.org
)
1 point
by
lawrenceyan
3 days ago
|
hide
|
past
|
pdf
|
discuss
152.
Practical Secrets Extraction Against Black-Box LLMs
(
arxiv.org
)
1 point
by
sbulaev
3 days ago
|
hide
|
past
|
pdf
|
discuss
153.
Shutdown Sabotage Propensities in Multi- Agent Systems
(
arxiv.org
)
1 point
by
baxtr
3 days ago
|
hide
|
past
|
pdf
|
discuss
154.
Pluralis Towards a Multicultural Multimodal, Multilingual Benchmark for AI Risk
(
arxiv.org
)
1 point
by
thinkevolve
4 days ago
|
hide
|
past
|
pdf
|
discuss
155.
AmpleGCG: Learning a Universal Generative Model for Jailbreaking
(
arxiv.org
)
1 point
by
Anon84
4 days ago
|
hide
|
past
|
pdf
|
discuss
156.
Learning How to Forget: Fine-Tuning for Long-Context Sparse Attention
(
arxiv.org
)
1 point
by
theanonymousone
4 days ago
|
hide
|
past
|
pdf
|
discuss
157.
Fathom: Per-query read depth for sparse decoding over offloaded KV caches
(
arxiv.org
)
1 point
by
vivekkalyanaran
4 days ago
|
hide
|
past
|
pdf
|
discuss
158.
User Model Extraction via Belief Self-Distillation
(
arxiv.org
)
1 point
by
sbulaev
5 days ago
|
hide
|
past
|
pdf
|
discuss
159.
EnigmaForge – an LLM benchmark where the question is hidden in the story
(
arxiv.org
)
1 point
by
robottwo
5 days ago
|
hide
|
past
|
pdf
|
discuss
160.
CliffCompaction: Cost-Efficient Compaction for Long-Horizon Coding Agents
(
arxiv.org
)
1 point
by
6bitquant
5 days ago
|
hide
|
past
|
pdf
|
discuss
161.
TuxBot: Semantic-Aware Online OS Tuning with Large Language Models
(
arxiv.org
)
1 point
by
matt_d
6 days ago
|
hide
|
past
|
pdf
|
discuss
162.
Just Ask Jev: Reinforcement Learning for Calibrated Decisions
(
arxiv.org
)
1 point
by
Anon84
7 days ago
|
hide
|
past
|
pdf
|
discuss
163.
Grow the Harness, Not the Context
(
arxiv.org
)
1 point
by
acossta
7 days ago
|
hide
|
past
|
pdf
|
discuss
164.
Codetta: High-Capacity, Keyless, and Undetectable Multi-Agent Collusion
(
arxiv.org
)
1 point
by
sbulaev
7 days ago
|
hide
|
past
|
pdf
|
discuss
165.
Skill-Guided Mining and Compilation of LLM Agent Traces
(
arxiv.org
)
1 point
by
nlpnerd
8 days ago
|
hide
|
past
|
pdf
|
discuss
166.
Artificial Kuramoto Oscillatory Neurons
(
arxiv.org
)
1 point
by
jerlendds
8 days ago
|
hide
|
past
|
pdf
|
discuss
167.
Entropy-Based Guided Collaboration in Heterogeneous LLM Multi-Agent Systems
(
arxiv.org
)
1 point
by
wslh
8 days ago
|
hide
|
past
|
pdf
|
discuss
168.
Control-Token Injection Suppresses Chain-of-Thought and Defeats Reasoning-Based
(
arxiv.org
)
1 point
by
sbulaev
8 days ago
|
hide
|
past
|
pdf
|
discuss
169.
Learning to Discover Interesting Mathematics
(
arxiv.org
)
1 point
by
E-Reverance
9 days ago
|
hide
|
past
|
pdf
|
discuss
170.
LensVLM: Selective Context Expansion for Compressed Visual Representation OfText
(
arxiv.org
)
1 point
by
KitN
9 days ago
|
hide
|
past
|
pdf
|
discuss
171.
Log-Depth Recurrent Language Modeling
(
arxiv.org
)
1 point
by
E-Reverance
10 days ago
|
hide
|
past
|
pdf
|
discuss
172.
Greedy Decoding Is Not Precision-Invariant: Cross-Precision Output Divergence
(
arxiv.org
)
1 point
by
sbulaev
10 days ago
|
hide
|
past
|
pdf
|
discuss
173.
LatentPort: Cross-model recurrent state transfer without prefix replay
(
arxiv.org
)
1 point
by
mmprotest
10 days ago
|
hide
|
past
|
pdf
|
discuss
174.
PICARD: Parsing Incrementally for Constrained Auto-Regressive Decoding from LLMs
(
arxiv.org
)
1 point
by
Bluestein
10 days ago
|
hide
|
past
|
pdf
|
discuss
175.
Specs cut defects in AI-generated code from 148 to 23 across five models
(
arxiv.org
)
1 point
by
sandeepdhuri
11 days ago
|
hide
|
past
|
pdf
|
discuss
176.
Emergent Collusion in Long-Horizon LLM Agent Interaction
(
arxiv.org
)
1 point
by
sbulaev
11 days ago
|
hide
|
past
|
pdf
|
discuss
177.
Measuring behavioral signals of LLM through psychometric profiling
(
arxiv.org
)
1 point
by
anigbrowl
11 days ago
|
hide
|
past
|
pdf
|
discuss
178.
A self-evolving agentic system for automated execution of biological protocols
(
arxiv.org
)
1 point
by
lawrenceyan
11 days ago
|
hide
|
past
|
pdf
|
discuss
179.
Agents That Edit Documents: Measuring Agentic PDF Forgery Against a Non-Agentic
(
arxiv.org
)
1 point
by
sbulaev
11 days ago
|
hide
|
past
|
pdf
|
discuss
180.
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses
(
arxiv.org
)
1 point
by
Betelbuddy
11 days ago
|
hide
|
past
|
pdf
|
discuss
More
About
|
RSS
|
RSS (all)
|
HN arXiv