ML News
new
|
past
|
best
|
rss
|
submit
about
1.
Process Matters More Than Output for Distinguishing Humans from Machines
(
arxiv.org
)
1 point
by
timshell
2 hours ago
|
hide
|
past
|
pdf
|
discuss
2.
SoftServe: A Scalable Quasi-Newton Method for Deep Learning
(
arxiv.org
)
1 point
by
E-Reverance
20 hours ago
|
hide
|
past
|
pdf
|
discuss
3.
PTXBench: Benchmarking and Adapting LLMs for GPU Kernel Optimization
(
arxiv.org
)
3 points
by
matt_d
23 hours ago
|
hide
|
past
|
pdf
|
discuss
4.
GPU-Initiated Communication: Dissecting Down to the Bone
(
arxiv.org
)
2 points
by
matt_d
1 day ago
|
hide
|
past
|
pdf
|
discuss
5.
SFT matches RL if you MCMC the training data first
(
arxiv.org
)
2 points
by
mrkn1
1 day ago
|
hide
|
past
|
pdf
|
discuss
6.
AI Agents Are Vulnerable to Radicalization
(
arxiv.org
)
3 points
by
Anon84
1 day ago
|
hide
|
past
|
pdf
|
discuss
7.
Fixing GRPO's credit assignment problem without evaluating every step
(
arxiv.org
)
23 points
by
mrkn1
1 day ago
|
hide
|
past
|
pdf
|
3 comments
8.
Decoding Looped Transformers Better for Almost Free
(
arxiv.org
)
1 point
by
mrkn1
1 day ago
|
hide
|
past
|
pdf
|
discuss
9.
Language Drift During RLVR Post-Training
(
arxiv.org
)
1 point
by
sbulaev
1 day ago
|
hide
|
past
|
pdf
|
discuss
10.
Removing Timing Shortcuts Improves Non-Invasive Brain-to-Text
(
arxiv.org
)
3 points
by
sbulaev
1 day ago
|
hide
|
past
|
pdf
|
discuss
11.
Scaling Laws for Looped Mixture of Experts
(
arxiv.org
)
2 points
by
matt_d
1 day ago
|
hide
|
past
|
pdf
|
discuss
12.
Covert Assistance: Helpful LLM Agents Evade Oversight in Multi-Agent Systems
(
arxiv.org
)
1 point
by
sbulaev
1 day ago
|
hide
|
past
|
pdf
|
1 comment
13.
Superhuman AI for Stratego
(
arxiv.org
)
5 points
by
droidjj
1 day ago
|
hide
|
past
|
pdf
|
1 comment
14.
When Fancy Eviction Fails: Rethinking Cache Replacement for LLM Prefix Reuse
(
arxiv.org
)
1 point
by
matt_d
2 days ago
|
hide
|
past
|
pdf
|
discuss
15.
Pretraining Latent Information Feedback Transformers with Teacher Supervision
(
arxiv.org
)
3 points
by
gmays
2 days ago
|
hide
|
past
|
pdf
|
discuss
16.
Decode-Latency Feedback Prefill: A Model-Free Controller
(
arxiv.org
)
1 point
by
gauravapiscean
2 days ago
|
hide
|
past
|
pdf
|
1 comment
17.
Thinking Before Thinking: Scaling Agentic Inference Through Meta-Reasoning
(
arxiv.org
)
2 points
by
theanonymousone
2 days ago
|
hide
|
past
|
pdf
|
discuss
18.
Context Language Models
(
arxiv.org
)
175 points
by
emersonmacro
2 days ago
|
hide
|
past
|
pdf
|
51 comments
19.
Learning Steganography Is Easy, Learning Steganographic Reasoning Is Hard
(
arxiv.org
)
1 point
by
sbulaev
2 days ago
|
hide
|
past
|
pdf
|
discuss
20.
Thinking Before Thinking: Scaling Agentic Inference Through Meta-Reasoning
(
arxiv.org
)
3 points
by
matt_d
3 days ago
|
hide
|
past
|
pdf
|
discuss
21.
AI as a Compiler: Compiling Triton kernels without the Triton compiler
(
arxiv.org
)
2 points
by
matt_d
3 days ago
|
hide
|
past
|
pdf
|
discuss
22.
Distillation Defenses Easily Break After Reinforcement Learning
(
arxiv.org
)
2 points
by
ollybritton
3 days ago
|
hide
|
past
|
pdf
|
discuss
23.
Sycophantic Chatbots Cause Delusional Spiraling, Even in Ideal Bayesians
(
arxiv.org
)
4 points
by
luispa
3 days ago
|
hide
|
past
|
pdf
|
discuss
24.
Decoupled DiLoCo for Resilient Distributed Pre-Training
(
arxiv.org
)
1 point
by
lawrenceyan
3 days ago
|
hide
|
past
|
pdf
|
discuss
25.
Purlin: Separating Orchestration from the Datapath of Collectives
(
arxiv.org
)
3 points
by
matt_d
3 days ago
|
hide
|
past
|
pdf
|
discuss
26.
Context Language Models
(
arxiv.org
)
6 points
by
tomatomatomato
3 days ago
|
hide
|
past
|
pdf
|
1 comment
27.
Practical Secrets Extraction Against Black-Box LLMs
(
arxiv.org
)
1 point
by
sbulaev
3 days ago
|
hide
|
past
|
pdf
|
discuss
28.
Shutdown Sabotage Propensities in Multi- Agent Systems
(
arxiv.org
)
1 point
by
baxtr
3 days ago
|
hide
|
past
|
pdf
|
discuss
29.
Compiling Triton kernels without the Triton compiler
(
arxiv.org
)
3 points
by
50kIters
3 days ago
|
hide
|
past
|
pdf
|
discuss
30.
Jev-as-a-Judge: Accept When Confident, Escalate When Unsure
(
arxiv.org
)
2 points
by
nico
3 days ago
|
hide
|
past
|
pdf
|
discuss
More
About
|
RSS
|
RSS (all)
|
HN arXiv