ML News
new
|
past
|
best
|
rss
|
submit
about
Stories from July 29, 2025 (UTC)
Go back a
day
,
month
, or
year
. Go forward a
day
.
1.
Supervised fine tuning on curated data is reinforcement learning
(
arxiv.org
)
71 points
by
GabrielBianconi
on Jul 29, 2025
|
hide
|
past
|
pdf
|
19 comments
2.
Query Agnostic Adversarial Triggers for Reasoning Models
(
arxiv.org
)
3 points
by
fzliu
on Jul 29, 2025
|
hide
|
past
|
pdf
|
discuss
3.
Language Model Can Be a Steganographic Privacy Leaking Agent
(
arxiv.org
)
3 points
by
dennis-tra
on Jul 29, 2025
|
hide
|
past
|
pdf
|
discuss
4.
TrimLLM: Progressive Layer Dropping for Domain-Specific LLMs
(
arxiv.org
)
2 points
by
pulkitsh1234
on Jul 29, 2025
|
hide
|
past
|
pdf
|
discuss
5.
SmallThinker: A Family of Efficient LLMs Natively Trained for Local Deployment
(
arxiv.org
)
2 points
by
limoce
on Jul 29, 2025
|
hide
|
past
|
pdf
|
discuss
About
|
RSS
|
RSS (all)
|
HN arXiv