ML News
new
|
past
|
best
|
rss
|
submit
about
91.
Open-Source E2E FHE Implementation for Privacy-Preserving Llama 3 8B Inference
(
arxiv.org
)
2 points
by
simonpure
6 days ago
|
hide
|
past
|
pdf
|
discuss
92.
What and Whose Knowledge? Measuring Epistemic Diversity in Large Language Models
(
arxiv.org
)
2 points
by
daniel_iversen
7 days ago
|
hide
|
past
|
pdf
|
discuss
93.
Synthetic Hospital: Physician-Validated Longitudinal EHR Benchmark
(
arxiv.org
)
2 points
by
simonpure
8 days ago
|
hide
|
past
|
pdf
|
discuss
94.
Instrumental Monitor Evasion Emerges Under Ordinary Task Pressure
(
arxiv.org
)
2 points
by
sbulaev
8 days ago
|
hide
|
past
|
pdf
|
discuss
95.
Extracting and Characterizing Hidden Chain-of-Thought in Frontier Models
(
arxiv.org
)
2 points
by
jumploops
8 days ago
|
hide
|
past
|
pdf
|
discuss
96.
Scaling Laws for Neural Language Models (first "scaling laws" paper from 2020)
(
arxiv.org
)
2 points
by
thoughtpeddler
8 days ago
|
hide
|
past
|
pdf
|
discuss
97.
Memory Control Signals Emerge Before Action in Long Horizon Agents
(
arxiv.org
)
2 points
by
simonpure
9 days ago
|
hide
|
past
|
pdf
|
discuss
98.
Double Descent and Malign Overfitting in Diffusion Models
(
arxiv.org
)
2 points
by
Betelbuddy
10 days ago
|
hide
|
past
|
pdf
|
discuss
99.
Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale
(
arxiv.org
)
2 points
by
raahelb
10 days ago
|
hide
|
past
|
pdf
|
discuss
100.
HySparse2: Hybrid Sparse Attention with Two-Level KV Sharing
(
arxiv.org
)
2 points
by
ksec
10 days ago
|
hide
|
past
|
pdf
|
discuss
101.
Training a Language Model End-to-End in Rust: An Experience Report
(
arxiv.org
)
2 points
by
Brajeshwar
10 days ago
|
hide
|
past
|
pdf
|
discuss
102.
Temporal Straightening for Latent Planning
(
arxiv.org
)
2 points
by
gmays
11 days ago
|
hide
|
past
|
pdf
|
discuss
103.
Scaling Discovery Through Test-Time Communication
(
arxiv.org
)
2 points
by
simonpure
12 days ago
|
hide
|
past
|
pdf
|
discuss
104.
Complex KDA:Understanding and Enhancing the Expressivity of Kimi Delta Attention
(
arxiv.org
)
2 points
by
jul8234
11 days ago
|
hide
|
past
|
pdf
|
discuss
105.
An Introduction to Compression-Based Machine Learning
(
arxiv.org
)
2 points
by
jackhurwitz
12 days ago
|
hide
|
past
|
pdf
|
discuss
106.
Steerable Cultural Preference Optimization of Reward Models
(
arxiv.org
)
2 points
by
measurablefunc
12 days ago
|
hide
|
past
|
pdf
|
discuss
107.
A Black-Box Audit of Provider-Side Token Inflation in LLM Services
(
arxiv.org
)
2 points
by
donk8r
14 days ago
|
hide
|
past
|
pdf
|
discuss
108.
Federated Learning Is Not Private for Google GBoard Next Word Prediction [pdf]
(
arxiv.org
)
2 points
by
thunderbong
15 days ago
|
hide
|
past
|
pdf
|
discuss
109.
Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents
(
arxiv.org
)
2 points
by
Betelbuddy
16 days ago
|
hide
|
past
|
pdf
|
discuss
110.
200K-Token LLM Serving on a 24 GiB Laptop with Just-in-Time State Management
(
arxiv.org
)
2 points
by
Betelbuddy
16 days ago
|
hide
|
past
|
pdf
|
discuss
111.
Competitive Market Behavior of LLMs
(
arxiv.org
)
2 points
by
paulpauper
17 days ago
|
hide
|
past
|
pdf
|
discuss
112.
Coding Agents Have Converged: Why the SWE-Bench Leaderboard Can No Longer Order
(
arxiv.org
)
2 points
by
sbulaev
17 days ago
|
hide
|
past
|
pdf
|
discuss
113.
Continual Learning Mechanisms Compose for Long-Horizon Memorization
(
arxiv.org
)
2 points
by
donk8r
17 days ago
|
hide
|
past
|
pdf
|
discuss
114.
Divide, Consult, Conquer: Capability Laundering Through Aligned LLMs
(
arxiv.org
)
2 points
by
sbulaev
18 days ago
|
hide
|
past
|
pdf
|
discuss
115.
AdaBoost Does Not Always Cycle
(
arxiv.org
)
2 points
by
gone35
18 days ago
|
hide
|
past
|
pdf
|
discuss
116.
StepAudio 3 Gen Technical Report
(
arxiv.org
)
2 points
by
gmays
18 days ago
|
hide
|
past
|
pdf
|
discuss
117.
Breaking the Token Ceiling
(
arxiv.org
)
2 points
by
sonabinu
18 days ago
|
hide
|
past
|
pdf
|
discuss
118.
NCP-ArchPreview: 8.9B latent LM matches OLMo-3-7B on 51% of tokens
(
arxiv.org
)
2 points
by
iamsyr
19 days ago
|
hide
|
past
|
pdf
|
discuss
119.
MOBA: Mixture of Block Attention for Long-Context LLMs
(
arxiv.org
)
2 points
by
ur-whale
20 days ago
|
hide
|
past
|
pdf
|
discuss
120.
Datasets for Large Language Models: A Comprehensive Survey
(
arxiv.org
)
2 points
by
Anon84
21 days ago
|
hide
|
past
|
pdf
|
discuss
More
About
|
RSS
|
RSS (all)
|
HN arXiv