about
91. Open-Source E2E FHE Implementation for Privacy-Preserving Llama 3 8B Inference (arxiv.org)
2 points by simonpure 6 days ago | hide | past | pdf | discuss
92. What and Whose Knowledge? Measuring Epistemic Diversity in Large Language Models (arxiv.org)
2 points by daniel_iversen 7 days ago | hide | past | pdf | discuss
93. Synthetic Hospital: Physician-Validated Longitudinal EHR Benchmark (arxiv.org)
2 points by simonpure 8 days ago | hide | past | pdf | discuss
94. Instrumental Monitor Evasion Emerges Under Ordinary Task Pressure (arxiv.org)
2 points by sbulaev 8 days ago | hide | past | pdf | discuss
95. Extracting and Characterizing Hidden Chain-of-Thought in Frontier Models (arxiv.org)
2 points by jumploops 8 days ago | hide | past | pdf | discuss
96. Scaling Laws for Neural Language Models (first "scaling laws" paper from 2020) (arxiv.org)
2 points by thoughtpeddler 8 days ago | hide | past | pdf | discuss
97. Memory Control Signals Emerge Before Action in Long Horizon Agents (arxiv.org)
2 points by simonpure 9 days ago | hide | past | pdf | discuss
98. Double Descent and Malign Overfitting in Diffusion Models (arxiv.org)
2 points by Betelbuddy 10 days ago | hide | past | pdf | discuss
99. Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale (arxiv.org)
2 points by raahelb 10 days ago | hide | past | pdf | discuss
100. HySparse2: Hybrid Sparse Attention with Two-Level KV Sharing (arxiv.org)
2 points by ksec 10 days ago | hide | past | pdf | discuss
101. Training a Language Model End-to-End in Rust: An Experience Report (arxiv.org)
2 points by Brajeshwar 10 days ago | hide | past | pdf | discuss
102. Temporal Straightening for Latent Planning (arxiv.org)
2 points by gmays 11 days ago | hide | past | pdf | discuss
103. Scaling Discovery Through Test-Time Communication (arxiv.org)
2 points by simonpure 12 days ago | hide | past | pdf | discuss
104. Complex KDA:Understanding and Enhancing the Expressivity of Kimi Delta Attention (arxiv.org)
2 points by jul8234 11 days ago | hide | past | pdf | discuss
105. An Introduction to Compression-Based Machine Learning (arxiv.org)
2 points by jackhurwitz 12 days ago | hide | past | pdf | discuss
106. Steerable Cultural Preference Optimization of Reward Models (arxiv.org)
2 points by measurablefunc 12 days ago | hide | past | pdf | discuss
107. A Black-Box Audit of Provider-Side Token Inflation in LLM Services (arxiv.org)
2 points by donk8r 14 days ago | hide | past | pdf | discuss
108. Federated Learning Is Not Private for Google GBoard Next Word Prediction [pdf] (arxiv.org)
2 points by thunderbong 15 days ago | hide | past | pdf | discuss
109. Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents (arxiv.org)
2 points by Betelbuddy 16 days ago | hide | past | pdf | discuss
110. 200K-Token LLM Serving on a 24 GiB Laptop with Just-in-Time State Management (arxiv.org)
2 points by Betelbuddy 16 days ago | hide | past | pdf | discuss
111. Competitive Market Behavior of LLMs (arxiv.org)
2 points by paulpauper 17 days ago | hide | past | pdf | discuss
112. Coding Agents Have Converged: Why the SWE-Bench Leaderboard Can No Longer Order (arxiv.org)
2 points by sbulaev 17 days ago | hide | past | pdf | discuss
113. Continual Learning Mechanisms Compose for Long-Horizon Memorization (arxiv.org)
2 points by donk8r 17 days ago | hide | past | pdf | discuss
114. Divide, Consult, Conquer: Capability Laundering Through Aligned LLMs (arxiv.org)
2 points by sbulaev 18 days ago | hide | past | pdf | discuss
115. AdaBoost Does Not Always Cycle (arxiv.org)
2 points by gone35 18 days ago | hide | past | pdf | discuss
116. StepAudio 3 Gen Technical Report (arxiv.org)
2 points by gmays 18 days ago | hide | past | pdf | discuss
117. Breaking the Token Ceiling (arxiv.org)
2 points by sonabinu 18 days ago | hide | past | pdf | discuss
118. NCP-ArchPreview: 8.9B latent LM matches OLMo-3-7B on 51% of tokens (arxiv.org)
2 points by iamsyr 19 days ago | hide | past | pdf | discuss
119. MOBA: Mixture of Block Attention for Long-Context LLMs (arxiv.org)
2 points by ur-whale 20 days ago | hide | past | pdf | discuss
120. Datasets for Large Language Models: A Comprehensive Survey (arxiv.org)
2 points by Anon84 21 days ago | hide | past | pdf | discuss