about
3151. A General Theoretical Paradigm to Understand Learning from Human Preferences (arxiv.org)
2 points by yenniejun111 on May 15, 2025 | hide | past | pdf | discuss
3152. DeepSeek-V3: Achieving Efficient LLM Scaling with 2,048 GPUs (arxiv.org)
7 points by qtwhat on May 15, 2025 | hide | past | pdf | 1 comment
3153. Online Isolation Forest (arxiv.org)
2 points by badmonster on May 15, 2025 | hide | past | pdf | discuss
3154. Understanding Perception and Reasoning Through Model Merging (arxiv.org)
2 points by veryluckyxyz on May 15, 2025 | hide | past | pdf | discuss
3155. LLMs get lost in multi-turn conversation (arxiv.org)
374 points by simonpure on May 15, 2025 | hide | past | pdf | 259 comments
3156. SEM-CTRL: Semantically Controlled Decoding (arxiv.org)
3 points by tough on May 14, 2025 | hide | past | pdf | discuss
3157. Bang for the Buck: Vector Search on Cloud CPUs (arxiv.org)
5 points by ashvardanian on May 14, 2025 | hide | past | pdf | discuss
3158. Do Words Reflect Beliefs? Evaluating Belief Depth in Large Language Models (arxiv.org)
1 point by PaulHoule on May 14, 2025 | hide | past | pdf | discuss
3159. IterGen: Iterative Semantic-Aware Structured LLM Generation with Backtracking (arxiv.org)
1 point by tough on May 14, 2025 | hide | past | pdf | discuss
3160. ROCODE: Integrating Backtracking Mechanism and Program Analysis in LLMs for Code (arxiv.org)
1 point by tough on May 14, 2025 | hide | past | pdf | discuss
3161. SRLCG: Self-Rectified Large-Scale Code Generation, CoT, Dynamic Backtracking (arxiv.org)
1 point by tough on May 14, 2025 | hide | past | pdf | discuss
3162. CRANE: Reasoning with Constrained LLM Generation (arxiv.org)
1 point by tough on May 13, 2025 | hide | past | pdf | discuss
3163. Type-constrained code generation with language models (arxiv.org)
257 points by tough on May 13, 2025 | hide | past | pdf | 127 comments
3164. AWRS SMC: Fast new algorithm for guiding LLMs as Bayesian inference (arxiv.org)
2 points by benlipkin on May 13, 2025 | hide | past | pdf | discuss
3165. Can Third-Parties Read Our Emotions? (arxiv.org)
2 points by PaulHoule on May 13, 2025 | hide | past | pdf | discuss
3166. Intellect-2: A Reasoning Model Trained Through Globally Decentralized RL (arxiv.org)
1 point by nkko on May 13, 2025 | hide | past | pdf | discuss
3167. Rethinking Memory in AI: Taxonomy, Operations, Topics, and Future Directions (arxiv.org)
5 points by wjSgoWPm5bWAhXB on May 13, 2025 | hide | past | pdf | discuss
3168. In-Context Learning can distort the relationship between likelihoods and fitness (arxiv.org)
1 point by PaulHoule on May 13, 2025 | hide | past | pdf | discuss
3169. LithOS: An Operating System for Efficient Machine Learning on GPUs (arxiv.org)
3 points by PaulHoule on May 13, 2025 | hide | past | pdf | discuss
3170. Base Models Beat Aligned Models at Randomness and Creativity (arxiv.org)
1 point by todsacerdoti on May 13, 2025 | hide | past | pdf | discuss
3171. Backslash: Rate Constrained Optimized Training of Large Language Models (arxiv.org)
3 points by PaulHoule on May 13, 2025 | hide | past | pdf | discuss
3172. TransMLA: Multi-head latent attention is all you need (arxiv.org)
123 points by ocean_moist on May 13, 2025 | hide | past | pdf | 32 comments
3173. Zero-shot forecasting of chaotic systems (arxiv.org)
2 points by wil3 on May 12, 2025 | hide | past | pdf | discuss
3174. LLMs Outperform Experts on Challenging Biology Benchmarks (arxiv.org)
1 point by belter on May 12, 2025 | hide | past | pdf | discuss
3175. Byte latent transformer: Patches scale better than tokens (2024) (arxiv.org)
107 points by dlojudice on May 12, 2025 | hide | past | pdf | 22 comments
3176. Large Language Models Are Autonomous Cyber Defenders (arxiv.org)
1 point by aibrother on May 11, 2025 | hide | past | pdf | discuss
3177. Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org)
88 points by leodriesch on May 11, 2025 | hide | past | pdf | 19 comments
3178. Learning Adaptive Parallel Reasoning with Language Models (arxiv.org)
2 points by PaulHoule on May 11, 2025 | hide | past | pdf | discuss
3179. Generating Physically Stable and Buildable Lego Designs from Text (arxiv.org)
1 point by aibrother on May 11, 2025 | hide | past | pdf | discuss
3180. Reasoning Models Don't Always Say What They Think (arxiv.org)
2 points by badmonster on May 10, 2025 | hide | past | pdf | discuss