about
3811. Advancing Language Model Reasoning Through RL and Inference Scaling (arxiv.org)
2 points by frozenseven on Jan 25, 2025 | hide | past | pdf | discuss
3812. DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL (arxiv.org)
1351 points by gradus_ad on Jan 25, 2025 | hide | past | pdf | 1056 comments
3813. Position Information Emerges in Causal Transformers Without Positional Encoding (arxiv.org)
3 points by PaulHoule on Jan 25, 2025 | hide | past | pdf | discuss
3814. Kimi K1.5 Technical Report (arxiv.org)
4 points by anticensor on Jan 25, 2025 | hide | past | pdf | discuss
3815. Generative Adversarial Neural Network Acceleration with Silicon Photonics (arxiv.org)
2 points by belter on Jan 24, 2025 | hide | past | pdf | discuss
3816. Hallucinations Can Improve Large Language Models in Drug Discovery (arxiv.org)
1 point by keepit on Jan 24, 2025 | hide | past | pdf | discuss
3817. Frontier AI systems have surpassed the self-replicating red line (arxiv.org)
1 point by programd on Jan 24, 2025 | hide | past | pdf | discuss
3818. Dissecting the NVIDIA Hopper Architecture through Microbenchmarking (arxiv.org)
2 points by matt_d on Jan 24, 2025 | hide | past | pdf | discuss
3819. Tell me about yourself: LLMs are aware of their learned behaviors (arxiv.org)
2 points by famouswaffles on Jan 24, 2025 | hide | past | pdf | discuss
3820. Evolution and the Knightian Blindspot of Machine Learning (arxiv.org)
2 points by jal278 on Jan 24, 2025 | hide | past | pdf | discuss
3821. Implicit Chain of Thought Reasoning via Knowledge Distillation (arxiv.org)
1 point by Jimmc414 on Jan 24, 2025 | hide | past | pdf | discuss
3822. UI-Tars: Pioneering Automated GUI Interaction with Native Agents (arxiv.org)
2 points by msoad on Jan 23, 2025 | hide | past | pdf | discuss
3823. Super Tiny Language Models (pdf) (arxiv.org)
3 points by nickpsecurity on Jan 23, 2025 | hide | past | pdf | 1 comment
3824. Ask HN: How Do You Align Evaluator LLMs with Subject Matter Experts? (arxiv.org)
2 points by MutedEstate45 on Jan 23, 2025 | hide | past | pdf | 1 comment
3825. Foundations of Large Language Models (arxiv.org)
219 points by pkoird on Jan 23, 2025 | hide | past | pdf | 20 comments
3826. The Mathematics of Artificial Intelligence (arxiv.org)
4 points by sieste on Jan 23, 2025 | hide | past | pdf | discuss
3827. Lossless Compression of Vector IDs for Approximate Nearest Neighbor Search (arxiv.org)
151 points by fzliu on Jan 22, 2025 | hide | past | pdf | 6 comments
3828. Test-time regression: a unifying framework for designing sequence models (arxiv.org)
1 point by ketothekingdom on Jan 22, 2025 | hide | past | pdf | discuss
3829. Can LLMs demonstrate behavioral self-awareness? (arxiv.org)
3 points by omarsar on Jan 22, 2025 | hide | past | pdf | 1 comment
3830. Evolving Deeper LLM Thinking (arxiv.org)
1 point by simonpure on Jan 22, 2025 | hide | past | pdf | discuss
3831. Physics of Skill Learning (arxiv.org)
1 point by elashri on Jan 22, 2025 | hide | past | pdf | discuss
3832. VideoWorld: Exploring Knowledge Learning from Unlabeled Video (arxiv.org)
2 points by TaurenHunter on Jan 22, 2025 | hide | past | pdf | discuss
3833. Flame: A small language model for spreadsheet formulas (2023) (arxiv.org)
117 points by azhenley on Jan 22, 2025 | hide | past | pdf | 18 comments
3834. Tensor Product Attention Is All You Need (arxiv.org)
160 points by eunos on Jan 22, 2025 | hide | past | pdf | 104 comments
3835. Training DNN in O(1) Memory Using Random Matrices (arxiv.org)
2 points by tnch on Jan 21, 2025 | hide | past | pdf | discuss
3836. Confidence-Based Estimators for Predictive Performance in Model Monitoring (arxiv.org)
1 point by santiviquez on Jan 21, 2025 | hide | past | pdf | discuss
3837. Vision-Language Models Do Not Understand Negation (arxiv.org)
2 points by kjhughes on Jan 21, 2025 | hide | past | pdf | discuss
3838. Accelerating Retrieval-Augmented Generation (arxiv.org)
1 point by handfuloflight on Jan 20, 2025 | hide | past | pdf | discuss
3839. Evolving Deeper LLM Thinking (arxiv.org)
12 points by hardmaru on Jan 20, 2025 | hide | past | pdf | discuss
3840. Enhancing Chat Language Models: Scaling High-Quality Instructional Conversations (arxiv.org)
1 point by TaurenHunter on Jan 19, 2025 | hide | past | pdf | discuss