about
3721. Test-time scaling new approach: extra test-time compute improves LLM reasoning (arxiv.org)
2 points by TaurenHunter on Feb 8, 2025 | hide | past | pdf | discuss
3722. CoCoNUT: Structural Code Understanding does not fall out of a tree (arxiv.org)
2 points by PaulHoule on Feb 8, 2025 | hide | past | pdf | discuss
3723. DocVLM: Make Your VLM an Efficient Reader (arxiv.org)
2 points by fzliu on Feb 7, 2025 | hide | past | pdf | discuss
3724. Transformers Boost the Performance of Decision Trees on Tabular Data (arxiv.org)
2 points by TaurenHunter on Feb 7, 2025 | hide | past | pdf | discuss
3725. Vision language models are blind (2024) [pdf] (arxiv.org)
3 points by thegeomaster on Feb 7, 2025 | hide | past | pdf | discuss
3726. IServe: An Intent-Based Serving System for LLMs (arxiv.org)
1 point by PaulHoule on Feb 7, 2025 | hide | past | pdf | discuss
3727. Global Optimization of Black-Box Functions with Unknown Lipschitz Constants (arxiv.org)
3 points by fofoz on Feb 7, 2025 | hide | past | pdf | discuss
3728. Gold-Medalist Performance in Solving Olympiad Geometry with AlphaGeometry2 (arxiv.org)
64 points by hnhn34 on Feb 7, 2025 | hide | past | pdf | 5 comments
3729. HippoRAG: Neurobiologically Inspired Long-Term Memory for LLMs (2024) (arxiv.org)
65 points by veryluckyxyz on Feb 7, 2025 | hide | past | pdf | 4 comments
3730. Ml.net (arxiv.org)
1 point by colonCapitalDee on Feb 7, 2025 | hide | past | pdf | discuss
3731. Robust autonomy emerges from self-play (arxiv.org)
140 points by reqo on Feb 7, 2025 | hide | past | pdf | 62 comments
3732. Fault Localization via Fine-Tuning LLMs with Mutation Generated Stack Traces (arxiv.org)
3 points by PaulHoule on Feb 7, 2025 | hide | past | pdf | discuss
3733. Meta AI's latest research: improved LLM reasoning with Latent Tokens (arxiv.org)
3 points by lessisgood123 on Feb 7, 2025 | hide | past | pdf | discuss
3734. Understanding Why Adam Outperforms SGD: Gradient Heterogeneity in Transformers (arxiv.org)
3 points by fofoz on Feb 6, 2025 | hide | past | pdf | discuss
3735. Demystifying Long Chain-of-Thought Reasoning in LLMs (arxiv.org)
2 points by sebg on Feb 6, 2025 | hide | past | pdf | discuss
3736. Develop AI Agents for System Engineering in Factorio (arxiv.org)
3 points by Jimmc414 on Feb 6, 2025 | hide | past | pdf | discuss
3737. Optimizing LLM Persuasion with Personalization and Fabricated Statistics (arxiv.org)
2 points by PaulHoule on Feb 6, 2025 | hide | past | pdf | discuss
3738. High-Fidelity Simultaneous Speech-to-Speech Translation (arxiv.org)
6 points by exgrv on Feb 6, 2025 | hide | past | pdf | 1 comment
3739. The Hyperfitting Phenomenon: Sharpening and Stabilizing LLMs (arxiv.org)
3 points by superidiot1932 on Feb 6, 2025 | hide | past | pdf | discuss
3740. LIMO: Less Is More for Reasoning (arxiv.org)
2 points by maksimur on Feb 6, 2025 | hide | past | pdf | discuss
3741. Pre-Trained Large Language Models Use Fourier Features for Addition (2024) (arxiv.org)
149 points by Kye on Feb 6, 2025 | hide | past | pdf | 40 comments
3742. SmolLM2: When Smol Goes Big – Data-Centric Training of a Small Language Model (arxiv.org)
1 point by rahimnathwani on Feb 6, 2025 | hide | past | pdf | discuss
3743. Harmonic Loss Trains Interpretable AI Models (arxiv.org)
1 point by fzliu on Feb 5, 2025 | hide | past | pdf | discuss
3744. Sundial: A Family of Highly Capable Time Series Foundation Models (arxiv.org)
3 points by fofoz on Feb 5, 2025 | hide | past | pdf | discuss
3745. Digital Agent outperforms o1 by 15% – trained with new RL-variant similar to R1 (arxiv.org)
11 points by let_tim_cook_ on Feb 5, 2025 | hide | past | pdf | discuss
3746. Harmonic Loss Trains Interpretable AI Models (arxiv.org)
4 points by lemonfever on Feb 5, 2025 | hide | past | pdf | 2 comments
3747. Over-Tokenized Transformer: Vocabulary Is Generally Worth Scaling (arxiv.org)
2 points by famouswaffles on Feb 4, 2025 | hide | past | pdf | discuss
3748. OmniHuman-1: Scaling-Up of One-Stage Conditioned Human Animation Models (arxiv.org)
1 point by neom on Feb 4, 2025 | hide | past | pdf | discuss
3749. Querying Databases with Function Calling (arxiv.org)
3 points by tosh on Feb 4, 2025 | hide | past | pdf | discuss
3750. DeepRAG: Thinking to retrieval step by step for large language models (arxiv.org)
191 points by fofoz on Feb 4, 2025 | hide | past | pdf | 29 comments