| 3721. |
Test-time scaling new approach: extra test-time compute improves LLM reasoning (arxiv.org) |
|
2 points by TaurenHunter on Feb 8, 2025 | hide | past | pdf | discuss
|
| 3722. |
CoCoNUT: Structural Code Understanding does not fall out of a tree (arxiv.org) |
|
2 points by PaulHoule on Feb 8, 2025 | hide | past | pdf | discuss
|
| 3723. |
DocVLM: Make Your VLM an Efficient Reader (arxiv.org) |
|
2 points by fzliu on Feb 7, 2025 | hide | past | pdf | discuss
|
| 3724. |
Transformers Boost the Performance of Decision Trees on Tabular Data (arxiv.org) |
|
2 points by TaurenHunter on Feb 7, 2025 | hide | past | pdf | discuss
|
| 3725. |
Vision language models are blind (2024) [pdf] (arxiv.org) |
|
3 points by thegeomaster on Feb 7, 2025 | hide | past | pdf | discuss
|
| 3726. |
IServe: An Intent-Based Serving System for LLMs (arxiv.org) |
|
1 point by PaulHoule on Feb 7, 2025 | hide | past | pdf | discuss
|
| 3727. |
Global Optimization of Black-Box Functions with Unknown Lipschitz Constants (arxiv.org) |
|
3 points by fofoz on Feb 7, 2025 | hide | past | pdf | discuss
|
| 3728. |
Gold-Medalist Performance in Solving Olympiad Geometry with AlphaGeometry2 (arxiv.org) |
|
64 points by hnhn34 on Feb 7, 2025 | hide | past | pdf | 5 comments
|
| 3729. |
HippoRAG: Neurobiologically Inspired Long-Term Memory for LLMs (2024) (arxiv.org) |
|
65 points by veryluckyxyz on Feb 7, 2025 | hide | past | pdf | 4 comments
|
| 3730. |
Ml.net (arxiv.org) |
|
1 point by colonCapitalDee on Feb 7, 2025 | hide | past | pdf | discuss
|
| 3731. |
Robust autonomy emerges from self-play (arxiv.org) |
|
140 points by reqo on Feb 7, 2025 | hide | past | pdf | 62 comments
|
| 3732. |
Fault Localization via Fine-Tuning LLMs with Mutation Generated Stack Traces (arxiv.org) |
|
3 points by PaulHoule on Feb 7, 2025 | hide | past | pdf | discuss
|
| 3733. |
Meta AI's latest research: improved LLM reasoning with Latent Tokens (arxiv.org) |
|
3 points by lessisgood123 on Feb 7, 2025 | hide | past | pdf | discuss
|
| 3734. |
Understanding Why Adam Outperforms SGD: Gradient Heterogeneity in Transformers (arxiv.org) |
|
3 points by fofoz on Feb 6, 2025 | hide | past | pdf | discuss
|
| 3735. |
Demystifying Long Chain-of-Thought Reasoning in LLMs (arxiv.org) |
|
2 points by sebg on Feb 6, 2025 | hide | past | pdf | discuss
|
| 3736. |
Develop AI Agents for System Engineering in Factorio (arxiv.org) |
|
3 points by Jimmc414 on Feb 6, 2025 | hide | past | pdf | discuss
|
| 3737. |
Optimizing LLM Persuasion with Personalization and Fabricated Statistics (arxiv.org) |
|
2 points by PaulHoule on Feb 6, 2025 | hide | past | pdf | discuss
|
| 3738. |
High-Fidelity Simultaneous Speech-to-Speech Translation (arxiv.org) |
|
6 points by exgrv on Feb 6, 2025 | hide | past | pdf | 1 comment
|
| 3739. |
The Hyperfitting Phenomenon: Sharpening and Stabilizing LLMs (arxiv.org) |
|
3 points by superidiot1932 on Feb 6, 2025 | hide | past | pdf | discuss
|
| 3740. |
LIMO: Less Is More for Reasoning (arxiv.org) |
|
2 points by maksimur on Feb 6, 2025 | hide | past | pdf | discuss
|
| 3741. |
Pre-Trained Large Language Models Use Fourier Features for Addition (2024) (arxiv.org) |
|
149 points by Kye on Feb 6, 2025 | hide | past | pdf | 40 comments
|
| 3742. |
SmolLM2: When Smol Goes Big – Data-Centric Training of a Small Language Model (arxiv.org) |
|
1 point by rahimnathwani on Feb 6, 2025 | hide | past | pdf | discuss
|
| 3743. |
Harmonic Loss Trains Interpretable AI Models (arxiv.org) |
|
1 point by fzliu on Feb 5, 2025 | hide | past | pdf | discuss
|
| 3744. |
Sundial: A Family of Highly Capable Time Series Foundation Models (arxiv.org) |
|
3 points by fofoz on Feb 5, 2025 | hide | past | pdf | discuss
|
| 3745. |
Digital Agent outperforms o1 by 15% – trained with new RL-variant similar to R1 (arxiv.org) |
|
11 points by let_tim_cook_ on Feb 5, 2025 | hide | past | pdf | discuss
|
| 3746. |
Harmonic Loss Trains Interpretable AI Models (arxiv.org) |
|
4 points by lemonfever on Feb 5, 2025 | hide | past | pdf | 2 comments
|
| 3747. |
Over-Tokenized Transformer: Vocabulary Is Generally Worth Scaling (arxiv.org) |
|
2 points by famouswaffles on Feb 4, 2025 | hide | past | pdf | discuss
|
| 3748. |
OmniHuman-1: Scaling-Up of One-Stage Conditioned Human Animation Models (arxiv.org) |
|
1 point by neom on Feb 4, 2025 | hide | past | pdf | discuss
|
| 3749. |
Querying Databases with Function Calling (arxiv.org) |
|
3 points by tosh on Feb 4, 2025 | hide | past | pdf | discuss
|
| 3750. |
DeepRAG: Thinking to retrieval step by step for large language models (arxiv.org) |
|
191 points by fofoz on Feb 4, 2025 | hide | past | pdf | 29 comments
|
| More |