| 3871. |
Decentralized Diffusion Models (arxiv.org) |
|
1 point by dvrp on Jan 15, 2025 | hide | past | pdf | 1 comment
|
| 3872. |
AutoGen Studio: A No-Code Developer Tool for Building Multi-Agent Systems (arxiv.org) |
|
2 points by azhenley on Jan 14, 2025 | hide | past | pdf | discuss
|
| 3873. |
MathReader: Text-to-Speech for Mathematical Documents [pdf] (arxiv.org) |
|
19 points by bikenaga on Jan 14, 2025 | hide | past | pdf | discuss
|
| 3874. |
Generative Flow Networks: Theory and Applications to Structure Learning (arxiv.org) |
|
2 points by lnyan on Jan 14, 2025 | hide | past | pdf | discuss
|
| 3875. |
Neural Network Verification Is a Programming Language Challenge (arxiv.org) |
|
2 points by xtoilette on Jan 14, 2025 | hide | past | pdf | discuss
|
| 3876. |
Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains (arxiv.org) |
|
1 point by fzliu on Jan 14, 2025 | hide | past | pdf | discuss
|
| 3877. |
Frontier Models are Capable of In-context Scheming (arxiv.org) |
|
2 points by wslh on Jan 13, 2025 | hide | past | pdf | discuss
|
| 3878. |
Titans: Learning to Memorize at Test Time (arxiv.org) |
|
115 points by birriel on Jan 13, 2025 | hide | past | pdf | 15 comments
|
| 3879. |
A Framework for Training and Deploying Language Models at the Edge Computers (arxiv.org) |
|
1 point by PaulHoule on Jan 13, 2025 | hide | past | pdf | discuss
|
| 3880. |
Data Poisoning in LLMs: Jailbreak-Tuning and Scaling Laws (arxiv.org) |
|
1 point by pera on Jan 13, 2025 | hide | past | pdf | discuss
|
| 3881. |
Quantifying Positional Biases in Text Embedding Models (arxiv.org) |
|
1 point by PaulHoule on Jan 13, 2025 | hide | past | pdf | discuss
|
| 3882. |
VideoRAG: Retrieval-Augmented Generation over Video Corpus (arxiv.org) |
|
4 points by t55 on Jan 13, 2025 | hide | past | pdf | discuss
|
| 3883. |
State Space Models Are Strong Text Rerankers (arxiv.org) |
|
1 point by PaulHoule on Jan 13, 2025 | hide | past | pdf | discuss
|
| 3884. |
Towards Backdoor Stealthiness in Model Parameter Space (arxiv.org) |
|
1 point by belter on Jan 13, 2025 | hide | past | pdf | discuss
|
| 3885. |
LLM forecasters rapidly approaching human-level performance (arxiv.org) |
|
1 point by drcwpl on Jan 13, 2025 | hide | past | pdf | 1 comment
|
| 3886. |
Question Answering is a Format; When is it Useful? (arxiv.org) |
|
1 point by jbarrow on Jan 13, 2025 | hide | past | pdf | discuss
|
| 3887. |
Inference Scaling vs. Reasoning: Analysis of Compute-Optimal LLM Problem-Solving (arxiv.org) |
|
2 points by PaulHoule on Jan 12, 2025 | hide | past | pdf | discuss
|
| 3888. |
Why Larger Language Models Do In-Context Learning Differently? (arxiv.org) |
|
2 points by belter on Jan 12, 2025 | hide | past | pdf | discuss
|
| 3889. |
Diffusion Models Generalize via Geometry-Adaptive Harmonic Representations (arxiv.org) |
|
3 points by belter on Jan 12, 2025 | hide | past | pdf | discuss
|
| 3890. |
Progress in Deep Learning: SGD Learns Parities Near the Computational Limit (arxiv.org) |
|
2 points by nabla9 on Jan 10, 2025 | hide | past | pdf | discuss
|
| 3891. |
The GAN is dead; long live the GAN - A Modern GAN Baseline (arxiv.org) |
|
3 points by IdealeZahlen on Jan 10, 2025 | hide | past | pdf | 1 comment
|
| 3892. |
Search-O1: Agentic Search-Enhanced Large Reasoning Models (arxiv.org) |
|
2 points by omarsar on Jan 10, 2025 | hide | past | pdf | discuss
|
| 3893. |
Learning how to think with Meta Chain-of-Thought (arxiv.org) |
|
229 points by drcwpl on Jan 10, 2025 | hide | past | pdf | 75 comments
|
| 3894. |
RAG with Differential Privacy (arxiv.org) |
|
2 points by ngrislain on Jan 10, 2025 | hide | past | pdf | discuss
|
| 3895. |
A Survey on LLMs with Some Insights on Their Capabilities and Limitations (arxiv.org) |
|
1 point by TaurenHunter on Jan 10, 2025 | hide | past | pdf | discuss
|
| 3896. |
Neural Parameter Estimation with Incomplete Data (arxiv.org) |
|
1 point by elashri on Jan 9, 2025 | hide | past | pdf | discuss
|
| 3897. |
Agent Laboratory: Using LLM Agents as Research Assistants (arxiv.org) |
|
1 point by omarsar on Jan 9, 2025 | hide | past | pdf | discuss
|
| 3898. |
Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking (arxiv.org) |
|
1 point by omarsar on Jan 9, 2025 | hide | past | pdf | discuss
|
| 3899. |
Searching Latent Program Spaces (arxiv.org) |
|
2 points by sva_ on Jan 9, 2025 | hide | past | pdf | discuss
|
| 3900. |
Towards System 2 Reasoning in LLMs: Learning How to Think Meta Chain-of-Thought (arxiv.org) |
|
1 point by Jimmc414 on Jan 9, 2025 | hide | past | pdf | discuss
|
| More |