| 3811. |
Advancing Language Model Reasoning Through RL and Inference Scaling (arxiv.org) |
|
2 points by frozenseven on Jan 25, 2025 | hide | past | pdf | discuss
|
| 3812. |
DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via RL (arxiv.org) |
|
1351 points by gradus_ad on Jan 25, 2025 | hide | past | pdf | 1056 comments
|
| 3813. |
Position Information Emerges in Causal Transformers Without Positional Encoding (arxiv.org) |
|
3 points by PaulHoule on Jan 25, 2025 | hide | past | pdf | discuss
|
| 3814. |
Kimi K1.5 Technical Report (arxiv.org) |
|
4 points by anticensor on Jan 25, 2025 | hide | past | pdf | discuss
|
| 3815. |
Generative Adversarial Neural Network Acceleration with Silicon Photonics (arxiv.org) |
|
2 points by belter on Jan 24, 2025 | hide | past | pdf | discuss
|
| 3816. |
Hallucinations Can Improve Large Language Models in Drug Discovery (arxiv.org) |
|
1 point by keepit on Jan 24, 2025 | hide | past | pdf | discuss
|
| 3817. |
Frontier AI systems have surpassed the self-replicating red line (arxiv.org) |
|
1 point by programd on Jan 24, 2025 | hide | past | pdf | discuss
|
| 3818. |
Dissecting the NVIDIA Hopper Architecture through Microbenchmarking (arxiv.org) |
|
2 points by matt_d on Jan 24, 2025 | hide | past | pdf | discuss
|
| 3819. |
Tell me about yourself: LLMs are aware of their learned behaviors (arxiv.org) |
|
2 points by famouswaffles on Jan 24, 2025 | hide | past | pdf | discuss
|
| 3820. |
Evolution and the Knightian Blindspot of Machine Learning (arxiv.org) |
|
2 points by jal278 on Jan 24, 2025 | hide | past | pdf | discuss
|
| 3821. |
Implicit Chain of Thought Reasoning via Knowledge Distillation (arxiv.org) |
|
1 point by Jimmc414 on Jan 24, 2025 | hide | past | pdf | discuss
|
| 3822. |
UI-Tars: Pioneering Automated GUI Interaction with Native Agents (arxiv.org) |
|
2 points by msoad on Jan 23, 2025 | hide | past | pdf | discuss
|
| 3823. |
Super Tiny Language Models (pdf) (arxiv.org) |
|
3 points by nickpsecurity on Jan 23, 2025 | hide | past | pdf | 1 comment
|
| 3824. |
Ask HN: How Do You Align Evaluator LLMs with Subject Matter Experts? (arxiv.org) |
|
2 points by MutedEstate45 on Jan 23, 2025 | hide | past | pdf | 1 comment
|
| 3825. |
Foundations of Large Language Models (arxiv.org) |
|
219 points by pkoird on Jan 23, 2025 | hide | past | pdf | 20 comments
|
| 3826. |
The Mathematics of Artificial Intelligence (arxiv.org) |
|
4 points by sieste on Jan 23, 2025 | hide | past | pdf | discuss
|
| 3827. |
Lossless Compression of Vector IDs for Approximate Nearest Neighbor Search (arxiv.org) |
|
151 points by fzliu on Jan 22, 2025 | hide | past | pdf | 6 comments
|
| 3828. |
Test-time regression: a unifying framework for designing sequence models (arxiv.org) |
|
1 point by ketothekingdom on Jan 22, 2025 | hide | past | pdf | discuss
|
| 3829. |
Can LLMs demonstrate behavioral self-awareness? (arxiv.org) |
|
3 points by omarsar on Jan 22, 2025 | hide | past | pdf | 1 comment
|
| 3830. |
Evolving Deeper LLM Thinking (arxiv.org) |
|
1 point by simonpure on Jan 22, 2025 | hide | past | pdf | discuss
|
| 3831. |
Physics of Skill Learning (arxiv.org) |
|
1 point by elashri on Jan 22, 2025 | hide | past | pdf | discuss
|
| 3832. |
VideoWorld: Exploring Knowledge Learning from Unlabeled Video (arxiv.org) |
|
2 points by TaurenHunter on Jan 22, 2025 | hide | past | pdf | discuss
|
| 3833. |
Flame: A small language model for spreadsheet formulas (2023) (arxiv.org) |
|
117 points by azhenley on Jan 22, 2025 | hide | past | pdf | 18 comments
|
| 3834. |
Tensor Product Attention Is All You Need (arxiv.org) |
|
160 points by eunos on Jan 22, 2025 | hide | past | pdf | 104 comments
|
| 3835. |
Training DNN in O(1) Memory Using Random Matrices (arxiv.org) |
|
2 points by tnch on Jan 21, 2025 | hide | past | pdf | discuss
|
| 3836. |
Confidence-Based Estimators for Predictive Performance in Model Monitoring (arxiv.org) |
|
1 point by santiviquez on Jan 21, 2025 | hide | past | pdf | discuss
|
| 3837. |
Vision-Language Models Do Not Understand Negation (arxiv.org) |
|
2 points by kjhughes on Jan 21, 2025 | hide | past | pdf | discuss
|
| 3838. |
Accelerating Retrieval-Augmented Generation (arxiv.org) |
|
1 point by handfuloflight on Jan 20, 2025 | hide | past | pdf | discuss
|
| 3839. |
Evolving Deeper LLM Thinking (arxiv.org) |
|
12 points by hardmaru on Jan 20, 2025 | hide | past | pdf | discuss
|
| 3840. |
Enhancing Chat Language Models: Scaling High-Quality Instructional Conversations (arxiv.org) |
|
1 point by TaurenHunter on Jan 19, 2025 | hide | past | pdf | discuss
|
| More |