| 2971. |
Reinforcement Learning to Train Large Language Models to Explain Human Decisions (arxiv.org) |
|
25 points by PaulHoule on Jun 7, 2025 | hide | past | pdf | discuss
|
| 2972. |
Prompting Techniques for Secure Code Generation (arxiv.org) |
|
6 points by saikatsg on Jun 7, 2025 | hide | past | pdf | discuss
|
| 2973. |
Why is AI hard and Physics simple? (arxiv.org) |
|
1 point by jxmorris12 on Jun 7, 2025 | hide | past | pdf | discuss
|
| 2974. |
Analog Foundation Models (arxiv.org) |
|
8 points by PaulHoule on Jun 7, 2025 | hide | past | pdf | 1 comment
|
| 2975. |
Log-Linear Attention (arxiv.org) |
|
41 points by sva_ on Jun 7, 2025 | hide | past | pdf | 3 comments
|
| 2976. |
OpenThoughts: Data Recipes for Reasoning Models (arxiv.org) |
|
3 points by Anon84 on Jun 7, 2025 | hide | past | pdf | discuss
|
| 2977. |
Lossless data compression by large models (arxiv.org) |
|
2 points by vitplister on Jun 7, 2025 | hide | past | pdf | discuss
|
| 2978. |
LLM-Explorer: Efficient and Affordable LLM-Based Exploration for Mobile Apps (arxiv.org) |
|
3 points by PaulHoule on Jun 6, 2025 | hide | past | pdf | discuss
|
| 2979. |
Holo1: Cost-Efficient Web Agent Powered by Open Weights (arxiv.org) |
|
3 points by marc-thibault on Jun 6, 2025 | hide | past | pdf | discuss
|
| 2980. |
Efficient Streaming Language Models with Attention Sinks (arxiv.org) |
|
5 points by jxmorris12 on Jun 6, 2025 | hide | past | pdf | discuss
|
| 2981. |
The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text (arxiv.org) |
|
4 points by yorwba on Jun 6, 2025 | hide | past | pdf | discuss
|
| 2982. |
A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples (arxiv.org) |
|
3 points by PaulHoule on Jun 6, 2025 | hide | past | pdf | discuss
|
| 2983. |
LongCodeBench: Evaluating Coding LLMs at 1M Context Windows (arxiv.org) |
|
25 points by PaulHoule on Jun 6, 2025 | hide | past | pdf | discuss
|
| 2984. |
Algebra Unveils Deep Learning – An Invitation to Neuroalgebraic Geometry (arxiv.org) |
|
13 points by IdealeZahlen on Jun 6, 2025 | hide | past | pdf | discuss
|
| 2985. |
Contrastive Flow Matching (arxiv.org) |
|
2 points by badmonster on Jun 6, 2025 | hide | past | pdf | 1 comment
|
| 2986. |
Quantum Mixed-State Self-Attention Network (arxiv.org) |
|
1 point by fs_tab on Jun 6, 2025 | hide | past | pdf | discuss
|
| 2987. |
What LLMss Don't Talk About: Empirical Study of Moderation & Censorship Practice (arxiv.org) |
|
4 points by superpupervlad on Jun 5, 2025 | hide | past | pdf | discuss
|
| 2988. |
RL in Name Only? Analyzing the Structural Assumptions in RL Post-Training (arxiv.org) |
|
2 points by porridgeraisin on Jun 5, 2025 | hide | past | pdf | discuss
|
| 2989. |
I-Con: A Unifying Framework for Representation Learning (arxiv.org) |
|
1 point by aerophilic on Jun 5, 2025 | hide | past | pdf | 1 comment
|
| 2990. |
Extreme Super-Resolution via Scale Autoregression and Preference Alignment (arxiv.org) |
|
14 points by Brajeshwar on Jun 5, 2025 | hide | past | pdf | 2 comments
|
| 2991. |
Questioning Representational Optimism in Deep Learning (arxiv.org) |
|
1 point by publicdaniel on Jun 5, 2025 | hide | past | pdf | 3 comments
|
| 2992. |
From tokens to thoughts: How LLMs and humans trade compression for meaning (arxiv.org) |
|
124 points by ggirelli on Jun 5, 2025 | hide | past | pdf | 25 comments
|
| 2993. |
Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models (arxiv.org) |
|
1 point by PaulHoule on Jun 5, 2025 | hide | past | pdf | discuss
|
| 2994. |
Predicting Empirical AI Research Outcomes with Language Models (arxiv.org) |
|
2 points by danielmorozoff on Jun 5, 2025 | hide | past | pdf | discuss
|
| 2995. |
Not all tokens are meant to be forgotten (arxiv.org) |
|
54 points by MarcoDewey on Jun 4, 2025 | hide | past | pdf | 23 comments
|
| 2996. |
Generating High-Performance Tensor Operators with Hardware Primitives (arxiv.org) |
|
2 points by PaulHoule on Jun 4, 2025 | hide | past | pdf | discuss
|
| 2997. |
Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents [pdf] (arxiv.org) |
|
1 point by _tk_ on Jun 4, 2025 | hide | past | pdf | discuss
|
| 2998. |
Leancode: Understanding Models Better for Code Simplification of Pre-Trained LLM (arxiv.org) |
|
1 point by PaulHoule on Jun 4, 2025 | hide | past | pdf | discuss
|
| 2999. |
Auto-Labeling Data for Object Detection (arxiv.org) |
|
6 points by nwlotz on Jun 4, 2025 | hide | past | pdf | 1 comment
|
| 3000. |
Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (arxiv.org) |
|
1 point by felineflock on Jun 4, 2025 | hide | past | pdf | discuss
|
| More |