about
2971. Reinforcement Learning to Train Large Language Models to Explain Human Decisions (arxiv.org)
25 points by PaulHoule on Jun 7, 2025 | hide | past | pdf | discuss
2972. Prompting Techniques for Secure Code Generation (arxiv.org)
6 points by saikatsg on Jun 7, 2025 | hide | past | pdf | discuss
2973. Why is AI hard and Physics simple? (arxiv.org)
1 point by jxmorris12 on Jun 7, 2025 | hide | past | pdf | discuss
2974. Analog Foundation Models (arxiv.org)
8 points by PaulHoule on Jun 7, 2025 | hide | past | pdf | 1 comment
2975. Log-Linear Attention (arxiv.org)
41 points by sva_ on Jun 7, 2025 | hide | past | pdf | 3 comments
2976. OpenThoughts: Data Recipes for Reasoning Models (arxiv.org)
3 points by Anon84 on Jun 7, 2025 | hide | past | pdf | discuss
2977. Lossless data compression by large models (arxiv.org)
2 points by vitplister on Jun 7, 2025 | hide | past | pdf | discuss
2978. LLM-Explorer: Efficient and Affordable LLM-Based Exploration for Mobile Apps (arxiv.org)
3 points by PaulHoule on Jun 6, 2025 | hide | past | pdf | discuss
2979. Holo1: Cost-Efficient Web Agent Powered by Open Weights (arxiv.org)
3 points by marc-thibault on Jun 6, 2025 | hide | past | pdf | discuss
2980. Efficient Streaming Language Models with Attention Sinks (arxiv.org)
5 points by jxmorris12 on Jun 6, 2025 | hide | past | pdf | discuss
2981. The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text (arxiv.org)
4 points by yorwba on Jun 6, 2025 | hide | past | pdf | discuss
2982. A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples (arxiv.org)
3 points by PaulHoule on Jun 6, 2025 | hide | past | pdf | discuss
2983. LongCodeBench: Evaluating Coding LLMs at 1M Context Windows (arxiv.org)
25 points by PaulHoule on Jun 6, 2025 | hide | past | pdf | discuss
2984. Algebra Unveils Deep Learning – An Invitation to Neuroalgebraic Geometry (arxiv.org)
13 points by IdealeZahlen on Jun 6, 2025 | hide | past | pdf | discuss
2985. Contrastive Flow Matching (arxiv.org)
2 points by badmonster on Jun 6, 2025 | hide | past | pdf | 1 comment
2986. Quantum Mixed-State Self-Attention Network (arxiv.org)
1 point by fs_tab on Jun 6, 2025 | hide | past | pdf | discuss
2987. What LLMss Don't Talk About: Empirical Study of Moderation & Censorship Practice (arxiv.org)
4 points by superpupervlad on Jun 5, 2025 | hide | past | pdf | discuss
2988. RL in Name Only? Analyzing the Structural Assumptions in RL Post-Training (arxiv.org)
2 points by porridgeraisin on Jun 5, 2025 | hide | past | pdf | discuss
2989. I-Con: A Unifying Framework for Representation Learning (arxiv.org)
1 point by aerophilic on Jun 5, 2025 | hide | past | pdf | 1 comment
2990. Extreme Super-Resolution via Scale Autoregression and Preference Alignment (arxiv.org)
14 points by Brajeshwar on Jun 5, 2025 | hide | past | pdf | 2 comments
2991. Questioning Representational Optimism in Deep Learning (arxiv.org)
1 point by publicdaniel on Jun 5, 2025 | hide | past | pdf | 3 comments
2992. From tokens to thoughts: How LLMs and humans trade compression for meaning (arxiv.org)
124 points by ggirelli on Jun 5, 2025 | hide | past | pdf | 25 comments
2993. Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models (arxiv.org)
1 point by PaulHoule on Jun 5, 2025 | hide | past | pdf | discuss
2994. Predicting Empirical AI Research Outcomes with Language Models (arxiv.org)
2 points by danielmorozoff on Jun 5, 2025 | hide | past | pdf | discuss
2995. Not all tokens are meant to be forgotten (arxiv.org)
54 points by MarcoDewey on Jun 4, 2025 | hide | past | pdf | 23 comments
2996. Generating High-Performance Tensor Operators with Hardware Primitives (arxiv.org)
2 points by PaulHoule on Jun 4, 2025 | hide | past | pdf | discuss
2997. Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents [pdf] (arxiv.org)
1 point by _tk_ on Jun 4, 2025 | hide | past | pdf | discuss
2998. Leancode: Understanding Models Better for Code Simplification of Pre-Trained LLM (arxiv.org)
1 point by PaulHoule on Jun 4, 2025 | hide | past | pdf | discuss
2999. Auto-Labeling Data for Object Detection (arxiv.org)
6 points by nwlotz on Jun 4, 2025 | hide | past | pdf | 1 comment
3000. Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces (arxiv.org)
1 point by felineflock on Jun 4, 2025 | hide | past | pdf | discuss