| 3181. |
Mixture-of-Transformers: Sparse and Scalable Architecture for Multi-Modal Models (arxiv.org) |
|
2 points by mfiguiere on May 10, 2025 | hide | past | pdf | discuss
|
| 3182. |
Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org) |
|
3 points by artninja1988 on May 10, 2025 | hide | past | pdf | discuss
|
| 3183. |
Advancing Conversational Diagnostic AI with Multimodal Reasoning (arxiv.org) |
|
2 points by rntn on May 10, 2025 | hide | past | pdf | discuss
|
| 3184. |
Scaling Laws for Scalable Oversight (arxiv.org) |
|
2 points by sonabinu on May 10, 2025 | hide | past | pdf | discuss
|
| 3185. |
RAGDoll: Efficient Offloading-Based Online RAG System on a Single GPU (arxiv.org) |
|
4 points by PaulHoule on May 10, 2025 | hide | past | pdf | discuss
|
| 3186. |
Learning to reason for long-form story generation (arxiv.org) |
|
1 point by paulpauper on May 9, 2025 | hide | past | pdf | discuss
|
| 3187. |
Imagining and building wise machines: The centrality of AI metacognition (arxiv.org) |
|
1 point by amichail on May 9, 2025 | hide | past | pdf | discuss
|
| 3188. |
Path to Multimodal Generalist: General-Level and General-Bench (arxiv.org) |
|
1 point by frozenseven on May 9, 2025 | hide | past | pdf | discuss
|
| 3189. |
Feeding LLM Annotations to Bert Classifiers at Your Own Risk (arxiv.org) |
|
2 points by PaulHoule on May 9, 2025 | hide | past | pdf | discuss
|
| 3190. |
Understanding Perception and Reasoning Through Model Merging (arxiv.org) |
|
1 point by aibrother on May 9, 2025 | hide | past | pdf | discuss
|
| 3191. |
MotionGlot: A Multi-Embodied Motion Generation Model (arxiv.org) |
|
1 point by programd on May 9, 2025 | hide | past | pdf | discuss
|
| 3192. |
Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org) |
|
3 points by sinuhe69 on May 9, 2025 | hide | past | pdf | 2 comments
|
| 3193. |
ZeroSearch: Incentivize the Search Capability of LLMs Without Searching (arxiv.org) |
|
2 points by ArminRS on May 9, 2025 | hide | past | pdf | discuss
|
| 3194. |
Exploring Compositional Generalization by Transformers (arxiv.org) |
|
1 point by PaulHoule on May 8, 2025 | hide | past | pdf | discuss
|
| 3195. |
The Role of the Gather-and-Aggregate Mechanism in Language Models (arxiv.org) |
|
2 points by PaulHoule on May 8, 2025 | hide | past | pdf | discuss
|
| 3196. |
Advances and Challenges in Foundation Agents (arxiv.org) |
|
3 points by Anon84 on May 8, 2025 | hide | past | pdf | discuss
|
| 3197. |
Machine Learning: A Lecture Note (arxiv.org) |
|
2 points by Anon84 on May 8, 2025 | hide | past | pdf | discuss
|
| 3198. |
A cycle-accurate systolic accelerator simulator for end-to-end system analysis (arxiv.org) |
|
1 point by PaulHoule on May 8, 2025 | hide | past | pdf | discuss
|
| 3199. |
OmniGIRL: A Multilingual and Multimodal Benchmark for GitHub Issue Resolution (arxiv.org) |
|
1 point by badmonster on May 8, 2025 | hide | past | pdf | discuss
|
| 3200. |
All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning (arxiv.org) |
|
1 point by jonbaer on May 8, 2025 | hide | past | pdf | discuss
|
| 3201. |
Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org) |
|
2 points by jonbaer on May 8, 2025 | hide | past | pdf | discuss
|
| 3202. |
Human-Like Episodic Memory for Infinite Context LLMs (arxiv.org) |
|
27 points by aibrother on May 8, 2025 | hide | past | pdf | discuss
|
| 3203. |
LLMs for Materials and Chemistry: 34 Real-World Examples (arxiv.org) |
|
15 points by yz-exodao on May 7, 2025 | hide | past | pdf | 1 comment
|
| 3204. |
Base Models Beat Aligned Models at Randomness and Creativity (arxiv.org) |
|
2 points by nkko on May 7, 2025 | hide | past | pdf | discuss
|
| 3205. |
Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org) |
|
3 points by distalx on May 7, 2025 | hide | past | pdf | discuss
|
| 3206. |
VR-CLI: Learning to Reason for Long-Form Story Generation (arxiv.org) |
|
2 points by andy12_ on May 7, 2025 | hide | past | pdf | discuss
|
| 3207. |
Language Representations Can Be What Recommenders Need: Findings and Potentials (arxiv.org) |
|
2 points by PaulHoule on May 6, 2025 | hide | past | pdf | discuss
|
| 3208. |
Towards Dataset Copyright Evasion Attack Against Personalized Diffusion Models (arxiv.org) |
|
3 points by badmonster on May 6, 2025 | hide | past | pdf | discuss
|
| 3209. |
DoomArena: A Framework for Testing AI Agents Against Evolving Security Threats (arxiv.org) |
|
10 points by PaulHoule on May 6, 2025 | hide | past | pdf | 2 comments
|
| 3210. |
Don't be lazy: CompleteP enables compute-efficient deep transformers (arxiv.org) |
|
4 points by nsdey on May 6, 2025 | hide | past | pdf | discuss
|
| More |