about
3181. Mixture-of-Transformers: Sparse and Scalable Architecture for Multi-Modal Models (arxiv.org)
2 points by mfiguiere on May 10, 2025 | hide | past | pdf | discuss
3182. Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org)
3 points by artninja1988 on May 10, 2025 | hide | past | pdf | discuss
3183. Advancing Conversational Diagnostic AI with Multimodal Reasoning (arxiv.org)
2 points by rntn on May 10, 2025 | hide | past | pdf | discuss
3184. Scaling Laws for Scalable Oversight (arxiv.org)
2 points by sonabinu on May 10, 2025 | hide | past | pdf | discuss
3185. RAGDoll: Efficient Offloading-Based Online RAG System on a Single GPU (arxiv.org)
4 points by PaulHoule on May 10, 2025 | hide | past | pdf | discuss
3186. Learning to reason for long-form story generation (arxiv.org)
1 point by paulpauper on May 9, 2025 | hide | past | pdf | discuss
3187. Imagining and building wise machines: The centrality of AI metacognition (arxiv.org)
1 point by amichail on May 9, 2025 | hide | past | pdf | discuss
3188. Path to Multimodal Generalist: General-Level and General-Bench (arxiv.org)
1 point by frozenseven on May 9, 2025 | hide | past | pdf | discuss
3189. Feeding LLM Annotations to Bert Classifiers at Your Own Risk (arxiv.org)
2 points by PaulHoule on May 9, 2025 | hide | past | pdf | discuss
3190. Understanding Perception and Reasoning Through Model Merging (arxiv.org)
1 point by aibrother on May 9, 2025 | hide | past | pdf | discuss
3191. MotionGlot: A Multi-Embodied Motion Generation Model (arxiv.org)
1 point by programd on May 9, 2025 | hide | past | pdf | discuss
3192. Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org)
3 points by sinuhe69 on May 9, 2025 | hide | past | pdf | 2 comments
3193. ZeroSearch: Incentivize the Search Capability of LLMs Without Searching (arxiv.org)
2 points by ArminRS on May 9, 2025 | hide | past | pdf | discuss
3194. Exploring Compositional Generalization by Transformers (arxiv.org)
1 point by PaulHoule on May 8, 2025 | hide | past | pdf | discuss
3195. The Role of the Gather-and-Aggregate Mechanism in Language Models (arxiv.org)
2 points by PaulHoule on May 8, 2025 | hide | past | pdf | discuss
3196. Advances and Challenges in Foundation Agents (arxiv.org)
3 points by Anon84 on May 8, 2025 | hide | past | pdf | discuss
3197. Machine Learning: A Lecture Note (arxiv.org)
2 points by Anon84 on May 8, 2025 | hide | past | pdf | discuss
3198. A cycle-accurate systolic accelerator simulator for end-to-end system analysis (arxiv.org)
1 point by PaulHoule on May 8, 2025 | hide | past | pdf | discuss
3199. OmniGIRL: A Multilingual and Multimodal Benchmark for GitHub Issue Resolution (arxiv.org)
1 point by badmonster on May 8, 2025 | hide | past | pdf | discuss
3200. All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning (arxiv.org)
1 point by jonbaer on May 8, 2025 | hide | past | pdf | discuss
3201. Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org)
2 points by jonbaer on May 8, 2025 | hide | past | pdf | discuss
3202. Human-Like Episodic Memory for Infinite Context LLMs (arxiv.org)
27 points by aibrother on May 8, 2025 | hide | past | pdf | discuss
3203. LLMs for Materials and Chemistry: 34 Real-World Examples (arxiv.org)
15 points by yz-exodao on May 7, 2025 | hide | past | pdf | 1 comment
3204. Base Models Beat Aligned Models at Randomness and Creativity (arxiv.org)
2 points by nkko on May 7, 2025 | hide | past | pdf | discuss
3205. Absolute Zero: Reinforced Self-Play Reasoning with Zero Data (arxiv.org)
3 points by distalx on May 7, 2025 | hide | past | pdf | discuss
3206. VR-CLI: Learning to Reason for Long-Form Story Generation (arxiv.org)
2 points by andy12_ on May 7, 2025 | hide | past | pdf | discuss
3207. Language Representations Can Be What Recommenders Need: Findings and Potentials (arxiv.org)
2 points by PaulHoule on May 6, 2025 | hide | past | pdf | discuss
3208. Towards Dataset Copyright Evasion Attack Against Personalized Diffusion Models (arxiv.org)
3 points by badmonster on May 6, 2025 | hide | past | pdf | discuss
3209. DoomArena: A Framework for Testing AI Agents Against Evolving Security Threats (arxiv.org)
10 points by PaulHoule on May 6, 2025 | hide | past | pdf | 2 comments
3210. Don't be lazy: CompleteP enables compute-efficient deep transformers (arxiv.org)
4 points by nsdey on May 6, 2025 | hide | past | pdf | discuss