about
3241. Graph Learning Will Lose Relevance Due to Poor Benchmarks (arxiv.org)
2 points by sonabinu on May 2, 2025 | hide | past | pdf | 1 comment
3242. Can Language Models Represent the Past Without Anachronism? (arxiv.org)
3 points by m-hodges on May 2, 2025 | hide | past | pdf | discuss
3243. Improving Instruct Models for Free: A Study on Partial Adaptation (arxiv.org)
1 point by PaulHoule on May 2, 2025 | hide | past | pdf | discuss
3244. ReVision: Video Generation with Explicit 3D Physics Modeling for Complex Motion (arxiv.org)
1 point by badmonster on May 2, 2025 | hide | past | pdf | discuss
3245. The Deep Learning Model of Higher-Lower-Order Cognition, Memory, and Affection (arxiv.org)
1 point by liamdgray on May 1, 2025 | hide | past | pdf | 1 comment
3246. Reinforcement Learning for Reasoning in LLMs with One Training Example (arxiv.org)
2 points by chrsw on May 1, 2025 | hide | past | pdf | discuss
3247. SWE-Smith: Scaling Data for Software Engineering Agents (arxiv.org)
1 point by s-macke on May 1, 2025 | hide | past | pdf | discuss
3248. Pre-Trained Security LLM 8B (arxiv.org)
6 points by YoOnoAP on May 1, 2025 | hide | past | pdf | discuss
3249. Proof or Bluff? Evaluating LLMs on 2025 USA Math Olympiad (arxiv.org)
3 points by godelski on May 1, 2025 | hide | past | pdf | discuss
3250. TesserAct: Learning 4D Embodied World Models (arxiv.org)
1 point by badmonster on May 1, 2025 | hide | past | pdf | discuss
3251. LLMs for Engineering: Teaching Models to Design High Powered Rockets (arxiv.org)
124 points by tamassimond on Apr 30, 2025 | hide | past | pdf | 45 comments
3252. Specialized text classification: classifying Open Banking transactions (arxiv.org)
1 point by PaulHoule on Apr 30, 2025 | hide | past | pdf | discuss
3253. RL for Reasoning in LLMs with One Training Example (arxiv.org)
3 points by simonpure on Apr 30, 2025 | hide | past | pdf | discuss
3254. EDGS: Eliminating Densification for Efficient Convergence of 3DGS (arxiv.org)
2 points by PaulHoule on Apr 30, 2025 | hide | past | pdf | discuss
3255. Let Me Grok for You: Accelerating Grokking via Embedding Transfer (arxiv.org)
2 points by PaulHoule on Apr 30, 2025 | hide | past | pdf | discuss
3256. The Leaderboard Illusion (arxiv.org)
184 points by pongogogo on Apr 30, 2025 | hide | past | pdf | 51 comments
3257. LIFT+: Lightweight Fine-Tuning for Long-Tail Learning (arxiv.org)
2 points by PaulHoule on Apr 30, 2025 | hide | past | pdf | discuss
3258. YoChameleon: Personalized Vision and Language Generation (arxiv.org)
2 points by badmonster on Apr 30, 2025 | hide | past | pdf | discuss
3259. Themisto: Jupyter-Based Runtime Benchmark (arxiv.org)
2 points by PaulHoule on Apr 29, 2025 | hide | past | pdf | discuss
3260. Efficient Memory Management for Large Language Model Serving with PagedAttention (arxiv.org)
2 points by sonabinu on Apr 29, 2025 | hide | past | pdf | discuss
3261. Fast-Slow Thinking for Large Vision-Language Model Reasoning (arxiv.org)
2 points by badmonster on Apr 29, 2025 | hide | past | pdf | discuss
3262. Worldmem: Long-Term Consistent World Simulation with Memory (arxiv.org)
1 point by PaulHoule on Apr 29, 2025 | hide | past | pdf | discuss
3263. CompleteMe: Reference-Based Human Image Completion (arxiv.org)
1 point by badmonster on Apr 29, 2025 | hide | past | pdf | discuss
3264. Vision Transformers Need Registers (arxiv.org)
94 points by felineflock on Apr 28, 2025 | hide | past | pdf | 9 comments
3265. A Coordination Framework of Small LLMs Matches Large LLMs in Data Synthesis (arxiv.org)
2 points by PaulHoule on Apr 28, 2025 | hide | past | pdf | discuss
3266. NodeRAG: Structuring Graph-Based RAG with Heterogeneous Nodes (arxiv.org)
2 points by PaulHoule on Apr 28, 2025 | hide | past | pdf | discuss
3267. When Gaussian Meets Surfel: Ultra-Fast High-Fidelity Radiance Field Rendering (arxiv.org)
2 points by ibobev on Apr 28, 2025 | hide | past | pdf | discuss
3268. LinPrim: Linear Primitives for Differentiable Volumetric Rendering (arxiv.org)
3 points by ibobev on Apr 28, 2025 | hide | past | pdf | discuss
3269. Inference-Aware Fine-Tuning for Best-of-N Sampling in Large Language Models (arxiv.org)
69 points by mfiguiere on Apr 28, 2025 | hide | past | pdf | 9 comments
3270. Sparks of Artificial General Intelligence: Early Experiments with GPT-4 (2023) (arxiv.org)
2 points by catskull on Apr 27, 2025 | hide | past | pdf | discuss