about
2491. Speed Always Wins: A Survey on Efficient Architectures for LLMs (arxiv.org)
2 points by Anon84 on Aug 27, 2025 | hide | past | pdf | discuss
2492. Audio-Visual Contact Classification for Tree Structures in Agriculture (arxiv.org)
3 points by PaulHoule on Aug 27, 2025 | hide | past | pdf | discuss
2493. 2-D Sparse Parallelism for Deep Learning Recommendation Model Training (arxiv.org)
1 point by PaulHoule on Aug 26, 2025 | hide | past | pdf | discuss
2494. Building Realistic Benchmarks for Evaluating Retrieval on Technical Documents (arxiv.org)
2 points by fzliu on Aug 26, 2025 | hide | past | pdf | discuss
2495. LLM Speed Up Breakthrough? (arxiv.org)
2 points by bilsbie on Aug 26, 2025 | hide | past | pdf | discuss
2496. Logit-Gap Steering: Efficient Short-Suffix Jailbreaks for Aligned LLMs (arxiv.org)
1 point by rntn on Aug 26, 2025 | hide | past | pdf | discuss
2497. Sequence Parallelism: Long Sequence Training from System Perspective (2021) (arxiv.org)
2 points by jxmorris12 on Aug 26, 2025 | hide | past | pdf | discuss
2498. Memento: Fine-tuning LLM Agents without Fine-tuning LLMs (arxiv.org)
1 point by vincirufus on Aug 26, 2025 | hide | past | pdf | 1 comment
2499. AnalogSeeker: An Open-Source Foundation Language Model for Analog Circuit Design (arxiv.org)
3 points by by_Seeing on Aug 25, 2025 | hide | past | pdf | discuss
2500. AlphaAgents: LLM Based Multi-Agents for Equity Portfolio Constructions (arxiv.org)
1 point by rbanffy on Aug 25, 2025 | hide | past | pdf | discuss
2501. AetherCode: Evaluating LLMs' Ability to Win in Premier Programming Competitions (arxiv.org)
2 points by pella on Aug 25, 2025 | hide | past | pdf | 1 comment
2502. Alvorada-Bench: Can Language Models Solve Brazilian University Entrance Exams? (arxiv.org)
1 point by henriquegodoy on Aug 25, 2025 | hide | past | pdf | discuss
2503. EvoGit: Decentralized multi-agent framework for software development (arxiv.org)
3 points by manualwise on Aug 25, 2025 | hide | past | pdf | 1 comment
2504. Hardwired-Neurons LPUs as General-Purpose Cognitive Substrates (arxiv.org)
1 point by chrsw on Aug 25, 2025 | hide | past | pdf | discuss
2505. Reinforcement learning and VR for personalized arachnophobia treatment (arxiv.org)
2 points by Kye on Aug 25, 2025 | hide | past | pdf | 1 comment
2506. Exploring LLM Confidence in Code Completion (arxiv.org)
2 points by Hard_Space on Aug 25, 2025 | hide | past | pdf | discuss
2507. AetherCode: Evaluating LLMs' Ability to Win in Premier Programming Competitions (arxiv.org)
1 point by limoce on Aug 25, 2025 | hide | past | pdf | discuss
2508. Zero-Shot Retrieval for Scalable Visual Search in a Two-Sided Marketplace (arxiv.org)
2 points by PaulHoule on Aug 24, 2025 | hide | past | pdf | discuss
2509. Evaluating Long-Term Conversational Memory of LLM Agents (arxiv.org)
1 point by handfuloflight on Aug 24, 2025 | hide | past | pdf | discuss
2510. BeyondWeb: Lessons from Scaling Synthetic Data for Trillion-Scale Pretraining (arxiv.org)
4 points by circuithunter on Aug 23, 2025 | hide | past | pdf | discuss
2511. Algorithmic Fairness Amid Social Determinants (arxiv.org)
2 points by PaulHoule on Aug 23, 2025 | hide | past | pdf | discuss
2512. LNS-Madam: Low-Precision Training in Log Using Multiplicative Weight Update (arxiv.org)
2 points by nabla9 on Aug 23, 2025 | hide | past | pdf | 1 comment
2513. Modeling Annotator Disagreement with Demographic-Aware Experts (arxiv.org)
2 points by PaulHoule on Aug 23, 2025 | hide | past | pdf | discuss
2514. Making LLMs Cheaper and Better via Performance-Efficiency Optimized Routing (arxiv.org)
130 points by omarsar on Aug 22, 2025 | hide | past | pdf | 28 comments
2515. How Random Is Random? Evaluating Randomness and Humaness of LLM Coin Flip (2024) (arxiv.org)
2 points by walterbell on Aug 22, 2025 | hide | past | pdf | discuss
2516. Part I: Tricks or Traps? A Deep Dive into RL for LLM Reasoning (arxiv.org)
1 point by Anon84 on Aug 22, 2025 | hide | past | pdf | discuss
2517. Is GPT-OSS Good? A Comprehensive Evaluation (arxiv.org)
2 points by dloss on Aug 22, 2025 | hide | past | pdf | discuss
2518. Intern-S1: A Scientific Multimodal Foundation Model (arxiv.org)
4 points by anothermathbozo on Aug 22, 2025 | hide | past | pdf | 1 comment
2519. DuPO: Enabling Reliable LLM Self-Verification via Dual Preference Optimization (arxiv.org)
1 point by simonpure on Aug 22, 2025 | hide | past | pdf | discuss
2520. FormalGrad: Integrating Formal Methods with Gradient-Based LLM Refinement (arxiv.org)
2 points by PaulHoule on Aug 21, 2025 | hide | past | pdf | discuss