about
2521. FormalGrad: Integrating Formal Methods with Gradient-Based LLM Refinement (arxiv.org)
2 points by PaulHoule on Aug 21, 2025 | hide | past | pdf | discuss
2522. FreshStack: Realistic benchmarks for evaluating retrieval on technical documents (arxiv.org)
4 points by fzliu on Aug 21, 2025 | hide | past | pdf | discuss
2523. R-Zero: Codes for R-Zero: Self-Evolving Reasoning LLM from Zero Data (arxiv.org)
2 points by bigwheels on Aug 21, 2025 | hide | past | pdf | 1 comment
2524. CCFC: Core and Core-Full-Core Dual-Track Defense for LLM Jailbreak Protection (arxiv.org)
1 point by summarity on Aug 21, 2025 | hide | past | pdf | discuss
2525. Beyond sensor data: Foundation models of behavioral data from wearables (arxiv.org)
230 points by brandonb on Aug 21, 2025 | hide | past | pdf | 54 comments
2526. OS-R1: Agentic Operating System Kernel Tuning with Reinforcement Learning (arxiv.org)
1 point by juanviera23 on Aug 21, 2025 | hide | past | pdf | discuss
2527. Scaling laws found in large generative medical event models (arxiv.org)
1 point by iloveoof on Aug 21, 2025 | hide | past | pdf | discuss
2528. Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens (arxiv.org)
1 point by freeqaz on Aug 21, 2025 | hide | past | pdf | discuss
2529. A Systematic Study of Post-Training Quantization for Diffusion LLMs (arxiv.org)
1 point by badmonster on Aug 21, 2025 | hide | past | pdf | discuss
2530. ComputerRL: Scaling Reinforcement Learning for Computer Use Agents (arxiv.org)
1 point by cjbarber on Aug 20, 2025 | hide | past | pdf | discuss
2531. Too Long, Didn't Model (arxiv.org)
2 points by squirrel on Aug 20, 2025 | hide | past | pdf | discuss
2532. Chain-of-Agents (arxiv.org)
2 points by omarsar on Aug 20, 2025 | hide | past | pdf | discuss
2533. Group Sequence Policy Optimization (arxiv.org)
2 points by kdavis on Aug 20, 2025 | hide | past | pdf | 1 comment
2534. A Survey on Diffusion Language Models (arxiv.org)
1 point by Anon84 on Aug 20, 2025 | hide | past | pdf | discuss
2535. AlphaSnake: Policy Iteration on a Nondeterministic NP-Hard MDP (arxiv.org)
1 point by kenny239 on Aug 20, 2025 | hide | past | pdf | discuss
2536. Virtuous Machines: Towards Artificial General Science (arxiv.org)
3 points by frozenseven on Aug 20, 2025 | hide | past | pdf | discuss
2537. Artifacts and Attention Sinks: Structured Approximations for Vision Transformers (arxiv.org)
1 point by PaulHoule on Aug 19, 2025 | hide | past | pdf | discuss
2538. Harnessing Large Language Models to Overcome Recommender System Challenges (arxiv.org)
2 points by PaulHoule on Aug 18, 2025 | hide | past | pdf | discuss
2539. Eyes Will Shut: A Vision-Based Next GPS Location Prediction Model (arxiv.org)
2 points by PaulHoule on Aug 18, 2025 | hide | past | pdf | discuss
2540. TREAD: Token Routing for Efficient Architecture-Agnostic Diffusion Training (arxiv.org)
37 points by fzliu on Aug 18, 2025 | hide | past | pdf | 6 comments
2541. The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text (arxiv.org)
2 points by bckr on Aug 18, 2025 | hide | past | pdf | discuss
2542. Toward Robust Hyper-Detailed Image Captioning (arxiv.org)
3 points by fzliu on Aug 18, 2025 | hide | past | pdf | discuss
2543. Caote: KV Cache Eviction for LLMs (arxiv.org)
3 points by bbzjk7 on Aug 18, 2025 | hide | past | pdf | discuss
2544. Profiling LLM Inference on Apple Silicon: A Quantization Perspective (arxiv.org)
2 points by diggan on Aug 17, 2025 | hide | past | pdf | discuss
2545. ISR: Invertible Symbolic Regression (2024) (arxiv.org)
7 points by liamdgray on Aug 17, 2025 | hide | past | pdf | 1 comment
2546. Composing Linear Layers from Irreducibles (arxiv.org)
2 points by liamdgray on Aug 16, 2025 | hide | past | pdf | 1 comment
2547. PyG 2.0: Scalable Learning on Real World Graphs (arxiv.org)
10 points by PaulHoule on Aug 16, 2025 | hide | past | pdf | 1 comment
2548. IFairy: The First 2-bit Complex LLM with All Parameters in \{\pm1, \pm i\} (arxiv.org)
4 points by Gathering6678 on Aug 16, 2025 | hide | past | pdf | 1 comment
2549. SiLQ: Simple Large Language Model Quantization-Aware Training (arxiv.org)
2 points by PaulHoule on Aug 15, 2025 | hide | past | pdf | discuss
2550. NoLiMa: Long-Context Evaluation Beyond Literal Matching (arxiv.org)
2 points by fzliu on Aug 15, 2025 | hide | past | pdf | discuss