| 2521. |
FormalGrad: Integrating Formal Methods with Gradient-Based LLM Refinement (arxiv.org) |
|
2 points by PaulHoule on Aug 21, 2025 | hide | past | pdf | discuss
|
| 2522. |
FreshStack: Realistic benchmarks for evaluating retrieval on technical documents (arxiv.org) |
|
4 points by fzliu on Aug 21, 2025 | hide | past | pdf | discuss
|
| 2523. |
R-Zero: Codes for R-Zero: Self-Evolving Reasoning LLM from Zero Data (arxiv.org) |
|
2 points by bigwheels on Aug 21, 2025 | hide | past | pdf | 1 comment
|
| 2524. |
CCFC: Core and Core-Full-Core Dual-Track Defense for LLM Jailbreak Protection (arxiv.org) |
|
1 point by summarity on Aug 21, 2025 | hide | past | pdf | discuss
|
| 2525. |
Beyond sensor data: Foundation models of behavioral data from wearables (arxiv.org) |
|
230 points by brandonb on Aug 21, 2025 | hide | past | pdf | 54 comments
|
| 2526. |
OS-R1: Agentic Operating System Kernel Tuning with Reinforcement Learning (arxiv.org) |
|
1 point by juanviera23 on Aug 21, 2025 | hide | past | pdf | discuss
|
| 2527. |
Scaling laws found in large generative medical event models (arxiv.org) |
|
1 point by iloveoof on Aug 21, 2025 | hide | past | pdf | discuss
|
| 2528. |
Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens (arxiv.org) |
|
1 point by freeqaz on Aug 21, 2025 | hide | past | pdf | discuss
|
| 2529. |
A Systematic Study of Post-Training Quantization for Diffusion LLMs (arxiv.org) |
|
1 point by badmonster on Aug 21, 2025 | hide | past | pdf | discuss
|
| 2530. |
ComputerRL: Scaling Reinforcement Learning for Computer Use Agents (arxiv.org) |
|
1 point by cjbarber on Aug 20, 2025 | hide | past | pdf | discuss
|
| 2531. |
Too Long, Didn't Model (arxiv.org) |
|
2 points by squirrel on Aug 20, 2025 | hide | past | pdf | discuss
|
| 2532. |
Chain-of-Agents (arxiv.org) |
|
2 points by omarsar on Aug 20, 2025 | hide | past | pdf | discuss
|
| 2533. |
Group Sequence Policy Optimization (arxiv.org) |
|
2 points by kdavis on Aug 20, 2025 | hide | past | pdf | 1 comment
|
| 2534. |
A Survey on Diffusion Language Models (arxiv.org) |
|
1 point by Anon84 on Aug 20, 2025 | hide | past | pdf | discuss
|
| 2535. |
AlphaSnake: Policy Iteration on a Nondeterministic NP-Hard MDP (arxiv.org) |
|
1 point by kenny239 on Aug 20, 2025 | hide | past | pdf | discuss
|
| 2536. |
Virtuous Machines: Towards Artificial General Science (arxiv.org) |
|
3 points by frozenseven on Aug 20, 2025 | hide | past | pdf | discuss
|
| 2537. |
Artifacts and Attention Sinks: Structured Approximations for Vision Transformers (arxiv.org) |
|
1 point by PaulHoule on Aug 19, 2025 | hide | past | pdf | discuss
|
| 2538. |
Harnessing Large Language Models to Overcome Recommender System Challenges (arxiv.org) |
|
2 points by PaulHoule on Aug 18, 2025 | hide | past | pdf | discuss
|
| 2539. |
Eyes Will Shut: A Vision-Based Next GPS Location Prediction Model (arxiv.org) |
|
2 points by PaulHoule on Aug 18, 2025 | hide | past | pdf | discuss
|
| 2540. |
TREAD: Token Routing for Efficient Architecture-Agnostic Diffusion Training (arxiv.org) |
|
37 points by fzliu on Aug 18, 2025 | hide | past | pdf | 6 comments
|
| 2541. |
The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text (arxiv.org) |
|
2 points by bckr on Aug 18, 2025 | hide | past | pdf | discuss
|
| 2542. |
Toward Robust Hyper-Detailed Image Captioning (arxiv.org) |
|
3 points by fzliu on Aug 18, 2025 | hide | past | pdf | discuss
|
| 2543. |
Caote: KV Cache Eviction for LLMs (arxiv.org) |
|
3 points by bbzjk7 on Aug 18, 2025 | hide | past | pdf | discuss
|
| 2544. |
Profiling LLM Inference on Apple Silicon: A Quantization Perspective (arxiv.org) |
|
2 points by diggan on Aug 17, 2025 | hide | past | pdf | discuss
|
| 2545. |
ISR: Invertible Symbolic Regression (2024) (arxiv.org) |
|
7 points by liamdgray on Aug 17, 2025 | hide | past | pdf | 1 comment
|
| 2546. |
Composing Linear Layers from Irreducibles (arxiv.org) |
|
2 points by liamdgray on Aug 16, 2025 | hide | past | pdf | 1 comment
|
| 2547. |
PyG 2.0: Scalable Learning on Real World Graphs (arxiv.org) |
|
10 points by PaulHoule on Aug 16, 2025 | hide | past | pdf | 1 comment
|
| 2548. |
IFairy: The First 2-bit Complex LLM with All Parameters in \{\pm1, \pm i\} (arxiv.org) |
|
4 points by Gathering6678 on Aug 16, 2025 | hide | past | pdf | 1 comment
|
| 2549. |
SiLQ: Simple Large Language Model Quantization-Aware Training (arxiv.org) |
|
2 points by PaulHoule on Aug 15, 2025 | hide | past | pdf | discuss
|
| 2550. |
NoLiMa: Long-Context Evaluation Beyond Literal Matching (arxiv.org) |
|
2 points by fzliu on Aug 15, 2025 | hide | past | pdf | discuss
|
| More |