| 2791. |
Cats Confuse Reasoning LLM: Query Agnostic Adversarial Triggers for Reasoning (arxiv.org) |
|
2 points by felineflock on Jul 5, 2025 | hide | past | pdf | discuss
|
| 2792. |
DiffuCoder: Understanding and Improving Masked Diffusion Models for Code (arxiv.org) |
|
9 points by simonpure on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2793. |
Fast and Simplex: 2-Simplicial Attention in Triton (arxiv.org) |
|
3 points by Pseudomanifold on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2794. |
Few-Shot Learning for Industrial Time Series: Screw-Fastening Process Monitoring (arxiv.org) |
|
1 point by PaulHoule on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2795. |
Can Large Language Models Play Text Games Well? (2023) (arxiv.org) |
|
70 points by willvarfar on Jul 4, 2025 | hide | past | pdf | 54 comments
|
| 2796. |
Establishing Best Practices for Building Rigorous Agentic Benchmarks (arxiv.org) |
|
2 points by consumer451 on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2797. |
Cats Confuse Reasoning LLM: Query Agnostic Adversarial Triggers for Reasoning (arxiv.org) |
|
4 points by ipnon on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2798. |
Attention is all you need (2017) (arxiv.org) |
|
1 point by simonebrunozzi on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2799. |
SegmentAnyMuscle: A muscle segmentation model across different locations in MRI (arxiv.org) |
|
2 points by michaefe on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2800. |
Hierarchical Reasoning Model (arxiv.org) |
|
2 points by mountainview on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2801. |
ML Conferences Should Establish a "Refutations and Critiques" Track (arxiv.org) |
|
3 points by distalx on Jul 4, 2025 | hide | past | pdf | discuss
|
| 2802. |
LoRA Fine-Tuning Without GPUs (arxiv.org) |
|
1 point by elashri on Jul 3, 2025 | hide | past | pdf | discuss
|
| 2803. |
High-fidelity simultaneous speech-to-speech translation (arxiv.org) |
|
115 points by Bluestein on Jul 3, 2025 | hide | past | pdf | 57 comments
|
| 2804. |
AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation (arxiv.org) |
|
1 point by badmonster on Jul 3, 2025 | hide | past | pdf | discuss
|
| 2805. |
Microsoft: Sequential Diagnosis with Language Models (arxiv.org) |
|
2 points by blopker on Jul 3, 2025 | hide | past | pdf | discuss
|
| 2806. |
AI for Scientific Search (arxiv.org) |
|
125 points by omarsar on Jul 3, 2025 | hide | past | pdf | 34 comments
|
| 2807. |
Red Teaming for Gen. AI, Report on a Copyright-Focused Exercise in Academic Med (arxiv.org) |
|
1 point by jjwen on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2808. |
Brain2Model Transfer: Training decision AI using the human brain as a teacher (arxiv.org) |
|
1 point by tomasgaquino on Jul 2, 2025 | hide | past | pdf | 1 comment
|
| 2809. |
MAIR: A Benchmark for Evaluating Instructed Retrieval (2024) (arxiv.org) |
|
1 point by fzliu on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2810. |
Radial Attention: Sparse Attention with Energy Decay for Long Video Generation (arxiv.org) |
|
2 points by yorwba on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2811. |
Transition Matching: Scalable and Flexible Generative Modeling (arxiv.org) |
|
2 points by lnyan on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2812. |
Your Language Model Can Handle Non-Canonical Tokenizations (arxiv.org) |
|
2 points by PaulHoule on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2813. |
MLE-Star: Machine Learning Engineering Agent via Search and Targeted Refinement (arxiv.org) |
|
1 point by PaulHoule on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2814. |
Programs as Singularities (arxiv.org) |
|
2 points by etiams on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2815. |
Large Language Models Don't Make Sense of Word Problems (arxiv.org) |
|
3 points by belter on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2816. |
Huawei releases an open weight model trained on Huawei Ascend GPUs (arxiv.org) |
|
321 points by buyucu on Jul 2, 2025 | hide | past | pdf | 333 comments
|
| 2817. |
Wider or Deeper? Scaling LLM Inference-Time Compute with Adaptive Tree Search (arxiv.org) |
|
3 points by vrm on Jul 2, 2025 | hide | past | pdf | discuss
|
| 2818. |
CoVE: Compressed Vocabulary Expansion Makes Better LLM-Based Recommender Systems (arxiv.org) |
|
2 points by PaulHoule on Jul 1, 2025 | hide | past | pdf | discuss
|
| 2819. |
Pangu Pro Moe: Mixture of Grouped Experts for Efficient Sparsity (arxiv.org) |
|
2 points by diggan on Jul 1, 2025 | hide | past | pdf | discuss
|
| 2820. |
An analytic theory of creativity in convolutional diffusion models (arxiv.org) |
|
2 points by jsenn on Jul 1, 2025 | hide | past | pdf | discuss
|
| More |