about
2791. Cats Confuse Reasoning LLM: Query Agnostic Adversarial Triggers for Reasoning (arxiv.org)
2 points by felineflock on Jul 5, 2025 | hide | past | pdf | discuss
2792. DiffuCoder: Understanding and Improving Masked Diffusion Models for Code (arxiv.org)
9 points by simonpure on Jul 4, 2025 | hide | past | pdf | discuss
2793. Fast and Simplex: 2-Simplicial Attention in Triton (arxiv.org)
3 points by Pseudomanifold on Jul 4, 2025 | hide | past | pdf | discuss
2794. Few-Shot Learning for Industrial Time Series: Screw-Fastening Process Monitoring (arxiv.org)
1 point by PaulHoule on Jul 4, 2025 | hide | past | pdf | discuss
2795. Can Large Language Models Play Text Games Well? (2023) (arxiv.org)
70 points by willvarfar on Jul 4, 2025 | hide | past | pdf | 54 comments
2796. Establishing Best Practices for Building Rigorous Agentic Benchmarks (arxiv.org)
2 points by consumer451 on Jul 4, 2025 | hide | past | pdf | discuss
2797. Cats Confuse Reasoning LLM: Query Agnostic Adversarial Triggers for Reasoning (arxiv.org)
4 points by ipnon on Jul 4, 2025 | hide | past | pdf | discuss
2798. Attention is all you need (2017) (arxiv.org)
1 point by simonebrunozzi on Jul 4, 2025 | hide | past | pdf | discuss
2799. SegmentAnyMuscle: A muscle segmentation model across different locations in MRI (arxiv.org)
2 points by michaefe on Jul 4, 2025 | hide | past | pdf | discuss
2800. Hierarchical Reasoning Model (arxiv.org)
2 points by mountainview on Jul 4, 2025 | hide | past | pdf | discuss
2801. ML Conferences Should Establish a "Refutations and Critiques" Track (arxiv.org)
3 points by distalx on Jul 4, 2025 | hide | past | pdf | discuss
2802. LoRA Fine-Tuning Without GPUs (arxiv.org)
1 point by elashri on Jul 3, 2025 | hide | past | pdf | discuss
2803. High-fidelity simultaneous speech-to-speech translation (arxiv.org)
115 points by Bluestein on Jul 3, 2025 | hide | past | pdf | 57 comments
2804. AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation (arxiv.org)
1 point by badmonster on Jul 3, 2025 | hide | past | pdf | discuss
2805. Microsoft: Sequential Diagnosis with Language Models (arxiv.org)
2 points by blopker on Jul 3, 2025 | hide | past | pdf | discuss
2806. AI for Scientific Search (arxiv.org)
125 points by omarsar on Jul 3, 2025 | hide | past | pdf | 34 comments
2807. Red Teaming for Gen. AI, Report on a Copyright-Focused Exercise in Academic Med (arxiv.org)
1 point by jjwen on Jul 2, 2025 | hide | past | pdf | discuss
2808. Brain2Model Transfer: Training decision AI using the human brain as a teacher (arxiv.org)
1 point by tomasgaquino on Jul 2, 2025 | hide | past | pdf | 1 comment
2809. MAIR: A Benchmark for Evaluating Instructed Retrieval (2024) (arxiv.org)
1 point by fzliu on Jul 2, 2025 | hide | past | pdf | discuss
2810. Radial Attention: Sparse Attention with Energy Decay for Long Video Generation (arxiv.org)
2 points by yorwba on Jul 2, 2025 | hide | past | pdf | discuss
2811. Transition Matching: Scalable and Flexible Generative Modeling (arxiv.org)
2 points by lnyan on Jul 2, 2025 | hide | past | pdf | discuss
2812. Your Language Model Can Handle Non-Canonical Tokenizations (arxiv.org)
2 points by PaulHoule on Jul 2, 2025 | hide | past | pdf | discuss
2813. MLE-Star: Machine Learning Engineering Agent via Search and Targeted Refinement (arxiv.org)
1 point by PaulHoule on Jul 2, 2025 | hide | past | pdf | discuss
2814. Programs as Singularities (arxiv.org)
2 points by etiams on Jul 2, 2025 | hide | past | pdf | discuss
2815. Large Language Models Don't Make Sense of Word Problems (arxiv.org)
3 points by belter on Jul 2, 2025 | hide | past | pdf | discuss
2816. Huawei releases an open weight model trained on Huawei Ascend GPUs (arxiv.org)
321 points by buyucu on Jul 2, 2025 | hide | past | pdf | 333 comments
2817. Wider or Deeper? Scaling LLM Inference-Time Compute with Adaptive Tree Search (arxiv.org)
3 points by vrm on Jul 2, 2025 | hide | past | pdf | discuss
2818. CoVE: Compressed Vocabulary Expansion Makes Better LLM-Based Recommender Systems (arxiv.org)
2 points by PaulHoule on Jul 1, 2025 | hide | past | pdf | discuss
2819. Pangu Pro Moe: Mixture of Grouped Experts for Efficient Sparsity (arxiv.org)
2 points by diggan on Jul 1, 2025 | hide | past | pdf | discuss
2820. An analytic theory of creativity in convolutional diffusion models (arxiv.org)
2 points by jsenn on Jul 1, 2025 | hide | past | pdf | discuss