about
2941. Text-to-LoRA: Instant Transformer Adaption (arxiv.org)
3 points by yurimo on Jun 12, 2025 | hide | past | pdf | 1 comment
2942. LLMs Can Write Efficient CUDA Kernels (arxiv.org)
2 points by MukundMohanK on Jun 12, 2025 | hide | past | pdf | discuss
2943. Institutional Books: A 242B token dataset from Harvard Library's collections (arxiv.org)
79 points by strangecasts on Jun 11, 2025 | hide | past | pdf | 22 comments
2944. Thermal Detection of People with Mobility Restrictions for Barrier Reduction (arxiv.org)
4 points by PaulHoule on Jun 11, 2025 | hide | past | pdf | discuss
2945. Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents (arxiv.org)
55 points by tzury on Jun 11, 2025 | hide | past | pdf | 16 comments
2946. The Common Pile v0.1: An 8TB dataset of public domain and openly licensed text (arxiv.org)
4 points by namanyayg on Jun 11, 2025 | hide | past | pdf | discuss
2947. TradingAgents: Multi-Agents LLM Financial Trading Framework (arxiv.org)
2 points by _vaporwave_ on Jun 11, 2025 | hide | past | pdf | discuss
2948. Tracr-Injection: Distilling Algorithms into Pre-Trained Language Models (arxiv.org)
1 point by PaulHoule on Jun 11, 2025 | hide | past | pdf | discuss
2949. Mixed-Chip Clusters Enable Efficient Large-Scale AI Training (arxiv.org)
3 points by MukundMohanK on Jun 11, 2025 | hide | past | pdf | discuss
2950. Rethinking Memory in AI: Taxonomy, Operations, Topics, and Future Directions (arxiv.org)
2 points by warthog on Jun 11, 2025 | hide | past | pdf | 1 comment
2951. Small Language Models Are the Future of Agentic AI (arxiv.org)
5 points by nnx on Jun 11, 2025 | hide | past | pdf | discuss
2952. Improving large language models with concept-aware fine-tuning (arxiv.org)
2 points by flyingmalamute on Jun 11, 2025 | hide | past | pdf | discuss
2953. Psychiatric Disorder Diagnosis System Using Wearable ECG Monitors (arxiv.org)
3 points by PaulHoule on Jun 11, 2025 | hide | past | pdf | discuss
2954. Qwen3 Embedding: Advancing Text Embedding and Reranking with Foundation Models (arxiv.org)
1 point by ot on Jun 11, 2025 | hide | past | pdf | discuss
2955. Talk to Your Slides: Efficient Slide Editing Agent with Large Language Models (arxiv.org)
1 point by PaulHoule on Jun 10, 2025 | hide | past | pdf | discuss
2956. Is (Selective) Round-to-Nearest Quantization All You Need? (arxiv.org)
2 points by PaulHoule on Jun 10, 2025 | hide | past | pdf | discuss
2957. JavelinGuard: Low-Cost Transformer Architectures for LLM Security (arxiv.org)
29 points by sharathr on Jun 10, 2025 | hide | past | pdf | 2 comments
2958. Robust agents learn causal world models (arxiv.org)
1 point by felineflock on Jun 10, 2025 | hide | past | pdf | discuss
2959. Geopolitical biases in LLMs (arxiv.org)
2 points by vokneruk on Jun 10, 2025 | hide | past | pdf | discuss
2960. Reinforcement Pre-Training (arxiv.org)
70 points by frozenseven on Jun 10, 2025 | hide | past | pdf | 18 comments
2961. Lossless Compression of LLMl-Generated Text via Next-Token Prediction (arxiv.org)
2 points by PaulHoule on Jun 9, 2025 | hide | past | pdf | discuss
2962. Is Perturbation-Based Image Protection Disruptive to Image Editing? (arxiv.org)
1 point by wslh on Jun 9, 2025 | hide | past | pdf | discuss
2963. Creating General User Models from Computer Use (arxiv.org)
3 points by australium on Jun 9, 2025 | hide | past | pdf | discuss
2964. Why is AI hard and Physics simple? (arxiv.org)
1 point by sebg on Jun 9, 2025 | hide | past | pdf | 1 comment
2965. An Extra RMSNorm Is All You Need for Fine Tuning to 1.58 Bits (arxiv.org)
1 point by PaulHoule on Jun 9, 2025 | hide | past | pdf | discuss
2966. Large Language Models Are Locally Linear Mappings (arxiv.org)
2 points by belter on Jun 9, 2025 | hide | past | pdf | discuss
2967. RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from LLMs (arxiv.org)
1 point by tinyspacewizard on Jun 9, 2025 | hide | past | pdf | discuss
2968. Do Large Language Models (Really) Need Statistical Foundations? (arxiv.org)
2 points by ggirelli on Jun 8, 2025 | hide | past | pdf | discuss
2969. LayerPeeler: Autoregressive Peeling for Layer-Wise Image Vectorization (arxiv.org)
5 points by andybak on Jun 8, 2025 | hide | past | pdf | discuss
2970. An SMT Formalization of Mixed-Precision Matrix Multiplication (arxiv.org)
1 point by matt_d on Jun 8, 2025 | hide | past | pdf | discuss