| 2941. |
Text-to-LoRA: Instant Transformer Adaption (arxiv.org) |
|
3 points by yurimo on Jun 12, 2025 | hide | past | pdf | 1 comment
|
| 2942. |
LLMs Can Write Efficient CUDA Kernels (arxiv.org) |
|
2 points by MukundMohanK on Jun 12, 2025 | hide | past | pdf | discuss
|
| 2943. |
Institutional Books: A 242B token dataset from Harvard Library's collections (arxiv.org) |
|
79 points by strangecasts on Jun 11, 2025 | hide | past | pdf | 22 comments
|
| 2944. |
Thermal Detection of People with Mobility Restrictions for Barrier Reduction (arxiv.org) |
|
4 points by PaulHoule on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2945. |
Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents (arxiv.org) |
|
55 points by tzury on Jun 11, 2025 | hide | past | pdf | 16 comments
|
| 2946. |
The Common Pile v0.1: An 8TB dataset of public domain and openly licensed text (arxiv.org) |
|
4 points by namanyayg on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2947. |
TradingAgents: Multi-Agents LLM Financial Trading Framework (arxiv.org) |
|
2 points by _vaporwave_ on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2948. |
Tracr-Injection: Distilling Algorithms into Pre-Trained Language Models (arxiv.org) |
|
1 point by PaulHoule on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2949. |
Mixed-Chip Clusters Enable Efficient Large-Scale AI Training (arxiv.org) |
|
3 points by MukundMohanK on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2950. |
Rethinking Memory in AI: Taxonomy, Operations, Topics, and Future Directions (arxiv.org) |
|
2 points by warthog on Jun 11, 2025 | hide | past | pdf | 1 comment
|
| 2951. |
Small Language Models Are the Future of Agentic AI (arxiv.org) |
|
5 points by nnx on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2952. |
Improving large language models with concept-aware fine-tuning (arxiv.org) |
|
2 points by flyingmalamute on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2953. |
Psychiatric Disorder Diagnosis System Using Wearable ECG Monitors (arxiv.org) |
|
3 points by PaulHoule on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2954. |
Qwen3 Embedding: Advancing Text Embedding and Reranking with Foundation Models (arxiv.org) |
|
1 point by ot on Jun 11, 2025 | hide | past | pdf | discuss
|
| 2955. |
Talk to Your Slides: Efficient Slide Editing Agent with Large Language Models (arxiv.org) |
|
1 point by PaulHoule on Jun 10, 2025 | hide | past | pdf | discuss
|
| 2956. |
Is (Selective) Round-to-Nearest Quantization All You Need? (arxiv.org) |
|
2 points by PaulHoule on Jun 10, 2025 | hide | past | pdf | discuss
|
| 2957. |
JavelinGuard: Low-Cost Transformer Architectures for LLM Security (arxiv.org) |
|
29 points by sharathr on Jun 10, 2025 | hide | past | pdf | 2 comments
|
| 2958. |
Robust agents learn causal world models (arxiv.org) |
|
1 point by felineflock on Jun 10, 2025 | hide | past | pdf | discuss
|
| 2959. |
Geopolitical biases in LLMs (arxiv.org) |
|
2 points by vokneruk on Jun 10, 2025 | hide | past | pdf | discuss
|
| 2960. |
Reinforcement Pre-Training (arxiv.org) |
|
70 points by frozenseven on Jun 10, 2025 | hide | past | pdf | 18 comments
|
| 2961. |
Lossless Compression of LLMl-Generated Text via Next-Token Prediction (arxiv.org) |
|
2 points by PaulHoule on Jun 9, 2025 | hide | past | pdf | discuss
|
| 2962. |
Is Perturbation-Based Image Protection Disruptive to Image Editing? (arxiv.org) |
|
1 point by wslh on Jun 9, 2025 | hide | past | pdf | discuss
|
| 2963. |
Creating General User Models from Computer Use (arxiv.org) |
|
3 points by australium on Jun 9, 2025 | hide | past | pdf | discuss
|
| 2964. |
Why is AI hard and Physics simple? (arxiv.org) |
|
1 point by sebg on Jun 9, 2025 | hide | past | pdf | 1 comment
|
| 2965. |
An Extra RMSNorm Is All You Need for Fine Tuning to 1.58 Bits (arxiv.org) |
|
1 point by PaulHoule on Jun 9, 2025 | hide | past | pdf | discuss
|
| 2966. |
Large Language Models Are Locally Linear Mappings (arxiv.org) |
|
2 points by belter on Jun 9, 2025 | hide | past | pdf | discuss
|
| 2967. |
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from LLMs (arxiv.org) |
|
1 point by tinyspacewizard on Jun 9, 2025 | hide | past | pdf | discuss
|
| 2968. |
Do Large Language Models (Really) Need Statistical Foundations? (arxiv.org) |
|
2 points by ggirelli on Jun 8, 2025 | hide | past | pdf | discuss
|
| 2969. |
LayerPeeler: Autoregressive Peeling for Layer-Wise Image Vectorization (arxiv.org) |
|
5 points by andybak on Jun 8, 2025 | hide | past | pdf | discuss
|
| 2970. |
An SMT Formalization of Mixed-Precision Matrix Multiplication (arxiv.org) |
|
1 point by matt_d on Jun 8, 2025 | hide | past | pdf | discuss
|
| More |