| 3001. |
Esoteric Language Models (arxiv.org) |
|
3 points by frozenseven on Jun 4, 2025 | hide | past | pdf | discuss
|
| 3002. |
The Evaluation of Engineering Artificial General Intelligence (arxiv.org) |
|
3 points by sirregex on Jun 3, 2025 | hide | past | pdf | discuss
|
| 3003. |
How much do language models memorize? (arxiv.org) |
|
58 points by mhmmmmmm on Jun 3, 2025 | hide | past | pdf | 1 comment
|
| 3004. |
TLOB: Dual Attention Transformer Predicts Price Trends from Order Book Data (arxiv.org) |
|
23 points by wertyk on Jun 3, 2025 | hide | past | pdf | discuss
|
| 3005. |
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics (arxiv.org) |
|
4 points by AdilZtn on Jun 3, 2025 | hide | past | pdf | discuss
|
| 3006. |
Transformers as Multi-Task Learners: Decoupling Features in Hidden Markov Models (arxiv.org) |
|
2 points by badmonster on Jun 3, 2025 | hide | past | pdf | 1 comment
|
| 3007. |
A Batch Size and Token NUM- BER Agnostic Learning Rate Scheduler (arxiv.org) |
|
2 points by veryluckyxyz on Jun 3, 2025 | hide | past | pdf | discuss
|
| 3008. |
Attack Breaks Permutation-Based Private Third-Party Inference Schemes for LLMs (arxiv.org) |
|
1 point by transpute on Jun 2, 2025 | hide | past | pdf | discuss
|
| 3009. |
3D CAD from Images, Text, and Point Clouds with RLVR (arxiv.org) |
|
13 points by vokneruk on Jun 2, 2025 | hide | past | pdf | 1 comment
|
| 3010. |
Agents in Charge of Managing Vending Machines: Short vs. Long-Term Coherence (arxiv.org) |
|
3 points by dpflan on Jun 2, 2025 | hide | past | pdf | discuss
|
| 3011. |
TiRex: Zero-Shot Forecasting Across Long and Short Horizons (arxiv.org) |
|
2 points by tosh on Jun 2, 2025 | hide | past | pdf | discuss
|
| 3012. |
WINA: Weight informed Neuron activation for accelerating LLM inference (arxiv.org) |
|
2 points by Ratelman on Jun 2, 2025 | hide | past | pdf | discuss
|
| 3013. |
Beyond the Black Box: Interpretability of LLMs in Finance (arxiv.org) |
|
67 points by ashater on Jun 2, 2025 | hide | past | pdf | 12 comments
|
| 3014. |
Yambda-5B – A Large-Scale Multi-Modal Dataset for Ranking and Retrieval (arxiv.org) |
|
3 points by SerCe on Jun 2, 2025 | hide | past | pdf | discuss
|
| 3015. |
TradeExpert, a trading framework that employs Mixture of Expert LLMs (arxiv.org) |
|
114 points by wertyk on Jun 2, 2025 | hide | past | pdf | 136 comments
|
| 3016. |
ReasoningGym: Reasoning Environments for RL with Verifiable Rewards (arxiv.org) |
|
105 points by t55 on Jun 2, 2025 | hide | past | pdf | 28 comments
|
| 3017. |
Feeding AI personas media diets improves prediction (arxiv.org) |
|
2 points by virtual_rf on Jun 2, 2025 | hide | past | pdf | discuss
|
| 3018. |
AI Persona Opinion Surveying in Healthcare (arxiv.org) |
|
1 point by virtual_rf on Jun 2, 2025 | hide | past | pdf | discuss
|
| 3019. |
Linear Layouts: Robust Code Generation of Efficient Tensor Computation Using F2 (arxiv.org) |
|
2 points by MarcoDewey on Jun 2, 2025 | hide | past | pdf | discuss
|
| 3020. |
LLMs replacing human participants harmfully misportray, flatten identity groups (arxiv.org) |
|
19 points by rntn on Jun 1, 2025 | hide | past | pdf | 9 comments
|
| 3021. |
Beyond Attention: Toward Machines with Intrinsic Higher Mental States (arxiv.org) |
|
67 points by holografix on Jun 1, 2025 | hide | past | pdf | 19 comments
|
| 3022. |
Planner: Generating Diversified Paragraph via Latent Language Diffusion Model (arxiv.org) |
|
1 point by nathan-barry on May 31, 2025 | hide | past | pdf | discuss
|
| 3023. |
SweRank: Software Issue Localization with Code Ranking (arxiv.org) |
|
1 point by PaulHoule on May 31, 2025 | hide | past | pdf | discuss
|
| 3024. |
A Tutorial on Meta-Reinforcement Learning (arxiv.org) |
|
1 point by Anon84 on May 31, 2025 | hide | past | pdf | discuss
|
| 3025. |
SoloSpeech: A high-quality target speech extractor (arxiv.org) |
|
2 points by hwang258 on May 31, 2025 | hide | past | pdf | discuss
|
| 3026. |
YOLO-World: Real-Time Open-Vocabulary Object Detection (arxiv.org) |
|
148 points by greesil on May 31, 2025 | hide | past | pdf | 52 comments
|
| 3027. |
TrojanStego: Your Language Model Can Be a Steganographic Agent (arxiv.org) |
|
1 point by Worta on May 31, 2025 | hide | past | pdf | discuss
|
| 3028. |
Atlas: Learning to Optimally Memorize the Context at Test Time (arxiv.org) |
|
43 points by famouswaffles on May 31, 2025 | hide | past | pdf | 4 comments
|
| 3029. |
Attention Is All You Need (arxiv.org) |
|
2 points by RyanShook on May 31, 2025 | hide | past | pdf | discuss
|
| 3030. |
Model-Preserving Adaptive Rounding (arxiv.org) |
|
5 points by sroy72 on May 30, 2025 | hide | past | pdf | discuss
|
| More |