| 1. |
Differential Transformer (arxiv.org) |
|
562 points by weirdcat on Oct 8, 2024 | hide | past | pdf | 177 comments
|
| 2. |
LLMs as Markov Chains (arxiv.org) |
|
5 points by akrymski on Oct 8, 2024 | hide | past | pdf | discuss
|
| 3. |
How Chinese Are Chinese Language Models? Lack of Language Policy in China's LLMs (arxiv.org) |
|
4 points by rntn on Oct 8, 2024 | hide | past | pdf | discuss
|
| 4. |
TableRAG: Million-Token Table Understanding with Language Models (arxiv.org) |
|
2 points by fzliu on Oct 8, 2024 | hide | past | pdf | discuss
|
| 5. |
Eliciting Better Multilingual Structured Reasoning from LLMs Through Code (arxiv.org) |
|
2 points by rntn on Oct 8, 2024 | hide | past | pdf | discuss
|
| 6. |
You Only Use Reactive Attention Slice for Long Context Retrieval (arxiv.org) |
|
2 points by PaulHoule on Oct 8, 2024 | hide | past | pdf | discuss
|
| 7. |
MobileLLM: Optimizing Subbillion Parameter Language Models for OnDevice UseCases (arxiv.org) |
|
2 points by rkwz on Oct 8, 2024 | hide | past | pdf | discuss
|
| 8. |
Large Language Models as Markov Chains (arxiv.org) |
|
1 point by tosh on Oct 8, 2024 | hide | past | pdf | discuss
|
| 9. |
Evil Geniuses: Delving into the Safety of LLM-Based Agents [pdf] (arxiv.org) |
|
1 point by squircle on Oct 8, 2024 | hide | past | pdf | discuss
|
| 10. |
I Bet You Did Not Mean That: Testing Semantic Importance via Betting (arxiv.org) |
|
1 point by Phantom_Core on Oct 8, 2024 | hide | past | pdf | discuss
|
| 11. |
Applying Quantum Autoencoders for Time Series Anomaly Detection (arxiv.org) |
|
1 point by hdvr on Oct 8, 2024 | hide | past | pdf | discuss
|
| 12. |
TableRAG: Million-Token Table Understanding with Language Models (arxiv.org) |
|
1 point by hdvr on Oct 8, 2024 | hide | past | pdf | discuss
|
| 13. |
Real Time Recommendation System with Collisionless Embedding Table (2022) (arxiv.org) |
|
1 point by janalsncm on Oct 8, 2024 | hide | past | pdf | discuss
|