| 1. |
σ-GPTs: A new approach to autoregressive models (arxiv.org) |
|
293 points by mehulashah on Jun 7, 2024 | hide | past | pdf | 93 comments
|
| 2. |
The illusion of state in state-space models (arxiv.org) |
|
60 points by canjobear on Jun 7, 2024 | hide | past | pdf | 59 comments
|
| 3. |
Open-Endedness Is Essential for Artificial Superhuman Intelligence (arxiv.org) |
|
7 points by artninja1988 on Jun 7, 2024 | hide | past | pdf | discuss
|
| 4. |
Graph Convolutional Branch and Bound (arxiv.org) |
|
6 points by lorenzos98 on Jun 7, 2024 | hide | past | pdf | discuss
|
| 5. |
Vision-LSTM: xLSTM as Generic Vision Backbone (arxiv.org) |
|
5 points by tosh on Jun 7, 2024 | hide | past | pdf | discuss
|
| 6. |
Open-Endedness Is Essential for Artificial Superhuman Intelligence (arxiv.org) |
|
5 points by tzury on Jun 7, 2024 | hide | past | pdf | discuss
|
| 7. |
Scalable MatMul-Free Language Modeling (arxiv.org) |
|
3 points by optimalsolver on Jun 7, 2024 | hide | past | pdf | discuss
|
| 8. |
The Impacts of Data, Ordering, and Intrinsic Dimensionality on Recall in HNSW (arxiv.org) |
|
3 points by jn2clark on Jun 7, 2024 | hide | past | pdf | discuss
|
| 9. |
Scalable Detection of Salient Entities in News Articles (arxiv.org) |
|
2 points by PaulHoule on Jun 7, 2024 | hide | past | pdf | discuss
|
| 10. |
Benchmarking the Energy Costs of Large Language Model Inference (2023) (arxiv.org) |
|
2 points by mcguire on Jun 7, 2024 | hide | past | pdf | discuss
|
| 11. |
Ask LLMs Directly, "What shapes your bias?" (arxiv.org) |
|
2 points by belter on Jun 7, 2024 | hide | past | pdf | discuss
|
| 12. |
Potential Field Based Deep Metric Learning (arxiv.org) |
|
2 points by PaulHoule on Jun 7, 2024 | hide | past | pdf | discuss
|
| 13. |
Will we run out of data? Limits of LLM scaling based on human-generated data (arxiv.org) |
|
1 point by Smith42 on Jun 7, 2024 | hide | past | pdf | 1 comment
|