|
|
Stories from October 3, 2024 (UTC)
|
| 1. |
Were RNNs all we needed? (arxiv.org) |
|
520 points by beefman on Oct 3, 2024 | hide | past | pdf | 260 comments
|
| 2. |
Serving 70B-scale LLMs efficiently on low-resource edge devices [pdf] (arxiv.org) |
|
248 points by simonpure on Oct 3, 2024 | hide | past | pdf | 58 comments
|
| 3. |
Thermodynamic Bayesian Inference (arxiv.org) |
|
7 points by aifer4 on Oct 3, 2024 | hide | past | pdf | 1 comment
|
| 4. |
Rasterized Edge Gradients: Handling Discontinuities Differentiably (arxiv.org) |
|
2 points by mfiguiere on Oct 3, 2024 | hide | past | pdf | discuss
|
| 5. |
Do Large Language Models Need a Content Delivery Network? (arxiv.org) |
|
2 points by PaulHoule on Oct 3, 2024 | hide | past | pdf | discuss
|
| 6. |
Towards Social AI: A Survey on Understanding Social Interactions (arxiv.org) |
|
1 point by PaulHoule on Oct 3, 2024 | hide | past | pdf | discuss
|
| 7. |
Evaluating Visual Perspective Taking in Vision Language Models (arxiv.org) |
|
1 point by PaulHoule on Oct 3, 2024 | hide | past | pdf | discuss
|
| 8. |
Recall: Empowering Multimodal Embedding for Edge Devices (arxiv.org) |
|
1 point by PaulHoule on Oct 3, 2024 | hide | past | pdf | discuss
|
| 9. |
To CoT or not? Chain-of-thought helps mainly on math and symbolic reasoning (arxiv.org) |
|
1 point by amichail on Oct 3, 2024 | hide | past | pdf | discuss
|
| 10. |
Emergent Abilities of Large Language Models (2022) (arxiv.org) |
|
1 point by squircle on Oct 3, 2024 | hide | past | pdf | discuss
|
| 11. |
Embers of Autoregression in OpenAI O1 (arxiv.org) |
|
1 point by hdvr on Oct 3, 2024 | hide | past | pdf | discuss
|
|