about
Stories from October 3, 2024 (UTC)
Go back a day, month, or year. Go forward a day.
1. Were RNNs all we needed? (arxiv.org)
520 points by beefman on Oct 3, 2024 | hide | past | pdf | 260 comments
2. Serving 70B-scale LLMs efficiently on low-resource edge devices [pdf] (arxiv.org)
248 points by simonpure on Oct 3, 2024 | hide | past | pdf | 58 comments
3. Thermodynamic Bayesian Inference (arxiv.org)
7 points by aifer4 on Oct 3, 2024 | hide | past | pdf | 1 comment
4. Rasterized Edge Gradients: Handling Discontinuities Differentiably (arxiv.org)
2 points by mfiguiere on Oct 3, 2024 | hide | past | pdf | discuss
5. Do Large Language Models Need a Content Delivery Network? (arxiv.org)
2 points by PaulHoule on Oct 3, 2024 | hide | past | pdf | discuss
6. Towards Social AI: A Survey on Understanding Social Interactions (arxiv.org)
1 point by PaulHoule on Oct 3, 2024 | hide | past | pdf | discuss
7. Evaluating Visual Perspective Taking in Vision Language Models (arxiv.org)
1 point by PaulHoule on Oct 3, 2024 | hide | past | pdf | discuss
8. Recall: Empowering Multimodal Embedding for Edge Devices (arxiv.org)
1 point by PaulHoule on Oct 3, 2024 | hide | past | pdf | discuss
9. To CoT or not? Chain-of-thought helps mainly on math and symbolic reasoning (arxiv.org)
1 point by amichail on Oct 3, 2024 | hide | past | pdf | discuss
10. Emergent Abilities of Large Language Models (2022) (arxiv.org)
1 point by squircle on Oct 3, 2024 | hide | past | pdf | discuss
11. Embers of Autoregression in OpenAI O1 (arxiv.org)
1 point by hdvr on Oct 3, 2024 | hide | past | pdf | discuss