about
Stories from April 16, 2024 (UTC)
Go back a day, month, or year. Go forward a day.
1. Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length (arxiv.org)
168 points by amichail on Apr 16, 2024 | hide | past | pdf | 28 comments
2. ResearchAgent: Iterative Research Idea Generation Using LLMs (arxiv.org)
124 points by milliondreams on Apr 16, 2024 | hide | past | pdf | 63 comments
3. TransformerFAM: Feedback attention is working memory (arxiv.org)
4 points by tosh on Apr 16, 2024 | hide | past | pdf | discuss
4. Megalodon: Efficient LLM Pretraining and Inference with Unlimited Context Length (arxiv.org)
4 points by tosh on Apr 16, 2024 | hide | past | pdf | discuss
5. LLM in a Flash: Efficient Large Language Model Inference with Limited Memory (arxiv.org)
2 points by abhinavk on Apr 16, 2024 | hide | past | pdf | discuss
6. Quality at a Glance: An Audit of Web-Crawled Multilingual Datasets (2022) (arxiv.org)
2 points by tosh on Apr 16, 2024 | hide | past | pdf | discuss
7. A Model-Based and Imitation Learning Deep Reinforcement Learning Hybrid (arxiv.org)
2 points by lucaspauker on Apr 16, 2024 | hide | past | pdf | discuss
8. Writing Wikipedia-Like Articles from Scratch with LLMs (arxiv.org)
1 point by talonx on Apr 16, 2024 | hide | past | pdf | discuss