about
Stories from May 31, 2023 (UTC)
Go back a day, month, or year. Go forward a day.
1. Grokking: Generalization Beyond Overfitting on Small Algorithmic Datasets (arxiv.org)
13 points by thunderbong on May 31, 2023 | hide | past | pdf | discuss
2. Whose Opinions Do Language Models Reflect? (arxiv.org)
4 points by rntn on May 31, 2023 | hide | past | pdf | 1 comment
3. Direct Preference Optimization: Your Language Model Is a Reward Model (arxiv.org)
4 points by tim_sw on May 31, 2023 | hide | past | pdf | discuss
4. Fine-Tuning Language Models with Just Forward Passes (arxiv.org)
3 points by famouswaffles on May 31, 2023 | hide | past | pdf | 1 comment
5. Intriguing Properties of Quantization at Scale (arxiv.org)
2 points by Jimmc414 on May 31, 2023 | hide | past | pdf | discuss
6. NetHack Is Hard to Hack (arxiv.org)
2 points by bikenaga on May 31, 2023 | hide | past | pdf | discuss
7. WikiChat: A Few-Shot LLM-Based Chatbot Grounded with Wikipedia (arxiv.org)
2 points by gdss on May 31, 2023 | hide | past | pdf | discuss
8. Concise Answers to Complex Questions: Summarization of Long-Form Answers (arxiv.org)
1 point by Jimmc414 on May 31, 2023 | hide | past | pdf | discuss
9. Outline, Then Details: Syntactically Guided Coarse-to-Fine Code Generation (arxiv.org)
1 point by PaulHoule on May 31, 2023 | hide | past | pdf | discuss