about
Stories from February 28, 2023 (UTC)
Go back a day, month, or year. Go forward a day.
1. Transformer learning explained: Coinductive guide to inductive transformer heads (arxiv.org)
91 points by adamnemecek on Feb 28, 2023 | hide | past | pdf | 28 comments
2. Language Is Not All You Need: Aligning Perception with Language Models (arxiv.org)
87 points by craftsquick on Feb 28, 2023 | hide | past | pdf | 25 comments
3. Improving large language models with external knowledge and automated feedback (arxiv.org)
15 points by PaulHoule on Feb 28, 2023 | hide | past | pdf | discuss
4. Testing AI on less frequent aspects of language reveals insensitivity to meaning (arxiv.org)
2 points by PaulHoule on Feb 28, 2023 | hide | past | pdf | discuss
5. Inseq: An Interpretability Toolkit for Sequence Generation Models (arxiv.org)
1 point by gsarti96 on Feb 28, 2023 | hide | past | pdf | discuss