about
Stories from April 1, 2024 (UTC)
Go back a day, month, or year. Go forward a day.
1. Jamba: A Hybrid Transformer-Mamba Language Model (arxiv.org)
74 points by eitanturok on Apr 1, 2024 | hide | past | pdf | 6 comments
2. Born with a Silver Spoon? Socioeconomic Bias in Large Language Models (arxiv.org)
3 points by PaulHoule on Apr 1, 2024 | hide | past | pdf | discuss
3. Topos of Transformer Networks – Thoughts? (arxiv.org)
3 points by cheekyfibonacci on Apr 1, 2024 | hide | past | pdf | 1 comment
4. ReALM: Reference Resolution as Language Modeling (arxiv.org)
2 points by davidbarker on Apr 1, 2024 | hide | past | pdf | discuss
5. Model Stock: All we need is just a few fine-tuned models (arxiv.org)
2 points by tosh on Apr 1, 2024 | hide | past | pdf | discuss
6. What Was Your Prompt? A Remote Keylogging Attack on AI Assistants (arxiv.org)
2 points by maayank on Apr 1, 2024 | hide | past | pdf | discuss
7. Long-form factuality in large language models (arxiv.org)
2 points by rntn on Apr 1, 2024 | hide | past | pdf | discuss
8. Agents Need Not Know Their Purpose (arxiv.org)
2 points by optimalsolver on Apr 1, 2024 | hide | past | pdf | discuss
9. Are We on the Right Way for Evaluating Large Vision-Language Models? (arxiv.org)
2 points by yhzan on Apr 1, 2024 | hide | past | pdf | discuss
10. Building a Vietnamese Language Model with Advanced Continual Pre-Training (arxiv.org)
1 point by PaulHoule on Apr 1, 2024 | hide | past | pdf | discuss
11. Unsolvable Problem Detection: Evaluating Trustworthiness of Vision LMs (arxiv.org)
1 point by villaaston1 on Apr 1, 2024 | hide | past | pdf | discuss
12. Editing Models with Task Arithmetic (2022) (arxiv.org)
1 point by tosh on Apr 1, 2024 | hide | past | pdf | discuss
13. Instruction-Following Evaluation for Large Language Models (arxiv.org)
1 point by tosh on Apr 1, 2024 | hide | past | pdf | discuss