| 1. |
Jamba: A Hybrid Transformer-Mamba Language Model (arxiv.org) |
|
74 points by eitanturok on Apr 1, 2024 | hide | past | pdf | 6 comments
|
| 2. |
Born with a Silver Spoon? Socioeconomic Bias in Large Language Models (arxiv.org) |
|
3 points by PaulHoule on Apr 1, 2024 | hide | past | pdf | discuss
|
| 3. |
Topos of Transformer Networks – Thoughts? (arxiv.org) |
|
3 points by cheekyfibonacci on Apr 1, 2024 | hide | past | pdf | 1 comment
|
| 4. |
ReALM: Reference Resolution as Language Modeling (arxiv.org) |
|
2 points by davidbarker on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5. |
Model Stock: All we need is just a few fine-tuned models (arxiv.org) |
|
2 points by tosh on Apr 1, 2024 | hide | past | pdf | discuss
|
| 6. |
What Was Your Prompt? A Remote Keylogging Attack on AI Assistants (arxiv.org) |
|
2 points by maayank on Apr 1, 2024 | hide | past | pdf | discuss
|
| 7. |
Long-form factuality in large language models (arxiv.org) |
|
2 points by rntn on Apr 1, 2024 | hide | past | pdf | discuss
|
| 8. |
Agents Need Not Know Their Purpose (arxiv.org) |
|
2 points by optimalsolver on Apr 1, 2024 | hide | past | pdf | discuss
|
| 9. |
Are We on the Right Way for Evaluating Large Vision-Language Models? (arxiv.org) |
|
2 points by yhzan on Apr 1, 2024 | hide | past | pdf | discuss
|
| 10. |
Building a Vietnamese Language Model with Advanced Continual Pre-Training (arxiv.org) |
|
1 point by PaulHoule on Apr 1, 2024 | hide | past | pdf | discuss
|
| 11. |
Unsolvable Problem Detection: Evaluating Trustworthiness of Vision LMs (arxiv.org) |
|
1 point by villaaston1 on Apr 1, 2024 | hide | past | pdf | discuss
|
| 12. |
Editing Models with Task Arithmetic (2022) (arxiv.org) |
|
1 point by tosh on Apr 1, 2024 | hide | past | pdf | discuss
|
| 13. |
Instruction-Following Evaluation for Large Language Models (arxiv.org) |
|
1 point by tosh on Apr 1, 2024 | hide | past | pdf | discuss
|