|
|
Stories from June 13, 2024 (UTC)
|
| 1. |
What If We Recaption Billions of Web Images with LLaMA-3? (arxiv.org) |
|
92 points by Jimmc414 on Jun 13, 2024 | hide | past | pdf | 40 comments
|
| 2. |
An Empirical Study of Mamba-Based Language Models (arxiv.org) |
|
43 points by panabee on Jun 13, 2024 | hide | past | pdf | 3 comments
|
| 3. |
Samba: Efficient Unlimited Context Language Modeling (arxiv.org) |
|
5 points by anon373839 on Jun 13, 2024 | hide | past | pdf | 1 comment
|
| 4. |
Creativity Has Left the Chat: The Price of Debiasing Language Models (arxiv.org) |
|
3 points by cubefox on Jun 13, 2024 | hide | past | pdf | discuss
|
| 5. |
Evolution Through Large Models (arxiv.org) |
|
3 points by 8organicbits on Jun 13, 2024 | hide | past | pdf | discuss
|
| 6. |
Modeling Boundedly Rational Agents with Latent Inference Budgets (2023) (arxiv.org) |
|
2 points by belter on Jun 13, 2024 | hide | past | pdf | discuss
|
| 7. |
Discovering Preference Optimization Algorithms with Large Language Models (arxiv.org) |
|
2 points by jonbaer on Jun 13, 2024 | hide | past | pdf | discuss
|
| 8. |
Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing (arxiv.org) |
|
2 points by alins on Jun 13, 2024 | hide | past | pdf | discuss
|
| 9. |
What's the Magic Word? A Control Theory of LLM Prompting (arxiv.org) |
|
1 point by sporadicjoke on Jun 13, 2024 | hide | past | pdf | discuss
|
| 10. |
Can Language Models Use Forecasting Strategies? (arxiv.org) |
|
1 point by PaulHoule on Jun 13, 2024 | hide | past | pdf | discuss
|
| 11. |
Proofread: Fixes All Errors with One Tap (arxiv.org) |
|
1 point by PaulHoule on Jun 13, 2024 | hide | past | pdf | discuss
|
|