|
|
Stories from July 11, 2024 (UTC)
|
| 1. |
Training a time series model using transformers at Datadog (arxiv.org) |
|
27 points by dbenamy on Jul 11, 2024 | hide | past | pdf | discuss
|
| 2. |
PaliGemma: A versatile 3B VLM for transfer (arxiv.org) |
|
5 points by tosh on Jul 11, 2024 | hide | past | pdf | discuss
|
| 3. |
Distilling System 2 into System 1 (arxiv.org) |
|
4 points by tosh on Jul 11, 2024 | hide | past | pdf | discuss
|
| 4. |
OpenDiLoCo: Open-Source Framework for Distributed Low-Communication Training (arxiv.org) |
|
4 points by Mougatine on Jul 11, 2024 | hide | past | pdf | discuss
|
| 5. |
CBT-LLM: A Chinese Large Language Model for Cognitive Behavioral Therapy (arxiv.org) |
|
3 points by aeontech on Jul 11, 2024 | hide | past | pdf | 1 comment
|
| 6. |
Cascade Reward Sampling for Efficient Decoding-Time Alignment (arxiv.org) |
|
3 points by Garcia98 on Jul 11, 2024 | hide | past | pdf | discuss
|
| 7. |
Mixture of a Million Experts (arxiv.org) |
|
3 points by quxinxin on Jul 11, 2024 | hide | past | pdf | discuss
|
| 8. |
Formal Aspects of Language Modeling (arxiv.org) |
|
2 points by Anon84 on Jul 11, 2024 | hide | past | pdf | discuss
|
| 9. |
When LLMs Play the Telephone Game: Cumulative Changes and Attractors in Iterated (arxiv.org) |
|
2 points by rcmcintosh on Jul 11, 2024 | hide | past | pdf | discuss
|
| 10. |
Facts About Building Retrieval Augmented Generation-Based Chatbots (arxiv.org) |
|
1 point by belter on Jul 11, 2024 | hide | past | pdf | discuss
|
|