| 5671. |
Tropical Geometry of Deep Neural Networks (arxiv.org) |
|
1 point by michelpp on Mar 9, 2024 | hide | past | pdf | discuss
|
| 5672. |
Algorithmic Complexities in Backpropagation and Tropical Neural Networks (arxiv.org) |
|
2 points by michelpp on Mar 9, 2024 | hide | past | pdf | discuss
|
| 5673. |
Self-Retrieval: Building an information retrieval system with one LLM (arxiv.org) |
|
200 points by PaulHoule on Mar 9, 2024 | hide | past | pdf | 29 comments
|
| 5674. |
Data Contamination and Evaluation Malpractices in Closed-Source LLMs (arxiv.org) |
|
2 points by adrianhoward on Mar 8, 2024 | hide | past | pdf | discuss
|
| 5675. |
Bootstrapping Cognitive Agents with a Large Language Model (arxiv.org) |
|
1 point by PaulHoule on Mar 8, 2024 | hide | past | pdf | discuss
|
| 5676. |
Gradient Low-Rank Projection (GaLore) reduces LLM training memory usage by 63% (arxiv.org) |
|
1 point by apsec112 on Mar 8, 2024 | hide | past | pdf | discuss
|
| 5677. |
Yi: Open Foundation Models (arxiv.org) |
|
1 point by tosh on Mar 8, 2024 | hide | past | pdf | discuss
|
| 5678. |
Can Large Language Models Reason and Plan? (arxiv.org) |
|
2 points by Jimmc414 on Mar 8, 2024 | hide | past | pdf | 1 comment
|
| 5679. |
The Unreasonable Effectiveness of Eccentric Automatic Prompts (arxiv.org) |
|
1 point by Jimmc414 on Mar 8, 2024 | hide | past | pdf | discuss
|
| 5680. |
Exposing Systemic Vulnerabilities of LLMs Through a Prompt Hacking Competition (arxiv.org) |
|
2 points by sohkamyung on Mar 8, 2024 | hide | past | pdf | discuss
|
| 5681. |
Machine learning and information theory concepts towards an AI Mathematician (arxiv.org) |
|
3 points by carlossouza on Mar 8, 2024 | hide | past | pdf | 1 comment
|
| 5682. |
PixArt-Σ: Weak-to-Strong Training of Diffusion Transformer for Text-to-Image (arxiv.org) |
|
1 point by artninja1988 on Mar 8, 2024 | hide | past | pdf | discuss
|
| 5683. |
PersonaLLM: Investigating the Ability of LLMs to Express Personality Traits (arxiv.org) |
|
1 point by swyx on Mar 7, 2024 | hide | past | pdf | discuss
|
| 5684. |
Prompt Decomposition and Reconstruction Makes Powerful LLM Jailbreakers (arxiv.org) |
|
2 points by PaulHoule on Mar 7, 2024 | hide | past | pdf | discuss
|
| 5685. |
Large language models surpass human experts in predicting neuroscience results (arxiv.org) |
|
2 points by MauranKilom on Mar 7, 2024 | hide | past | pdf | discuss
|
| 5686. |
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection (arxiv.org) |
|
6 points by victormustar on Mar 7, 2024 | hide | past | pdf | discuss
|
| 5687. |
RepairLLaMA: Efficient Representations and Fine-Tuned Adapter for Program Repair (arxiv.org) |
|
2 points by andre15silva on Mar 7, 2024 | hide | past | pdf | discuss
|
| 5688. |
Learning to Decode Collaboratively with Multiple Language Models (arxiv.org) |
|
2 points by padolsey on Mar 7, 2024 | hide | past | pdf | discuss
|
| 5689. |
SaulLM-7B: A Pioneering Large Language Model for Law (arxiv.org) |
|
3 points by telmop on Mar 7, 2024 | hide | past | pdf | discuss
|
| 5690. |
Curious Decline of Linguistic Diversity: Training LLMs on Synthetic Text (2023) (arxiv.org) |
|
2 points by 1vuio0pswjnm7 on Mar 7, 2024 | hide | past | pdf | 1 comment
|
| 5691. |
How Do LLMs Answer Multiple-Choice Questions Without the Question? (arxiv.org) |
|
2 points by smusamashah on Mar 6, 2024 | hide | past | pdf | discuss
|
| 5692. |
MacGyver: Are Large Language Models Creative Problem Solvers? (arxiv.org) |
|
1 point by ulrischa on Mar 6, 2024 | hide | past | pdf | 1 comment
|
| 5693. |
LDB: A Large Language Model Debugger via Verifying Runtime Execution (arxiv.org) |
|
1 point by dmarchand90 on Mar 6, 2024 | hide | past | pdf | 1 comment
|
| 5694. |
Design2Code: How Far Are We from Automating Front-End Engineering? (arxiv.org) |
|
1 point by victormustar on Mar 6, 2024 | hide | past | pdf | discuss
|
| 5695. |
Sora: Review on Background, Tech, Limits, and Opportunities of Vision Models (arxiv.org) |
|
33 points by shinryudbz on Mar 6, 2024 | hide | past | pdf | 2 comments
|
| 5696. |
If in a Crowdsourced Data Annotation Pipeline, a GPT-4 (arxiv.org) |
|
1 point by fatso784 on Mar 5, 2024 | hide | past | pdf | discuss
|
| 5697. |
StarCoder 2 and the Stack v2 (arxiv.org) |
|
1 point by tosh on Mar 5, 2024 | hide | past | pdf | discuss
|
| 5698. |
Dialect prejudice predicts AI decisions about people's character (arxiv.org) |
|
4 points by saassiopeia on Mar 5, 2024 | hide | past | pdf | 2 comments
|
| 5699. |
Why do tree-based models still outperform deep learning on tabular data? (2022) (arxiv.org) |
|
212 points by tosh on Mar 5, 2024 | hide | past | pdf | 111 comments
|
| 5700. |
LLM Ensemble Prediction Capabilities Match Human Crowd Accuracy (arxiv.org) |
|
1 point by SchoeneggerP on Mar 5, 2024 | hide | past | pdf | 2 comments
|
| More |