| 4981. |
The Prompt a Systematic Survey of Prompting Techniques (arxiv.org) |
|
1 point by tosh on Jun 14, 2024 | hide | past | pdf | discuss
|
| 4982. |
What's the Magic Word? A Control Theory of LLM Prompting (arxiv.org) |
|
1 point by sporadicjoke on Jun 13, 2024 | hide | past | pdf | discuss
|
| 4983. |
Creativity Has Left the Chat: The Price of Debiasing Language Models (arxiv.org) |
|
3 points by cubefox on Jun 13, 2024 | hide | past | pdf | discuss
|
| 4984. |
An Empirical Study of Mamba-Based Language Models (arxiv.org) |
|
43 points by panabee on Jun 13, 2024 | hide | past | pdf | 3 comments
|
| 4985. |
Can Language Models Use Forecasting Strategies? (arxiv.org) |
|
1 point by PaulHoule on Jun 13, 2024 | hide | past | pdf | discuss
|
| 4986. |
Proofread: Fixes All Errors with One Tap (arxiv.org) |
|
1 point by PaulHoule on Jun 13, 2024 | hide | past | pdf | discuss
|
| 4987. |
Samba: Efficient Unlimited Context Language Modeling (arxiv.org) |
|
5 points by anon373839 on Jun 13, 2024 | hide | past | pdf | 1 comment
|
| 4988. |
Modeling Boundedly Rational Agents with Latent Inference Budgets (2023) (arxiv.org) |
|
2 points by belter on Jun 13, 2024 | hide | past | pdf | discuss
|
| 4989. |
Evolution Through Large Models (arxiv.org) |
|
3 points by 8organicbits on Jun 13, 2024 | hide | past | pdf | discuss
|
| 4990. |
Discovering Preference Optimization Algorithms with Large Language Models (arxiv.org) |
|
2 points by jonbaer on Jun 13, 2024 | hide | past | pdf | discuss
|
| 4991. |
Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing (arxiv.org) |
|
2 points by alins on Jun 13, 2024 | hide | past | pdf | discuss
|
| 4992. |
What If We Recaption Billions of Web Images with LLaMA-3? (arxiv.org) |
|
92 points by Jimmc414 on Jun 13, 2024 | hide | past | pdf | 40 comments
|
| 4993. |
21.2× faster than llama.cpp? plus 40% memory usage reduction (arxiv.org) |
|
43 points by helloericsf on Jun 12, 2024 | hide | past | pdf | 14 comments
|
| 4994. |
Poco: Policy Composition from and for Heterogeneous Robot Learning (arxiv.org) |
|
2 points by rntn on Jun 12, 2024 | hide | past | pdf | discuss
|
| 4995. |
Accessing Math Solutions via Monte Carlo Self-Refine with LLaMa-3 8B (arxiv.org) |
|
100 points by belter on Jun 12, 2024 | hide | past | pdf | 9 comments
|
| 4996. |
Gzip Predicts Data-Dependent Scaling Laws (arxiv.org) |
|
3 points by RafelMri on Jun 12, 2024 | hide | past | pdf | discuss
|
| 4997. |
Improve Math Reasoning in Language Models by Automated Process Supervision (arxiv.org) |
|
3 points by amichail on Jun 12, 2024 | hide | past | pdf | discuss
|
| 4998. |
Samba: Simple Hybrid State Space Models (arxiv.org) |
|
3 points by jasondavies on Jun 12, 2024 | hide | past | pdf | discuss
|
| 4999. |
Can you protect your prompt from being stolen? (arxiv.org) |
|
2 points by m0g1cian on Jun 12, 2024 | hide | past | pdf | discuss
|
| 5000. |
The Prompt a Systematic Survey of Prompting Techniques (arxiv.org) |
|
2 points by sebg on Jun 12, 2024 | hide | past | pdf | discuss
|
| 5001. |
Large Language Models' Detection of Political Orientation in Newspapers (arxiv.org) |
|
2 points by PaulHoule on Jun 12, 2024 | hide | past | pdf | discuss
|
| 5002. |
Turbo Sparse: Achieving LLM SOTA Performance with Minimal Activated Parameters (arxiv.org) |
|
11 points by limoce on Jun 12, 2024 | hide | past | pdf | discuss
|
| 5003. |
TextGrad: Automatic "Differentiation" via Text (arxiv.org) |
|
2 points by jasondavies on Jun 12, 2024 | hide | past | pdf | discuss
|
| 5004. |
Zero-Shot Image Editing with Reference Imitation (arxiv.org) |
|
1 point by Jimmc414 on Jun 12, 2024 | hide | past | pdf | 1 comment
|
| 5005. |
An LLM-Based Recommender System Environment (arxiv.org) |
|
2 points by PaulHoule on Jun 12, 2024 | hide | past | pdf | discuss
|
| 5006. |
PyTorch-IE: Fast and Reproducible Prototyping for Information Extraction (arxiv.org) |
|
2 points by PaulHoule on Jun 11, 2024 | hide | past | pdf | discuss
|
| 5007. |
Self-Supervised Visual Grounding of Sound and Language (arxiv.org) |
|
2 points by belter on Jun 11, 2024 | hide | past | pdf | 1 comment
|
| 5008. |
Google: Towards a Personal Health Large Language Model (arxiv.org) |
|
3 points by tosh on Jun 11, 2024 | hide | past | pdf | 1 comment
|
| 5009. |
PowerInfer-2: Fast Large Language Model Inference on a Smartphone (arxiv.org) |
|
1 point by limoce on Jun 11, 2024 | hide | past | pdf | discuss
|
| 5010. |
GenAI Arena: An Open Evaluation Platform for Generative Models (arxiv.org) |
|
2 points by belter on Jun 11, 2024 | hide | past | pdf | discuss
|
| More |