| 5041. |
Empirical influence functions to understand the logic of fine-tuning (arxiv.org) |
|
1 point by j6m8 on Jun 5, 2024 | hide | past | pdf | discuss
|
| 5042. |
Simple tasks showing reasoning breakdown in state-of-the-art LLMs (arxiv.org) |
|
375 points by tosh on Jun 5, 2024 | hide | past | pdf | 380 comments
|
| 5043. |
To Believe or Not Believe Your LLM (arxiv.org) |
|
58 points by josh-sematic on Jun 5, 2024 | hide | past | pdf | 17 comments
|
| 5044. |
Knockout: A simple way to handle missing inputs (arxiv.org) |
|
3 points by jasondavies on Jun 5, 2024 | hide | past | pdf | discuss
|
| 5045. |
Tiny Time Mixers (TTMs) (arxiv.org) |
|
3 points by nojito on Jun 5, 2024 | hide | past | pdf | discuss
|
| 5046. |
The Geometry of Categorical and Hierarchical Concepts in Large Language Models (arxiv.org) |
|
3 points by davedx on Jun 4, 2024 | hide | past | pdf | discuss
|
| 5047. |
Understanding Recall in HNSW Search (arxiv.org) |
|
3 points by esleightholm on Jun 4, 2024 | hide | past | pdf | discuss
|
| 5048. |
LLMs can learn self-restraint through iterative self-reflection (arxiv.org) |
|
3 points by PaulHoule on Jun 4, 2024 | hide | past | pdf | discuss
|
| 5049. |
SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales (arxiv.org) |
|
28 points by jasondavies on Jun 4, 2024 | hide | past | pdf | 9 comments
|
| 5050. |
Diffusion on Syntax Trees for Program Synthesis (arxiv.org) |
|
2 points by 23B1 on Jun 4, 2024 | hide | past | pdf | 1 comment
|
| 5051. |
LLMs achieve adult human performance on higher-order theory of mind tasks (arxiv.org) |
|
2 points by tosh on Jun 3, 2024 | hide | past | pdf | 1 comment
|
| 5052. |
Grokfast: Accelerated Grokking by Amplifying Slow Gradients (arxiv.org) |
|
117 points by johnsutor on Jun 3, 2024 | hide | past | pdf | 39 comments
|
| 5053. |
Leveraging Human Revisions for Improving Text-to-Layout Models (arxiv.org) |
|
2 points by PaulHoule on Jun 3, 2024 | hide | past | pdf | discuss
|
| 5054. |
Evaluating Gemini Models for Dangerous Capabilities (arxiv.org) |
|
3 points by fishfish on Jun 3, 2024 | hide | past | pdf | discuss
|
| 5055. |
Kotlin ML Pack: Technical Report (arxiv.org) |
|
1 point by belter on Jun 3, 2024 | hide | past | pdf | discuss
|
| 5056. |
There and Back Again: The AI Alignment Paradox (arxiv.org) |
|
2 points by belter on Jun 3, 2024 | hide | past | pdf | discuss
|
| 5057. |
Transformers are SSMs (Mamba-2) (arxiv.org) |
|
2 points by jasondavies on Jun 3, 2024 | hide | past | pdf | discuss
|
| 5058. |
Is Complexity an Illusion? (arxiv.org) |
|
2 points by broyojo on Jun 3, 2024 | hide | past | pdf | 1 comment
|
| 5059. |
Faithful Logical Reasoning via Symbolic Chain-of-Thought (arxiv.org) |
|
2 points by burakemir on Jun 3, 2024 | hide | past | pdf | discuss
|
| 5060. |
Swarm Parallelism: Training Large Models on Poorly Connected Devices (arxiv.org) |
|
2 points by kwindla on Jun 3, 2024 | hide | past | pdf | discuss
|
| 5061. |
Agent Hospital: A Simulacrum of Hospital with Evolvable Medical Agents (arxiv.org) |
|
2 points by jonbaer on Jun 2, 2024 | hide | past | pdf | discuss
|
| 5062. |
ToonCrafter: Generative Cartoon Interpolation (arxiv.org) |
|
3 points by lmariscal on Jun 2, 2024 | hide | past | pdf | 2 comments
|
| 5063. |
Contextual Position Encoding: Learning to Count What's Important (arxiv.org) |
|
2 points by sebzim4500 on Jun 2, 2024 | hide | past | pdf | 1 comment
|
| 5064. |
A survey on text generation using generative adversarial networks (arxiv.org) |
|
2 points by Anon84 on Jun 2, 2024 | hide | past | pdf | discuss
|
| 5065. |
Analogies Explained: Towards Understanding Word Embeddings (arxiv.org) |
|
1 point by Anon84 on Jun 2, 2024 | hide | past | pdf | discuss
|
| 5066. |
Is Model Collapse Inevitable? Breaking the Curse of Recursion (arxiv.org) |
|
2 points by colinprince on Jun 2, 2024 | hide | past | pdf | discuss
|
| 5067. |
Is In-Context Learning Sufficient for Instruction Following in LLMs? (arxiv.org) |
|
2 points by max-andr on Jun 1, 2024 | hide | past | pdf | 1 comment
|
| 5068. |
LLMs achieve adult human performance on higher-order theory of mind tasks (arxiv.org) |
|
1 point by amichail on May 31, 2024 | hide | past | pdf | 2 comments
|
| 5069. |
Compressed-Language Models for Understanding Compressed File Formats: JPEG (arxiv.org) |
|
3 points by jasondavies on May 31, 2024 | hide | past | pdf | discuss
|
| 5070. |
Sparse Maximal Update Parameterization (arxiv.org) |
|
2 points by cs-fan-101 on May 31, 2024 | hide | past | pdf | discuss
|
| More |