| 4141. |
The structure of the token space for large language models (arxiv.org) |
|
2 points by pizza on Nov 24, 2024 | hide | past | pdf | discuss
|
| 4142. |
Evaluating the Robustness of Analogical Reasoning in Large Language Models (arxiv.org) |
|
1 point by benchmarkist on Nov 23, 2024 | hide | past | pdf | discuss
|
| 4143. |
Learning Lossless Compression for High Bit-Depth Volumetric Medical Image (arxiv.org) |
|
2 points by ksec on Nov 23, 2024 | hide | past | pdf | discuss
|
| 4144. |
Unlocking State-Tracking in Linear RNNs Through Negative Eigenvalues (arxiv.org) |
|
2 points by jul8234 on Nov 23, 2024 | hide | past | pdf | 1 comment
|
| 4145. |
DroidSpeak: Enhancing Cross-LLM Communication (arxiv.org) |
|
17 points by jonbaer on Nov 23, 2024 | hide | past | pdf | 2 comments
|
| 4146. |
A Critique of Unfounded Skepticism Around AI for Chip Design (arxiv.org) |
|
3 points by utopcell on Nov 23, 2024 | hide | past | pdf | discuss
|
| 4147. |
Samurai: Adapting Segment Anything Model for Zero-Shot Visual Tracking (arxiv.org) |
|
55 points by fzliu on Nov 22, 2024 | hide | past | pdf | discuss
|
| 4148. |
Driven by Compression Progress (arxiv.org) |
|
2 points by handfuloflight on Nov 22, 2024 | hide | past | pdf | discuss
|
| 4149. |
Model-Based Transfer Learning for Contextual Reinforcement Learning (arxiv.org) |
|
2 points by handfuloflight on Nov 22, 2024 | hide | past | pdf | discuss
|
| 4150. |
Improved GUI Grounding via Iterative Narrowing (arxiv.org) |
|
2 points by sandwichsphinx on Nov 22, 2024 | hide | past | pdf | discuss
|
| 4151. |
Marco-O1: Towards Open Reasoning Models for Open-Ended Solutions (arxiv.org) |
|
2 points by lnyan on Nov 22, 2024 | hide | past | pdf | discuss
|
| 4152. |
Automating LLM Development with LLMs (arxiv.org) |
|
1 point by __tuxi__ on Nov 22, 2024 | hide | past | pdf | 1 comment
|
| 4153. |
Medical Video Generation for Disease Progression Simulation (arxiv.org) |
|
2 points by IrohXu on Nov 22, 2024 | hide | past | pdf | discuss
|
| 4154. |
Adding Error Bars to Evals: A Statistical Approach to Language Model Evaluations (arxiv.org) |
|
2 points by mnk47 on Nov 22, 2024 | hide | past | pdf | discuss
|
| 4155. |
Cramming: Training a Language Model on a Single GPU in One Day (arxiv.org) |
|
3 points by andai on Nov 21, 2024 | hide | past | pdf | discuss
|
| 4156. |
Generative Agent Simulations of 1k People (arxiv.org) |
|
1 point by klaussilveira on Nov 21, 2024 | hide | past | pdf | discuss
|
| 4157. |
WhisperNER: Unified Open Named Entity and Speech Recognition (arxiv.org) |
|
133 points by timbilt on Nov 21, 2024 | hide | past | pdf | 17 comments
|
| 4158. |
Generating Science from AI-Powered Automated Falsification (arxiv.org) |
|
1 point by omarsar on Nov 21, 2024 | hide | past | pdf | discuss
|
| 4159. |
SEFD: Semantic-Enhanced Framework for Detecting LLM-Generated Text (arxiv.org) |
|
1 point by sandwichsphinx on Nov 21, 2024 | hide | past | pdf | discuss
|
| 4160. |
Generative Agent Simulations of 1k People (arxiv.org) |
|
1 point by suprgeek on Nov 21, 2024 | hide | past | pdf | discuss
|
| 4161. |
Wave Network: An Ultra-Small Language Model (arxiv.org) |
|
27 points by PaulHoule on Nov 21, 2024 | hide | past | pdf | 4 comments
|
| 4162. |
MolGrapher: Graph-Based Visual Recognition of Chemical Structures (2023) (arxiv.org) |
|
2 points by sandwichsphinx on Nov 20, 2024 | hide | past | pdf | discuss
|
| 4163. |
Procedural Knowledge in Pretraining Drives Reasoning in Large Language Models (arxiv.org) |
|
2 points by pongogogo on Nov 20, 2024 | hide | past | pdf | discuss
|
| 4164. |
Birdie: Advancing State Space Models with Reward-Driven Objectives and Curricula (arxiv.org) |
|
3 points by PaulHoule on Nov 20, 2024 | hide | past | pdf | 1 comment
|
| 4165. |
Adding Error Bars to Evals: A Statistical Approach to Language Model Evaluations (arxiv.org) |
|
1 point by paradite on Nov 20, 2024 | hide | past | pdf | discuss
|
| 4166. |
Animal Behavior Analysis Methods Using Deep Learning: A Survey (arxiv.org) |
|
3 points by ctoth on Nov 19, 2024 | hide | past | pdf | discuss
|
| 4167. |
The Dawn of AI-Native EDA: Opportunities and Challenges of Large Circuit Models (arxiv.org) |
|
3 points by petra on Nov 19, 2024 | hide | past | pdf | discuss
|
| 4168. |
Is Flash Attention Stable? (No) (arxiv.org) |
|
1 point by anewhnaccount2 on Nov 19, 2024 | hide | past | pdf | discuss
|
| 4169. |
LLMs' Political Leaning and Their Influence on Voters (arxiv.org) |
|
1 point by tristanMatthias on Nov 19, 2024 | hide | past | pdf | discuss
|
| 4170. |
RAG and Beyond: How to Make Your LLMs Use External Data More Wisely (arxiv.org) |
|
1 point by gnabgib on Nov 19, 2024 | hide | past | pdf | discuss
|
| More |