about
4141. The structure of the token space for large language models (arxiv.org)
2 points by pizza on Nov 24, 2024 | hide | past | pdf | discuss
4142. Evaluating the Robustness of Analogical Reasoning in Large Language Models (arxiv.org)
1 point by benchmarkist on Nov 23, 2024 | hide | past | pdf | discuss
4143. Learning Lossless Compression for High Bit-Depth Volumetric Medical Image (arxiv.org)
2 points by ksec on Nov 23, 2024 | hide | past | pdf | discuss
4144. Unlocking State-Tracking in Linear RNNs Through Negative Eigenvalues (arxiv.org)
2 points by jul8234 on Nov 23, 2024 | hide | past | pdf | 1 comment
4145. DroidSpeak: Enhancing Cross-LLM Communication (arxiv.org)
17 points by jonbaer on Nov 23, 2024 | hide | past | pdf | 2 comments
4146. A Critique of Unfounded Skepticism Around AI for Chip Design (arxiv.org)
3 points by utopcell on Nov 23, 2024 | hide | past | pdf | discuss
4147. Samurai: Adapting Segment Anything Model for Zero-Shot Visual Tracking (arxiv.org)
55 points by fzliu on Nov 22, 2024 | hide | past | pdf | discuss
4148. Driven by Compression Progress (arxiv.org)
2 points by handfuloflight on Nov 22, 2024 | hide | past | pdf | discuss
4149. Model-Based Transfer Learning for Contextual Reinforcement Learning (arxiv.org)
2 points by handfuloflight on Nov 22, 2024 | hide | past | pdf | discuss
4150. Improved GUI Grounding via Iterative Narrowing (arxiv.org)
2 points by sandwichsphinx on Nov 22, 2024 | hide | past | pdf | discuss
4151. Marco-O1: Towards Open Reasoning Models for Open-Ended Solutions (arxiv.org)
2 points by lnyan on Nov 22, 2024 | hide | past | pdf | discuss
4152. Automating LLM Development with LLMs (arxiv.org)
1 point by __tuxi__ on Nov 22, 2024 | hide | past | pdf | 1 comment
4153. Medical Video Generation for Disease Progression Simulation (arxiv.org)
2 points by IrohXu on Nov 22, 2024 | hide | past | pdf | discuss
4154. Adding Error Bars to Evals: A Statistical Approach to Language Model Evaluations (arxiv.org)
2 points by mnk47 on Nov 22, 2024 | hide | past | pdf | discuss
4155. Cramming: Training a Language Model on a Single GPU in One Day (arxiv.org)
3 points by andai on Nov 21, 2024 | hide | past | pdf | discuss
4156. Generative Agent Simulations of 1k People (arxiv.org)
1 point by klaussilveira on Nov 21, 2024 | hide | past | pdf | discuss
4157. WhisperNER: Unified Open Named Entity and Speech Recognition (arxiv.org)
133 points by timbilt on Nov 21, 2024 | hide | past | pdf | 17 comments
4158. Generating Science from AI-Powered Automated Falsification (arxiv.org)
1 point by omarsar on Nov 21, 2024 | hide | past | pdf | discuss
4159. SEFD: Semantic-Enhanced Framework for Detecting LLM-Generated Text (arxiv.org)
1 point by sandwichsphinx on Nov 21, 2024 | hide | past | pdf | discuss
4160. Generative Agent Simulations of 1k People (arxiv.org)
1 point by suprgeek on Nov 21, 2024 | hide | past | pdf | discuss
4161. Wave Network: An Ultra-Small Language Model (arxiv.org)
27 points by PaulHoule on Nov 21, 2024 | hide | past | pdf | 4 comments
4162. MolGrapher: Graph-Based Visual Recognition of Chemical Structures (2023) (arxiv.org)
2 points by sandwichsphinx on Nov 20, 2024 | hide | past | pdf | discuss
4163. Procedural Knowledge in Pretraining Drives Reasoning in Large Language Models (arxiv.org)
2 points by pongogogo on Nov 20, 2024 | hide | past | pdf | discuss
4164. Birdie: Advancing State Space Models with Reward-Driven Objectives and Curricula (arxiv.org)
3 points by PaulHoule on Nov 20, 2024 | hide | past | pdf | 1 comment
4165. Adding Error Bars to Evals: A Statistical Approach to Language Model Evaluations (arxiv.org)
1 point by paradite on Nov 20, 2024 | hide | past | pdf | discuss
4166. Animal Behavior Analysis Methods Using Deep Learning: A Survey (arxiv.org)
3 points by ctoth on Nov 19, 2024 | hide | past | pdf | discuss
4167. The Dawn of AI-Native EDA: Opportunities and Challenges of Large Circuit Models (arxiv.org)
3 points by petra on Nov 19, 2024 | hide | past | pdf | discuss
4168. Is Flash Attention Stable? (No) (arxiv.org)
1 point by anewhnaccount2 on Nov 19, 2024 | hide | past | pdf | discuss
4169. LLMs' Political Leaning and Their Influence on Voters (arxiv.org)
1 point by tristanMatthias on Nov 19, 2024 | hide | past | pdf | discuss
4170. RAG and Beyond: How to Make Your LLMs Use External Data More Wisely (arxiv.org)
1 point by gnabgib on Nov 19, 2024 | hide | past | pdf | discuss