| 4081. |
Effective Strategies for Mitigating Hallucinations in LLMs for Data Analytics (arxiv.org) |
|
2 points by micrum on Dec 3, 2024 | hide | past | pdf | discuss
|
| 4082. |
AI Meets Antimatter: Unveiling Antihydrogen Annihilations (arxiv.org) |
|
1 point by throwawayed1 on Dec 3, 2024 | hide | past | pdf | discuss
|
| 4083. |
LLMs as Method Actors: A Model for Prompt Engineering and Architecture (arxiv.org) |
|
2 points by gronky_ on Dec 3, 2024 | hide | past | pdf | 1 comment
|
| 4084. |
Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning (arxiv.org) |
|
11 points by rntn on Dec 3, 2024 | hide | past | pdf | 1 comment
|
| 4085. |
Beyond Language Models: Byte Models Are Digital World Simulators (arxiv.org) |
|
4 points by bob1029 on Dec 3, 2024 | hide | past | pdf | discuss
|
| 4086. |
Auto-RAG (arxiv.org) |
|
3 points by omarsar on Dec 3, 2024 | hide | past | pdf | discuss
|
| 4087. |
Star: Synthesis of Tailored Architectures (arxiv.org) |
|
1 point by simonpure on Dec 3, 2024 | hide | past | pdf | discuss
|
| 4088. |
Neural simulation-based inference for parameter estimation in ATLAS (arxiv.org) |
|
1 point by pizza on Dec 3, 2024 | hide | past | pdf | discuss
|
| 4089. |
Demo: Decoupled Momentum Optimization (arxiv.org) |
|
1 point by simonpure on Dec 3, 2024 | hide | past | pdf | discuss
|
| 4090. |
Modeling AdaGrad, RMSProp, and Adam with Integro-Differential Equations (arxiv.org) |
|
2 points by PaulHoule on Dec 2, 2024 | hide | past | pdf | discuss
|
| 4091. |
Generative Agent Simulations of 1k People (arxiv.org) |
|
2 points by wslh on Dec 2, 2024 | hide | past | pdf | discuss
|
| 4092. |
Large Language Model-Brained GUI Agents: A Survey (arxiv.org) |
|
6 points by pongogogo on Dec 2, 2024 | hide | past | pdf | discuss
|
| 4093. |
Demo: Decoupled Momentum Optimization – training neural networks in parallel (arxiv.org) |
|
3 points by sergiotapia on Dec 2, 2024 | hide | past | pdf | discuss
|
| 4094. |
Reverse Thinking Makes LLMs Stronger Reasoners (arxiv.org) |
|
4 points by omarsar on Dec 2, 2024 | hide | past | pdf | discuss
|
| 4095. |
Do LLMs Perform Latent Multi-Hop Reasoning Without Exploiting Shortcuts? (arxiv.org) |
|
1 point by fzliu on Dec 2, 2024 | hide | past | pdf | discuss
|
| 4096. |
Procedural knowledge in pretraining drives reasoning in large language models (arxiv.org) |
|
248 points by reqo on Dec 1, 2024 | hide | past | pdf | 101 comments
|
| 4097. |
Challenges and Applications of Large Language Models (arxiv.org) |
|
1 point by t55 on Dec 1, 2024 | hide | past | pdf | discuss
|
| 4098. |
DynaSaur: Large Language Agents Beyond Predefined Actions (arxiv.org) |
|
128 points by surprisetalk on Dec 1, 2024 | hide | past | pdf | 31 comments
|
| 4099. |
The Curse of Recursion: Training on generated data makes models forget (2023) (arxiv.org) |
|
122 points by surprisetalk on Dec 1, 2024 | hide | past | pdf | 107 comments
|
| 4100. |
Streaming Deep Reinforcement Learning Works (arxiv.org) |
|
2 points by surprisetalk on Dec 1, 2024 | hide | past | pdf | discuss
|
| 4101. |
DrugAgent: AI-Aided Drug Discovery Programming Through LLM Multi-Agent Collab (arxiv.org) |
|
2 points by surprisetalk on Dec 1, 2024 | hide | past | pdf | discuss
|
| 4102. |
Large Language Models as Markov Chains (arxiv.org) |
|
75 points by Anon84 on Nov 30, 2024 | hide | past | pdf | 54 comments
|
| 4103. |
SnapMem: Snapshot-Based 3D Scene Memory for Embodied Exploration and Reasoning (arxiv.org) |
|
2 points by sandwichsphinx on Nov 30, 2024 | hide | past | pdf | discuss
|
| 4104. |
Generative Agent Simulations of 1k People (arxiv.org) |
|
1 point by doener on Nov 30, 2024 | hide | past | pdf | discuss
|
| 4105. |
Stable, Fast, Automatic Learning Algorithm for Predictive Coding Networks [pdf] (arxiv.org) |
|
3 points by kelseyfrog on Nov 29, 2024 | hide | past | pdf | discuss
|
| 4106. |
Super-Resolution with Low-Bit Quantization (arxiv.org) |
|
2 points by RobinHirst11 on Nov 29, 2024 | hide | past | pdf | discuss
|
| 4107. |
DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving (arxiv.org) |
|
2 points by RobinHirst11 on Nov 29, 2024 | hide | past | pdf | discuss
|
| 4108. |
Enhancing JEPAs with Spatial Conditioning: Efficient Representation Learning (arxiv.org) |
|
3 points by sandwichsphinx on Nov 29, 2024 | hide | past | pdf | discuss
|
| 4109. |
CleaR: Robust and Generalized Parameter-Efficient Fine-Tuning for Noisy Labels (arxiv.org) |
|
23 points by PaulHoule on Nov 29, 2024 | hide | past | pdf | discuss
|
| 4110. |
Llama Guard 3-1B-INT4: Compact and Efficient Safeguard for Human-AI Conversation (arxiv.org) |
|
2 points by sandwichsphinx on Nov 29, 2024 | hide | past | pdf | discuss
|
| More |