about
4081. Effective Strategies for Mitigating Hallucinations in LLMs for Data Analytics (arxiv.org)
2 points by micrum on Dec 3, 2024 | hide | past | pdf | discuss
4082. AI Meets Antimatter: Unveiling Antihydrogen Annihilations (arxiv.org)
1 point by throwawayed1 on Dec 3, 2024 | hide | past | pdf | discuss
4083. LLMs as Method Actors: A Model for Prompt Engineering and Architecture (arxiv.org)
2 points by gronky_ on Dec 3, 2024 | hide | past | pdf | 1 comment
4084. Kolmogorov Complexity, and the Role of Inductive Biases in Machine Learning (arxiv.org)
11 points by rntn on Dec 3, 2024 | hide | past | pdf | 1 comment
4085. Beyond Language Models: Byte Models Are Digital World Simulators (arxiv.org)
4 points by bob1029 on Dec 3, 2024 | hide | past | pdf | discuss
4086. Auto-RAG (arxiv.org)
3 points by omarsar on Dec 3, 2024 | hide | past | pdf | discuss
4087. Star: Synthesis of Tailored Architectures (arxiv.org)
1 point by simonpure on Dec 3, 2024 | hide | past | pdf | discuss
4088. Neural simulation-based inference for parameter estimation in ATLAS (arxiv.org)
1 point by pizza on Dec 3, 2024 | hide | past | pdf | discuss
4089. Demo: Decoupled Momentum Optimization (arxiv.org)
1 point by simonpure on Dec 3, 2024 | hide | past | pdf | discuss
4090. Modeling AdaGrad, RMSProp, and Adam with Integro-Differential Equations (arxiv.org)
2 points by PaulHoule on Dec 2, 2024 | hide | past | pdf | discuss
4091. Generative Agent Simulations of 1k People (arxiv.org)
2 points by wslh on Dec 2, 2024 | hide | past | pdf | discuss
4092. Large Language Model-Brained GUI Agents: A Survey (arxiv.org)
6 points by pongogogo on Dec 2, 2024 | hide | past | pdf | discuss
4093. Demo: Decoupled Momentum Optimization – training neural networks in parallel (arxiv.org)
3 points by sergiotapia on Dec 2, 2024 | hide | past | pdf | discuss
4094. Reverse Thinking Makes LLMs Stronger Reasoners (arxiv.org)
4 points by omarsar on Dec 2, 2024 | hide | past | pdf | discuss
4095. Do LLMs Perform Latent Multi-Hop Reasoning Without Exploiting Shortcuts? (arxiv.org)
1 point by fzliu on Dec 2, 2024 | hide | past | pdf | discuss
4096. Procedural knowledge in pretraining drives reasoning in large language models (arxiv.org)
248 points by reqo on Dec 1, 2024 | hide | past | pdf | 101 comments
4097. Challenges and Applications of Large Language Models (arxiv.org)
1 point by t55 on Dec 1, 2024 | hide | past | pdf | discuss
4098. DynaSaur: Large Language Agents Beyond Predefined Actions (arxiv.org)
128 points by surprisetalk on Dec 1, 2024 | hide | past | pdf | 31 comments
4099. The Curse of Recursion: Training on generated data makes models forget (2023) (arxiv.org)
122 points by surprisetalk on Dec 1, 2024 | hide | past | pdf | 107 comments
4100. Streaming Deep Reinforcement Learning Works (arxiv.org)
2 points by surprisetalk on Dec 1, 2024 | hide | past | pdf | discuss
4101. DrugAgent: AI-Aided Drug Discovery Programming Through LLM Multi-Agent Collab (arxiv.org)
2 points by surprisetalk on Dec 1, 2024 | hide | past | pdf | discuss
4102. Large Language Models as Markov Chains (arxiv.org)
75 points by Anon84 on Nov 30, 2024 | hide | past | pdf | 54 comments
4103. SnapMem: Snapshot-Based 3D Scene Memory for Embodied Exploration and Reasoning (arxiv.org)
2 points by sandwichsphinx on Nov 30, 2024 | hide | past | pdf | discuss
4104. Generative Agent Simulations of 1k People (arxiv.org)
1 point by doener on Nov 30, 2024 | hide | past | pdf | discuss
4105. Stable, Fast, Automatic Learning Algorithm for Predictive Coding Networks [pdf] (arxiv.org)
3 points by kelseyfrog on Nov 29, 2024 | hide | past | pdf | discuss
4106. Super-Resolution with Low-Bit Quantization (arxiv.org)
2 points by RobinHirst11 on Nov 29, 2024 | hide | past | pdf | discuss
4107. DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving (arxiv.org)
2 points by RobinHirst11 on Nov 29, 2024 | hide | past | pdf | discuss
4108. Enhancing JEPAs with Spatial Conditioning: Efficient Representation Learning (arxiv.org)
3 points by sandwichsphinx on Nov 29, 2024 | hide | past | pdf | discuss
4109. CleaR: Robust and Generalized Parameter-Efficient Fine-Tuning for Noisy Labels (arxiv.org)
23 points by PaulHoule on Nov 29, 2024 | hide | past | pdf | discuss
4110. Llama Guard 3-1B-INT4: Compact and Efficient Safeguard for Human-AI Conversation (arxiv.org)
2 points by sandwichsphinx on Nov 29, 2024 | hide | past | pdf | discuss