about
4231. An Empirical Model of Large-Batch Training (2018) (arxiv.org)
1 point by fzliu on Nov 7, 2024 | hide | past | pdf | discuss
4232. Won't Get Fooled Again: Answering Questions with False Premises (2023) (arxiv.org)
1 point by sandwichsphinx on Nov 7, 2024 | hide | past | pdf | discuss
4233. Physics-informed Shadowgraph Network: End-to-end Density Field Reconstruction (arxiv.org)
19 points by mixeden on Nov 7, 2024 | hide | past | pdf | 1 comment
4234. LLMs Orchestrating Structured Reasoning Achieve Kaggle Grandmaster Level (arxiv.org)
3 points by YeGoblynQueenne on Nov 7, 2024 | hide | past | pdf | discuss
4235. How Far Is Video Generation from World Model: A Physical Law Perspective (arxiv.org)
1 point by YeGoblynQueenne on Nov 7, 2024 | hide | past | pdf | discuss
4236. Evaluating the world model implicit in a generative model (arxiv.org)
159 points by dsubburam on Nov 7, 2024 | hide | past | pdf | 45 comments
4237. ChemBench: Evaluating LLMs Against Expert Chemists [New Results] (arxiv.org)
2 points by kjappelbaum on Nov 6, 2024 | hide | past | pdf | 1 comment
4238. Advancing Interpretability in Text Classification Through Prototype Learning (arxiv.org)
2 points by PaulHoule on Nov 6, 2024 | hide | past | pdf | discuss
4239. Ask, and it shall be given: Turing completeness of prompting (arxiv.org)
1 point by sandwichsphinx on Nov 6, 2024 | hide | past | pdf | discuss
4240. GenXD: Generating Any 3D and 4D Scenes (arxiv.org)
9 points by sandwichsphinx on Nov 5, 2024 | hide | past | pdf | discuss
4241. Leveraging Large Language Models for Advanced Multilingual Text-to-Speech (arxiv.org)
1 point by mnk47 on Nov 5, 2024 | hide | past | pdf | 1 comment
4242. PatternBoost: Constructions in Mathematics with a Little Help from AI (arxiv.org)
1 point by adas0693 on Nov 5, 2024 | hide | past | pdf | 2 comments
4243. TextLap: Customizing Language Models for Text-to-Layout Planning (arxiv.org)
8 points by PaulHoule on Nov 5, 2024 | hide | past | pdf | discuss
4244. Hunyuan-Large: An Open-Source Moe Model with 52B Activated Parameters (arxiv.org)
5 points by belter on Nov 5, 2024 | hide | past | pdf | discuss
4245. WebRL: Training LLM Web Agents via Self-Evolving Online Reinforcement Learning (arxiv.org)
23 points by theredsix on Nov 5, 2024 | hide | past | pdf | 1 comment
4246. A Survey on Generative Diffusion Model (arxiv.org)
3 points by lapnect on Nov 5, 2024 | hide | past | pdf | discuss
4247. Image-Goal Representations Atomic Control Units for Foundation Model Embodied AI (arxiv.org)
2 points by sandwichsphinx on Nov 5, 2024 | hide | past | pdf | discuss
4248. Can Large Language Models generalize analogy solving like people can? (arxiv.org)
3 points by belter on Nov 5, 2024 | hide | past | pdf | 1 comment
4249. Enhancing Long Context Performance in LLMs Through Inner Loop Query Mechanism (arxiv.org)
2 points by PaulHoule on Nov 5, 2024 | hide | past | pdf | discuss
4250. Exponential Separation Between Quantum and Quantum-Inspired Algorithms for ML (arxiv.org)
1 point by fuglede_ on Nov 5, 2024 | hide | past | pdf | discuss
4251. Accuracy-Performance Trade-Offs in LLM Quantization (arxiv.org)
1 point by belter on Nov 5, 2024 | hide | past | pdf | 1 comment
4252. Text Embedding Benchmark (2022) (arxiv.org)
2 points by todsacerdoti on Nov 5, 2024 | hide | past | pdf | discuss
4253. URAvatar: Universal Relightable Gaussian Codec Avatars (arxiv.org)
2 points by sandwichsphinx on Nov 4, 2024 | hide | past | pdf | discuss
4254. An End-to-End Model with Adaptive Filtering for Retrieval-Augmented Generation (arxiv.org)
2 points by foweltschmerz on Nov 4, 2024 | hide | past | pdf | discuss
4255. VibeCheck: Discover and Quantify Qualitative Differences in LLMs (arxiv.org)
1 point by PaulHoule on Nov 4, 2024 | hide | past | pdf | discuss
4256. Interpretable Online Log Analysis Using LLMs with Prompt Strategies (arxiv.org)
1 point by krunck on Nov 4, 2024 | hide | past | pdf | discuss
4257. MarsCode Agent: AI-Native Automated Bug Fixing (arxiv.org)
1 point by zshanhui on Nov 4, 2024 | hide | past | pdf | discuss
4258. Alice's Adventures in a Differentiable Wonderland (PDF) (arxiv.org)
1 point by yarapavan on Nov 4, 2024 | hide | past | pdf | discuss
4259. Personalization of Large Language Models: A Survey (arxiv.org)
1 point by sandwichsphinx on Nov 4, 2024 | hide | past | pdf | discuss
4260. The Persistence of Neural Collapse Despite Low-Rank Bias (arxiv.org)
1 point by Wheatman on Nov 4, 2024 | hide | past | pdf | 1 comment