about
4381. Discovering Global Lyapunov functions using symbolic transformers (arxiv.org)
2 points by DarmokJalad1701 on Oct 17, 2024 | hide | past | pdf | 1 comment
4382. Sample what you can't compress; image auto-encoders wihtout GANs (arxiv.org)
19 points by vighneshb on Oct 17, 2024 | hide | past | pdf | 4 comments
4383. Thinking LLMs: General Instruction Following with Thought Generation (arxiv.org)
2 points by IdealeZahlen on Oct 16, 2024 | hide | past | pdf | discuss
4384. Model Swarms: Collaborative Search to Adapt LLM Experts via Swarm Intelligence (arxiv.org)
2 points by fzliu on Oct 16, 2024 | hide | past | pdf | discuss
4385. Language Models Encode Numbers Using Digit Representations in Base 10 (arxiv.org)
2 points by beanvessel on Oct 16, 2024 | hide | past | pdf | discuss
4386. A Hitchhiker's Guide to Scaling Law Estimation (arxiv.org)
1 point by belter on Oct 16, 2024 | hide | past | pdf | discuss
4387. Tackling the Abstraction and Reasoning Corpus with Vision Transformers (arxiv.org)
2 points by apsec112 on Oct 15, 2024 | hide | past | pdf | discuss
4388. Data-Prep-Kit: getting your data ready for LLM application development (arxiv.org)
1 point by PaulHoule on Oct 15, 2024 | hide | past | pdf | discuss
4389. Paged KV-Cache Compression with Variable Compression Rates per Attention Head (arxiv.org)
2 points by PaulHoule on Oct 15, 2024 | hide | past | pdf | discuss
4390. Running LLMs with 3.3M Context Tokens on a Single GPU (arxiv.org)
14 points by Van_Chopiszt on Oct 15, 2024 | hide | past | pdf | 1 comment
4391. Representation Matters (arxiv.org)
2 points by quxinxin on Oct 15, 2024 | hide | past | pdf | discuss
4392. Quantifying Political Neutrality in LLM-Generated News Summaries (arxiv.org)
1 point by Hard_Space on Oct 15, 2024 | hide | past | pdf | discuss
4393. Intelligence at the Edge of Chaos [evidence of intelligence in LLMs] (arxiv.org)
1 point by fallingknife on Oct 15, 2024 | hide | past | pdf | discuss
4394. Lie Algebra Canonicalization: Equivariant Neural Ops Under Arbitrary Lie Groups (arxiv.org)
1 point by lnyan on Oct 15, 2024 | hide | past | pdf | discuss
4395. Thinking LLMs: General Instruction Following with Thought Generation (arxiv.org)
2 points by mnk47 on Oct 15, 2024 | hide | past | pdf | 1 comment
4396. Cheating Automatic LLM Benchmarks (arxiv.org)
3 points by jneagu on Oct 15, 2024 | hide | past | pdf | discuss
4397. Agents Thinking Fast and Slow: A Talker-Reasoner Architecture (arxiv.org)
3 points by rntn on Oct 14, 2024 | hide | past | pdf | discuss
4398. AutoML-Agent: A Multi-Agent LLM Framework for Full-Pipeline AutoML (arxiv.org)
2 points by rntn on Oct 14, 2024 | hide | past | pdf | discuss
4399. Knowledge Graph Based Agent for Complex, Knowledge-Intensive QA in Medicine (arxiv.org)
3 points by rntn on Oct 14, 2024 | hide | past | pdf | discuss
4400. Meissonic, High-Resolution Text-to-Image Synthesis on consumer graphics cards (arxiv.org)
65 points by jinqueeny on Oct 14, 2024 | hide | past | pdf | 4 comments
4401. DeepSeek: Advancing theorem proving in LLMs through large-scale synthetic data (arxiv.org)
186 points by hhs on Oct 14, 2024 | hide | past | pdf | 54 comments
4402. The Dynamics of Social Conventions in LLM Populations (arxiv.org)
2 points by Anon84 on Oct 14, 2024 | hide | past | pdf | discuss
4403. What's the Magic Word? A Control Theory of LLM Prompting (arxiv.org)
1 point by mnk47 on Oct 14, 2024 | hide | past | pdf | discuss
4404. Learning Category Trees for ID-Based Recommendation: Differentiable VQ (arxiv.org)
4 points by RicoElectrico on Oct 14, 2024 | hide | past | pdf | discuss
4405. Capabilities and Limitations of Large Language Models for Cultural Commonsense (arxiv.org)
2 points by rntn on Oct 13, 2024 | hide | past | pdf | discuss
4406. D-Edit, an image editing framework supporting text, image, mask-based editing (arxiv.org)
1 point by jinqueeny on Oct 13, 2024 | hide | past | pdf | discuss
4407. Gödel Agent: A self-referential agent framework for recursive self-improvement (arxiv.org)
81 points by tkgally on Oct 13, 2024 | hide | past | pdf | 29 comments
4408. Machine learning and information theory concepts towards an AI Mathematician (arxiv.org)
109 points by marojejian on Oct 12, 2024 | hide | past | pdf | 19 comments
4409. DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines (arxiv.org)
2 points by ulrischa on Oct 12, 2024 | hide | past | pdf | discuss
4410. Fine-Tuning Vision Classifiers on a Budget (arxiv.org)
1 point by PaulHoule on Oct 11, 2024 | hide | past | pdf | discuss