about
4561. Awes, Laws, and Flaws from Today's LLM Research (arxiv.org)
2 points by PaulHoule on Sep 12, 2024 | hide | past | pdf | discuss
4562. Can LLMs Generate Novel Research Ideas? (arxiv.org)
50 points by kmdupree on Sep 12, 2024 | hide | past | pdf | 81 comments
4563. Hecaton: Training and Finetuning LLMs with Scalable Chiplet Systems (arxiv.org)
2 points by PaulHoule on Sep 12, 2024 | hide | past | pdf | discuss
4564. LLMs Will Always Hallucinate, and We Need to Live with This (arxiv.org)
4 points by kvee on Sep 12, 2024 | hide | past | pdf | discuss
4565. Denoising: Building-Block for Imaging, Inverse Problems, and Machine Learning (arxiv.org)
2 points by sebg on Sep 12, 2024 | hide | past | pdf | discuss
4566. A Simple Convergence Proof of Adam and Adagrad (arxiv.org)
1 point by sebg on Sep 12, 2024 | hide | past | pdf | discuss
4567. Towards Large Language Models as Copilots for Theorem Proving in Lean (arxiv.org)
3 points by yeesian on Sep 12, 2024 | hide | past | pdf | discuss
4568. A Review of Graph Neural Networks in Epidemic Modeling (arxiv.org)
1 point by Anon84 on Sep 11, 2024 | hide | past | pdf | discuss
4569. Can LLMs Generate Novel Research Ideas? (arxiv.org)
2 points by vinni2 on Sep 11, 2024 | hide | past | pdf | discuss
4570. MemLong: Memory-Augmented Retrieval for Long Text Modeling (arxiv.org)
1 point by PaulHoule on Sep 11, 2024 | hide | past | pdf | discuss
4571. LLMs Outpace Humans in Novel Idea Generation (arxiv.org)
5 points by RafelMri on Sep 11, 2024 | hide | past | pdf | discuss
4572. Tutorial on diffusion models for imaging and vision (arxiv.org)
221 points by Anon84 on Sep 10, 2024 | hide | past | pdf | 18 comments
4573. Leveraging Large Language Models for Solving Rare MIP Challenges (arxiv.org)
2 points by _artifice_ on Sep 10, 2024 | hide | past | pdf | discuss
4574. HLogformer: A Hierarchical Transformer for Representing Log Data (arxiv.org)
4 points by PaulHoule on Sep 10, 2024 | hide | past | pdf | discuss
4575. ChartEye: A Deep Learning Framework for Chart Information Extraction (arxiv.org)
38 points by PaulHoule on Sep 10, 2024 | hide | past | pdf | discuss
4576. Deductive Verification for Chain-of-Thought Reasoning in LLMs (arxiv.org)
80 points by smooke on Sep 10, 2024 | hide | past | pdf | 20 comments
4577. State and Action Factorization in Power Grids (arxiv.org)
2 points by DevScout on Sep 10, 2024 | hide | past | pdf | discuss
4578. Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study (arxiv.org)
3 points by hardmaru on Sep 10, 2024 | hide | past | pdf | discuss
4579. The AdEMAMix Optimizer: Better, Faster, Older (arxiv.org)
2 points by andy12_ on Sep 10, 2024 | hide | past | pdf | discuss
4580. MMR: Evaluating Reading Ability of Large Multimodal Models (arxiv.org)
3 points by PaulHoule on Sep 9, 2024 | hide | past | pdf | discuss
4581. LLM-Assisted Labeling Function Generation for Semantic Type Detection (arxiv.org)
1 point by PaulHoule on Sep 9, 2024 | hide | past | pdf | discuss
4582. Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study (arxiv.org)
4 points by handfuloflight on Sep 9, 2024 | hide | past | pdf | discuss
4583. Transfusion: Predict the next token and diffuse images with one multimodal model (arxiv.org)
122 points by fzliu on Sep 9, 2024 | hide | past | pdf | 10 comments
4584. ReMamba: Equip Mamba with Effective Long-Sequence Modeling (arxiv.org)
40 points by PaulHoule on Sep 9, 2024 | hide | past | pdf | 1 comment
4585. ServerlessLLM: Low-Latency Serverless Inference for Large Language Models (arxiv.org)
2 points by geuds on Sep 9, 2024 | hide | past | pdf | discuss
4586. Diffusion models are real-time game engines (arxiv.org)
1 point by sargstuff on Sep 9, 2024 | hide | past | pdf | discuss
4587. A Benchmark for Learned Cardinality Estimation in Relational Databases (arxiv.org)
2 points by PaulHoule on Sep 7, 2024 | hide | past | pdf | discuss
4588. Flux That Plays Music (arxiv.org)
2 points by johnsutor on Sep 7, 2024 | hide | past | pdf | discuss
4589. A Comparative Study on Large Language Models for Log Parsing (arxiv.org)
1 point by belter on Sep 7, 2024 | hide | past | pdf | discuss
4590. Memory Is All You Need: An Overview of Compute-in-Memory Architectures for LLMs (arxiv.org)
4 points by fulafel on Sep 7, 2024 | hide | past | pdf | discuss