about
4861. Review of Mechanistic Interpretability for Transformer-Based Language Models (arxiv.org)
2 points by belter on Jul 7, 2024 | hide | past | pdf | discuss
4862. Reasoning in Large Language Models: A Geometric Perspective (arxiv.org)
214 points by belter on Jul 7, 2024 | hide | past | pdf | 171 comments
4863. An Efficient NAS-Based Approach for Handling Imbalanced Datasets (arxiv.org)
1 point by PaulHoule on Jul 7, 2024 | hide | past | pdf | discuss
4864. MobileLLM: Optimizing Sub-Billion Parameter Language Models for On-Device Use (arxiv.org)
3 points by tosh on Jul 7, 2024 | hide | past | pdf | discuss
4865. T-FREE Tokenizer-Free LLMs via Sparse Representation Memory-Efficient Embeddings (arxiv.org)
3 points by Bluestein on Jul 6, 2024 | hide | past | pdf | discuss
4866. Data curation via joint example selection (arxiv.org)
3 points by jonbaer on Jul 6, 2024 | hide | past | pdf | discuss
4867. Can LLMs Generate Visualizations with Dataless Prompts? (arxiv.org)
3 points by PaulHoule on Jul 6, 2024 | hide | past | pdf | discuss
4868. Verif.ai: Open-Source Generative Q&A System with Referenced Answers (arxiv.org)
2 points by nikolamilosevic on Jul 6, 2024 | hide | past | pdf | discuss
4869. ConvNet vs Transformer, Supervised vs CLIP (arxiv.org)
2 points by fzliu on Jul 6, 2024 | hide | past | pdf | discuss
4870. Spurious Reconstruction of Visual Perception from Brain Activity (arxiv.org)
1 point by DanielleMolloy on Jul 5, 2024 | hide | past | pdf | 1 comment
4871. LLM Agents can Autonomously Exploit One-day Vulnerabili-ties [pdf] (arxiv.org)
4 points by gsky on Jul 5, 2024 | hide | past | pdf | 1 comment
4872. Improve Mathematical Logic in Language Models by Automated Process Supervision (arxiv.org)
2 points by bryanrasmussen on Jul 4, 2024 | hide | past | pdf | 1 comment
4873. LLMMatDesign – Gen AI for Materials (arxiv.org)
4 points by socratic1 on Jul 4, 2024 | hide | past | pdf | discuss
4874. Radiology Report Generation and Evaluation with Layman's Terms (arxiv.org)
1 point by PaulHoule on Jul 4, 2024 | hide | past | pdf | discuss
4875. GraphReader: Building Graph-Based Agent to Enhance Long-Context Abilities of LLM (arxiv.org)
2 points by theptip on Jul 4, 2024 | hide | past | pdf | discuss
4876. Mental Modeling of Reinforcement Learning Agents by Language Models (arxiv.org)
7 points by piecerough on Jul 3, 2024 | hide | past | pdf | discuss
4877. Reducing the Memory Footprint of 3D Gaussian Splatting (arxiv.org)
2 points by PaulHoule on Jul 3, 2024 | hide | past | pdf | discuss
4878. ColPali: Efficient Document Retrieval with Vision Language Models (arxiv.org)
2 points by alphabetting on Jul 2, 2024 | hide | past | pdf | discuss
4879. UniGen: Unified Modeling for Generating Autonomous Driving Scenarios (arxiv.org)
1 point by ra7 on Jul 2, 2024 | hide | past | pdf | discuss
4880. AI Agents That Matter (arxiv.org)
4 points by randomwalker on Jul 2, 2024 | hide | past | pdf | discuss
4881. MoonshotAI unveils Kimi's large-scale LLM serving architecture (arxiv.org)
18 points by slothfulhamster on Jul 2, 2024 | hide | past | pdf | 1 comment
4882. Generalist Lightweight Model for Various Information Extraction Tasks (arxiv.org)
1 point by PaulHoule on Jul 2, 2024 | hide | past | pdf | discuss
4883. xLSTM-UNet can be an Effective 2D and 3D Medical Image Segmentation Backbone (arxiv.org)
2 points by tosh on Jul 2, 2024 | hide | past | pdf | discuss
4884. Large language models have developed a higher-order theory of mind (arxiv.org)
17 points by some-unique-x on Jul 2, 2024 | hide | past | pdf | 4 comments
4885. Segment Any Text (arxiv.org)
1 point by jasondavies on Jul 2, 2024 | hide | past | pdf | discuss
4886. FTTN: Feature-Targeted Testing for Numerical Properties of Matrix Accelerators (arxiv.org)
2 points by matt_d on Jul 2, 2024 | hide | past | pdf | discuss
4887. The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse (arxiv.org)
1 point by ijk on Jul 1, 2024 | hide | past | pdf | discuss
4888. What is the best model? Application-driven Evaluation for Large Language Models (arxiv.org)
1 point by PaulHoule on Jul 1, 2024 | hide | past | pdf | discuss
4889. The Remarkable Robustness of LLMs: Stages of Inference? (arxiv.org)
2 points by belter on Jun 30, 2024 | hide | past | pdf | discuss
4890. Newswire: A large-scale structured database of a century of historical news (arxiv.org)
165 points by h2odragon on Jun 30, 2024 | hide | past | pdf | 39 comments