about
4051. Arctic-Embed 2.0: Multilingual Retrieval Without Compromise (arxiv.org)
2 points by fzliu on Dec 10, 2024 | hide | past | pdf | discuss
4052. Can OpenAI o1 outperform humans in higher-order cognitive thinking? [pdf] (arxiv.org)
2 points by bikenaga on Dec 10, 2024 | hide | past | pdf | discuss
4053. Training LLMs to Reason in a Continuous Latent Space (arxiv.org)
283 points by omarsar on Dec 10, 2024 | hide | past | pdf | 114 comments
4054. Flow Matching Guide and Code (arxiv.org)
3 points by lnyan on Dec 10, 2024 | hide | past | pdf | discuss
4055. PyPIM: Integrating Digital Processing-in-Memory from Microarchitectural Design (arxiv.org)
2 points by pauloxnet on Dec 10, 2024 | hide | past | pdf | 1 comment
4056. LightLLM: A Versatile Large Language Model for Predictive Light Sensing (arxiv.org)
1 point by PaulHoule on Dec 9, 2024 | hide | past | pdf | discuss
4057. GhostRNN: Reducing State Redundancy in RNN with Cheap Operations (arxiv.org)
2 points by PaulHoule on Dec 9, 2024 | hide | past | pdf | discuss
4058. Semantic Retrieval at Walmart (arxiv.org)
2 points by sonabinu on Dec 9, 2024 | hide | past | pdf | 1 comment
4059. Transformers Struggle to Learn to Search (arxiv.org)
2 points by omarsar on Dec 9, 2024 | hide | past | pdf | discuss
4060. The Llama 3 Herd of Models (arxiv.org)
1 point by cloudsql on Dec 9, 2024 | hide | past | pdf | discuss
4061. Reinforcement Learning: An Overview (arxiv.org)
6 points by killme2008 on Dec 9, 2024 | hide | past | pdf | 1 comment
4062. RoboHanger: Learning Generalizable Robotic Hanger Insertion for Diverse Garments (arxiv.org)
1 point by sandwichsphinx on Dec 9, 2024 | hide | past | pdf | discuss
4063. RL, But Don't Do Anything I Wouldn't Do (arxiv.org)
1 point by optimalsolver on Dec 8, 2024 | hide | past | pdf | 2 comments
4064. Vulnerability of LLM Benchmarks: Do They Accurately Reflect True LLM Performance (arxiv.org)
2 points by sandwichsphinx on Dec 8, 2024 | hide | past | pdf | discuss
4065. The Alignment Problem from a Deep Learning Perspective (arxiv.org)
3 points by occamschainsaw on Dec 8, 2024 | hide | past | pdf | discuss
4066. Amplifying human performance in combinatorial competitive programming (arxiv.org)
2 points by bmc7505 on Dec 7, 2024 | hide | past | pdf | discuss
4067. Reverse Thinking Makes LLMs Stronger Reasoners (arxiv.org)
2 points by Anon84 on Dec 7, 2024 | hide | past | pdf | discuss
4068. Gradient Routing: Masking Gradients to Localize Computation in Neural Networks (arxiv.org)
2 points by Eugeleo on Dec 7, 2024 | hide | past | pdf | discuss
4069. Enhancing Mathematical Reasoning in LLMs with Background Operators (arxiv.org)
1 point by triska on Dec 7, 2024 | hide | past | pdf | discuss
4070. Densing Law of LLMs (arxiv.org)
1 point by cyp0633 on Dec 7, 2024 | hide | past | pdf | discuss
4071. Mapping the Podcast Ecosystem with the Structured Podcast Research Corpus (arxiv.org)
2 points by avyfain on Dec 6, 2024 | hide | past | pdf | discuss
4072. Intriguing Properties of Robust Classification (arxiv.org)
1 point by lamename on Dec 6, 2024 | hide | past | pdf | discuss
4073. FlashAttention on a Napkin:A Diagrammatic Approach to Deep Learning IO-Awareness (arxiv.org)
5 points by mpweiher on Dec 6, 2024 | hide | past | pdf | discuss
4074. Compressing Large Language Models Using Low Rank and Low Precision Decomposition (arxiv.org)
2 points by westurner on Dec 5, 2024 | hide | past | pdf | 1 comment
4075. PowerGraph: A power grid benchmark dataset for graph neural networks (arxiv.org)
2 points by sandwichsphinx on Dec 5, 2024 | hide | past | pdf | discuss
4076. Distinguishing Ignorance from Error in LLM Hallucinations (arxiv.org)
2 points by micrum on Dec 4, 2024 | hide | past | pdf | discuss
4077. Generative Agent Simulations of 1k People (arxiv.org)
1 point by amrrs on Dec 4, 2024 | hide | past | pdf | discuss
4078. The Surprising Effectiveness of Test-Time Training for Abstract Reasoning (arxiv.org)
1 point by surprisetalk on Dec 4, 2024 | hide | past | pdf | discuss
4079. OpenHumanVid: Large-Scale High-Quality Dataset Enhancing Human-Centric Video Gen (arxiv.org)
2 points by sandwichsphinx on Dec 4, 2024 | hide | past | pdf | discuss
4080. Effective Strategies for Mitigating Hallucinations in LLMs for Data Analytics (arxiv.org)
2 points by micrum on Dec 3, 2024 | hide | past | pdf | discuss