about
3931. Proof of Thought: Neurosymbolic Program Synthesis for Interpretable Reasoning (arxiv.org)
4 points by barthelomew on Jan 4, 2025 | hide | past | pdf | 1 comment
3932. Cache-Augmented Generation (CAG) (arxiv.org)
1 point by asah on Jan 4, 2025 | hide | past | pdf | discuss
3933. In Defense of Smart Algorithms over Hardware Acceleration for Large-Scale AI (arxiv.org)
2 points by gone35 on Jan 4, 2025 | hide | past | pdf | discuss
3934. Unlocking the Potential of Large Language Models in Data-Scarce Contexts (arxiv.org)
1 point by PaulHoule on Jan 4, 2025 | hide | past | pdf | discuss
3935. European Space Agency Benchmark for Anomaly Detection in Satellite Telemetry (arxiv.org)
3 points by sarusso on Jan 3, 2025 | hide | past | pdf | discuss
3936. A path to O1 open source (arxiv.org)
133 points by bchelli on Jan 3, 2025 | hide | past | pdf | 80 comments
3937. Algorithmic Language Models with Neurally Compiled Libraries (arxiv.org)
1 point by wseqyrku on Jan 3, 2025 | hide | past | pdf | discuss
3938. 2 OLMo 2 Furious (arxiv.org)
4 points by lavabender on Jan 3, 2025 | hide | past | pdf | 1 comment
3939. Medec: A Benchmark for Medical Error Detection and Correction in Clinical Notes (arxiv.org)
2 points by gone35 on Jan 3, 2025 | hide | past | pdf | discuss
3940. InvestorBench: A Benchmark for Financial Decision-Making Tasks with Agents (arxiv.org)
1 point by xianshou on Jan 3, 2025 | hide | past | pdf | discuss
3941. Meta: Memory Layers at Scale (arxiv.org)
4 points by georgehill on Jan 2, 2025 | hide | past | pdf | discuss
3942. TinyStories: How Small Can Language Models Be and Still Speak Coherent English? (2023) (arxiv.org)
218 points by tzury on Jan 2, 2025 | hide | past | pdf | 104 comments
3943. Reinforcement Learning for Multi-Intersection Traffic Signal Control (arxiv.org)
1 point by PaulHoule on Jan 2, 2025 | hide | past | pdf | discuss
3944. MVQ: Efficient DNN Compression and Acceleration with Masked Vector Quantization (arxiv.org)
2 points by PaulHoule on Jan 2, 2025 | hide | past | pdf | discuss
3945. Generative Modeling with Explicit Memory (arxiv.org)
2 points by PaulHoule on Jan 2, 2025 | hide | past | pdf | discuss
3946. The Overthinking of O1-Like LLMs (arxiv.org)
3 points by omarsar on Jan 2, 2025 | hide | past | pdf | 1 comment
3947. Why transformers are obviously good models of language (arxiv.org)
6 points by jxmorris12 on Jan 2, 2025 | hide | past | pdf | discuss
3948. Scaling of Search and Learning (arxiv.org)
2 points by jonbaer on Jan 2, 2025 | hide | past | pdf | discuss
3949. 1.58-Bit Flux (arxiv.org)
2 points by reynaldi on Jan 1, 2025 | hide | past | pdf | 1 comment
3950. DeepSeek-V2: A Strong, Economical, and Efficient MOE Language Model (arxiv.org)
3 points by sonabinu on Jan 1, 2025 | hide | past | pdf | discuss
3951. Identifying and Manipulating LLM Personality Traits via Activation Engineering (arxiv.org)
23 points by rntn on Dec 31, 2024 | hide | past | pdf | 9 comments
3952. Re-Bench: Evaluating ML agents against human ML experts (arxiv.org)
2 points by marojejian on Dec 31, 2024 | hide | past | pdf | 1 comment
3953. Unifying Generative and Dense Retrieval for Sequential Recommendation (arxiv.org)
4 points by ashvardanian on Dec 31, 2024 | hide | past | pdf | discuss
3954. Mulberry: Empowering MLLM with o1-like Reasoning (arxiv.org)
3 points by Anon84 on Dec 31, 2024 | hide | past | pdf | discuss
3955. Beyond Gradient Averaging in Parallel Optimization (arxiv.org)
96 points by shinryudbz on Dec 30, 2024 | hide | past | pdf | 41 comments
3956. Optimizing Fantasy Sports Team Selection with Deep Reinforcement Learning (arxiv.org)
1 point by lucaspauker on Dec 30, 2024 | hide | past | pdf | discuss
3957. AgreeMate: Teaching LLMs to Haggle (arxiv.org)
2 points by rntn on Dec 30, 2024 | hide | past | pdf | discuss
3958. The Unreasonable Effectiveness of Open Science in AI: A Replication Study (arxiv.org)
2 points by belter on Dec 30, 2024 | hide | past | pdf | 1 comment
3959. How Well Do LLMs Generate Code for Different Application Domains? (arxiv.org)
81 points by belter on Dec 30, 2024 | hide | past | pdf | 25 comments
3960. Syzygy: Dual Code-Test C to Rust Translation Using LLMs and Dynamic Analysis (arxiv.org)
7 points by belter on Dec 30, 2024 | hide | past | pdf | 2 comments