| 3931. |
Proof of Thought: Neurosymbolic Program Synthesis for Interpretable Reasoning (arxiv.org) |
|
4 points by barthelomew on Jan 4, 2025 | hide | past | pdf | 1 comment
|
| 3932. |
Cache-Augmented Generation (CAG) (arxiv.org) |
|
1 point by asah on Jan 4, 2025 | hide | past | pdf | discuss
|
| 3933. |
In Defense of Smart Algorithms over Hardware Acceleration for Large-Scale AI (arxiv.org) |
|
2 points by gone35 on Jan 4, 2025 | hide | past | pdf | discuss
|
| 3934. |
Unlocking the Potential of Large Language Models in Data-Scarce Contexts (arxiv.org) |
|
1 point by PaulHoule on Jan 4, 2025 | hide | past | pdf | discuss
|
| 3935. |
European Space Agency Benchmark for Anomaly Detection in Satellite Telemetry (arxiv.org) |
|
3 points by sarusso on Jan 3, 2025 | hide | past | pdf | discuss
|
| 3936. |
A path to O1 open source (arxiv.org) |
|
133 points by bchelli on Jan 3, 2025 | hide | past | pdf | 80 comments
|
| 3937. |
Algorithmic Language Models with Neurally Compiled Libraries (arxiv.org) |
|
1 point by wseqyrku on Jan 3, 2025 | hide | past | pdf | discuss
|
| 3938. |
2 OLMo 2 Furious (arxiv.org) |
|
4 points by lavabender on Jan 3, 2025 | hide | past | pdf | 1 comment
|
| 3939. |
Medec: A Benchmark for Medical Error Detection and Correction in Clinical Notes (arxiv.org) |
|
2 points by gone35 on Jan 3, 2025 | hide | past | pdf | discuss
|
| 3940. |
InvestorBench: A Benchmark for Financial Decision-Making Tasks with Agents (arxiv.org) |
|
1 point by xianshou on Jan 3, 2025 | hide | past | pdf | discuss
|
| 3941. |
Meta: Memory Layers at Scale (arxiv.org) |
|
4 points by georgehill on Jan 2, 2025 | hide | past | pdf | discuss
|
| 3942. |
TinyStories: How Small Can Language Models Be and Still Speak Coherent English? (2023) (arxiv.org) |
|
218 points by tzury on Jan 2, 2025 | hide | past | pdf | 104 comments
|
| 3943. |
Reinforcement Learning for Multi-Intersection Traffic Signal Control (arxiv.org) |
|
1 point by PaulHoule on Jan 2, 2025 | hide | past | pdf | discuss
|
| 3944. |
MVQ: Efficient DNN Compression and Acceleration with Masked Vector Quantization (arxiv.org) |
|
2 points by PaulHoule on Jan 2, 2025 | hide | past | pdf | discuss
|
| 3945. |
Generative Modeling with Explicit Memory (arxiv.org) |
|
2 points by PaulHoule on Jan 2, 2025 | hide | past | pdf | discuss
|
| 3946. |
The Overthinking of O1-Like LLMs (arxiv.org) |
|
3 points by omarsar on Jan 2, 2025 | hide | past | pdf | 1 comment
|
| 3947. |
Why transformers are obviously good models of language (arxiv.org) |
|
6 points by jxmorris12 on Jan 2, 2025 | hide | past | pdf | discuss
|
| 3948. |
Scaling of Search and Learning (arxiv.org) |
|
2 points by jonbaer on Jan 2, 2025 | hide | past | pdf | discuss
|
| 3949. |
1.58-Bit Flux (arxiv.org) |
|
2 points by reynaldi on Jan 1, 2025 | hide | past | pdf | 1 comment
|
| 3950. |
DeepSeek-V2: A Strong, Economical, and Efficient MOE Language Model (arxiv.org) |
|
3 points by sonabinu on Jan 1, 2025 | hide | past | pdf | discuss
|
| 3951. |
Identifying and Manipulating LLM Personality Traits via Activation Engineering (arxiv.org) |
|
23 points by rntn on Dec 31, 2024 | hide | past | pdf | 9 comments
|
| 3952. |
Re-Bench: Evaluating ML agents against human ML experts (arxiv.org) |
|
2 points by marojejian on Dec 31, 2024 | hide | past | pdf | 1 comment
|
| 3953. |
Unifying Generative and Dense Retrieval for Sequential Recommendation (arxiv.org) |
|
4 points by ashvardanian on Dec 31, 2024 | hide | past | pdf | discuss
|
| 3954. |
Mulberry: Empowering MLLM with o1-like Reasoning (arxiv.org) |
|
3 points by Anon84 on Dec 31, 2024 | hide | past | pdf | discuss
|
| 3955. |
Beyond Gradient Averaging in Parallel Optimization (arxiv.org) |
|
96 points by shinryudbz on Dec 30, 2024 | hide | past | pdf | 41 comments
|
| 3956. |
Optimizing Fantasy Sports Team Selection with Deep Reinforcement Learning (arxiv.org) |
|
1 point by lucaspauker on Dec 30, 2024 | hide | past | pdf | discuss
|
| 3957. |
AgreeMate: Teaching LLMs to Haggle (arxiv.org) |
|
2 points by rntn on Dec 30, 2024 | hide | past | pdf | discuss
|
| 3958. |
The Unreasonable Effectiveness of Open Science in AI: A Replication Study (arxiv.org) |
|
2 points by belter on Dec 30, 2024 | hide | past | pdf | 1 comment
|
| 3959. |
How Well Do LLMs Generate Code for Different Application Domains? (arxiv.org) |
|
81 points by belter on Dec 30, 2024 | hide | past | pdf | 25 comments
|
| 3960. |
Syzygy: Dual Code-Test C to Rust Translation Using LLMs and Dynamic Analysis (arxiv.org) |
|
7 points by belter on Dec 30, 2024 | hide | past | pdf | 2 comments
|
| More |