about
3961. Transformers Meet Neural Algorithmic Reasoners (arxiv.org)
1 point by rntn on Dec 30, 2024 | hide | past | pdf | discuss
3962. Efficient Generative Modeling with Residual Vector Quantization-Based Tokens (arxiv.org)
1 point by PaulHoule on Dec 29, 2024 | hide | past | pdf | discuss
3963. Modeling health trajectories with a transformer-based deep learning model (arxiv.org)
1 point by PaulHoule on Dec 29, 2024 | hide | past | pdf | discuss
3964. ReAct: Synergizing Reasoning and Acting in Language Models (arxiv.org)
3 points by sebg on Dec 28, 2024 | hide | past | pdf | discuss
3965. Empirical Study of Test Generation with LLM's (arxiv.org)
40 points by nickpsecurity on Dec 28, 2024 | hide | past | pdf | 36 comments
3966. A Many Objective Problem Where Crossover Is Provably Indispensable (arxiv.org)
1 point by crete on Dec 28, 2024 | hide | past | pdf | discuss
3967. Frontier AI systems have surpassed the self-replicating red line (arxiv.org)
1 point by blacktulip on Dec 28, 2024 | hide | past | pdf | 1 comment
3968. Measuring and Understanding LLM Identity Confusion (arxiv.org)
21 points by _____k on Dec 28, 2024 | hide | past | pdf | 1 comment
3969. Armada: Augmented Reality for Robot Manipulation and Robot-Free Data Acquisition (arxiv.org)
6 points by sandwichsphinx on Dec 28, 2024 | hide | past | pdf | discuss
3970. Explaining Large Language Models Decisions Using Shapley Values (arxiv.org)
89 points by veryluckyxyz on Dec 28, 2024 | hide | past | pdf | 19 comments
3971. The Famine of Forte: Few Search Problems Greatly Favor Your Algorithm (2016) (arxiv.org)
1 point by pizza on Dec 28, 2024 | hide | past | pdf | discuss
3972. Rethinking Uncertainty Estimation in Natural Language Generation (arxiv.org)
1 point by sroussey on Dec 27, 2024 | hide | past | pdf | discuss
3973. Language Model as Visual Explainer (arxiv.org)
5 points by PaulHoule on Dec 27, 2024 | hide | past | pdf | discuss
3974. DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts LMs (arxiv.org)
2 points by doener on Dec 27, 2024 | hide | past | pdf | discuss
3975. Towards Unsupervised Learning Scheme for Efficiently Solving Parameterized MIPS (arxiv.org)
1 point by crete on Dec 27, 2024 | hide | past | pdf | discuss
3976. Are Language Models Actually Useful for Time Series Forecasting? (arxiv.org)
7 points by cl42 on Dec 27, 2024 | hide | past | pdf | 7 comments
3977. When Every Token Counts: Optimal Segmentation for Low-Resource Language Models (arxiv.org)
1 point by PaulHoule on Dec 27, 2024 | hide | past | pdf | discuss
3978. Improving feature interactions at Pinterest under industry constraints (arxiv.org)
1 point by PaulHoule on Dec 26, 2024 | hide | past | pdf | discuss
3979. Deliberation in Latent Space via Differentiable Cache Augmentation [pdf] (arxiv.org)
9 points by lawrenceyan on Dec 26, 2024 | hide | past | pdf | discuss
3980. Monolith: Real Time Recommendation System with Collisionless Embedding Table (arxiv.org)
1 point by ianrahman on Dec 26, 2024 | hide | past | pdf | discuss
3981. Phi-4 Technical Report (arxiv.org)
2 points by veryluckyxyz on Dec 25, 2024 | hide | past | pdf | discuss
3982. AlphaPruning: Using Heavy-Tailed Self Regularization for Improved LLM Pruning (arxiv.org)
1 point by selimthegrim on Dec 25, 2024 | hide | past | pdf | discuss
3983. Brain-to-Text Benchmark '24 (arxiv.org)
2 points by UCdallasGA on Dec 24, 2024 | hide | past | pdf | discuss
3984. LearnLM (arxiv.org)
2 points by UCdallasGA on Dec 24, 2024 | hide | past | pdf | discuss
3985. Arbitrary-Steps Image Super-Resolution via Diffusion Inversion (arxiv.org)
3 points by ChrisArchitect on Dec 23, 2024 | hide | past | pdf | discuss
3986. A High-Performance In-Browser LLM Inference Engine (arxiv.org)
1 point by UCdallasGA on Dec 23, 2024 | hide | past | pdf | discuss
3987. Adversarial policies beat superhuman Go AIs (2023) (arxiv.org)
306 points by amichail on Dec 23, 2024 | hide | past | pdf | 139 comments
3988. Offline Reinforcement Learning for LLM Multi-Step Reasoning (arxiv.org)
111 points by belter on Dec 23, 2024 | hide | past | pdf | 9 comments
3989. SWE-Bench+: Enhanced Coding Benchmark for LLMs (arxiv.org)
3 points by zeroonetwothree on Dec 23, 2024 | hide | past | pdf | discuss
3990. Selective Adapter Freezing for Memory-Efficient Fine-Tuning of Language Models (arxiv.org)
2 points by PaulHoule on Dec 22, 2024 | hide | past | pdf | discuss