| 3781. |
Over-Tokenized Transformer: Vocabulary Is Worth Scaling [pdf] (arxiv.org) |
|
2 points by nickpsecurity on Jan 30, 2025 | hide | past | pdf | discuss
|
| 3782. |
Reasoning-Enhanced Self-Training for Long-Form Personalized Text Generation (arxiv.org) |
|
1 point by PaulHoule on Jan 30, 2025 | hide | past | pdf | discuss
|
| 3783. |
Open Problems in Machine Unlearning for AI Safety (arxiv.org) |
|
1 point by RTFPaper on Jan 30, 2025 | hide | past | pdf | discuss
|
| 3784. |
Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values? (arxiv.org) |
|
1 point by batterylake on Jan 30, 2025 | hide | past | pdf | 1 comment
|
| 3785. |
Parametric Retrieval Augmented Generation (arxiv.org) |
|
3 points by batterylake on Jan 30, 2025 | hide | past | pdf | 1 comment
|
| 3786. |
Is Your Image a Good Storyteller? (arxiv.org) |
|
4 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3787. |
Hardware Implemented Accelerator Design in ReRAM Analog Computing Without ADCs (arxiv.org) |
|
2 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3788. |
Enhancing In-Context Learning with Differentiated and Reweighting Objectives (arxiv.org) |
|
2 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3789. |
Large Language Model Training Using FP4 Quantization (arxiv.org) |
|
2 points by t55 on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3790. |
Supervised Fine-Tuning Memorizes, RL Generalizes (arxiv.org) |
|
1 point by t55 on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3791. |
Small Language Models for Document Layout Generation and Classification (arxiv.org) |
|
2 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3792. |
People who use ChatGPT for writing are robust detectors of AI-generated text (arxiv.org) |
|
2 points by eatonphil on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3793. |
Shrink the longest: improving latent space isotropy with symplicial geometry (arxiv.org) |
|
2 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3794. |
Open Problems in Mechanistic Interpretability (arxiv.org) |
|
2 points by RTFPaper on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3795. |
Matrix Calculus (For Machine Learning and Beyond) (arxiv.org) |
|
1 point by sebg on Jan 29, 2025 | hide | past | pdf | discuss
|
| 3796. |
Self-Replicating AI (lab experiments) (arxiv.org) |
|
3 points by slow_typist on Jan 29, 2025 | hide | past | pdf | 2 comments
|
| 3797. |
Auto-Differentiating Any LLM Workflow: A Farewell to Manual Prompting (arxiv.org) |
|
137 points by meame2010 on Jan 29, 2025 | hide | past | pdf | 32 comments
|
| 3798. |
3D scene reconstruction in adverse weather conditions via Gaussian splatting (arxiv.org) |
|
53 points by PaulHoule on Jan 28, 2025 | hide | past | pdf | 14 comments
|
| 3799. |
Evolution and the Knightian Blindspot of Machine Learning (arxiv.org) |
|
1 point by jarrattp on Jan 28, 2025 | hide | past | pdf | discuss
|
| 3800. |
ComMer: A Framework for Compressing and Merging User Data for Personalization (arxiv.org) |
|
1 point by PaulHoule on Jan 28, 2025 | hide | past | pdf | discuss
|
| 3801. |
Towards Large Reasoning Models: A Survey of Reinforced Reasoning with LLMs (arxiv.org) |
|
1 point by keepit on Jan 28, 2025 | hide | past | pdf | discuss
|
| 3802. |
DeepSeekMath: Pushing Limits of Mathematical Reasoning in Open Language Models (arxiv.org) |
|
1 point by tosh on Jan 28, 2025 | hide | past | pdf | 1 comment
|
| 3803. |
DeepSeek-V3 Technical Report (arxiv.org) |
|
3 points by doener on Jan 28, 2025 | hide | past | pdf | discuss
|
| 3804. |
Grounding Text-to-Image Models for Controlled High-Quality Image Generation (arxiv.org) |
|
1 point by GenAI-research on Jan 28, 2025 | hide | past | pdf | 1 comment
|
| 3805. |
RL and Transformer = a General-Purpose Problem Solver (arxiv.org) |
|
3 points by keepit on Jan 27, 2025 | hide | past | pdf | discuss
|
| 3806. |
Granger Causality Detection with Kolmogorov-Arnold Networks (arxiv.org) |
|
1 point by Anon84 on Jan 27, 2025 | hide | past | pdf | discuss
|
| 3807. |
Autonomy-of-Experts Models (ArXiv) (arxiv.org) |
|
2 points by erdaltoprak on Jan 27, 2025 | hide | past | pdf | discuss
|
| 3808. |
The Matrix Calculus You Need for Deep Learning (2018) (arxiv.org) |
|
3 points by mpweiher on Jan 26, 2025 | hide | past | pdf | discuss
|
| 3809. |
DeepSeekMath (arxiv.org) |
|
6 points by sonabinu on Jan 26, 2025 | hide | past | pdf | discuss
|
| 3810. |
Kimi K1.5: Scaling Reinforcement Learning with LLMs (arxiv.org) |
|
4 points by anjneymidha on Jan 26, 2025 | hide | past | pdf | discuss
|
| More |