about
3781. Over-Tokenized Transformer: Vocabulary Is Worth Scaling [pdf] (arxiv.org)
2 points by nickpsecurity on Jan 30, 2025 | hide | past | pdf | discuss
3782. Reasoning-Enhanced Self-Training for Long-Form Personalized Text Generation (arxiv.org)
1 point by PaulHoule on Jan 30, 2025 | hide | past | pdf | discuss
3783. Open Problems in Machine Unlearning for AI Safety (arxiv.org)
1 point by RTFPaper on Jan 30, 2025 | hide | past | pdf | discuss
3784. Mind the Value-Action Gap: Do LLMs Act in Alignment with Their Values? (arxiv.org)
1 point by batterylake on Jan 30, 2025 | hide | past | pdf | 1 comment
3785. Parametric Retrieval Augmented Generation (arxiv.org)
3 points by batterylake on Jan 30, 2025 | hide | past | pdf | 1 comment
3786. Is Your Image a Good Storyteller? (arxiv.org)
4 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
3787. Hardware Implemented Accelerator Design in ReRAM Analog Computing Without ADCs (arxiv.org)
2 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
3788. Enhancing In-Context Learning with Differentiated and Reweighting Objectives (arxiv.org)
2 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
3789. Large Language Model Training Using FP4 Quantization (arxiv.org)
2 points by t55 on Jan 29, 2025 | hide | past | pdf | discuss
3790. Supervised Fine-Tuning Memorizes, RL Generalizes (arxiv.org)
1 point by t55 on Jan 29, 2025 | hide | past | pdf | discuss
3791. Small Language Models for Document Layout Generation and Classification (arxiv.org)
2 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
3792. People who use ChatGPT for writing are robust detectors of AI-generated text (arxiv.org)
2 points by eatonphil on Jan 29, 2025 | hide | past | pdf | discuss
3793. Shrink the longest: improving latent space isotropy with symplicial geometry (arxiv.org)
2 points by PaulHoule on Jan 29, 2025 | hide | past | pdf | discuss
3794. Open Problems in Mechanistic Interpretability (arxiv.org)
2 points by RTFPaper on Jan 29, 2025 | hide | past | pdf | discuss
3795. Matrix Calculus (For Machine Learning and Beyond) (arxiv.org)
1 point by sebg on Jan 29, 2025 | hide | past | pdf | discuss
3796. Self-Replicating AI (lab experiments) (arxiv.org)
3 points by slow_typist on Jan 29, 2025 | hide | past | pdf | 2 comments
3797. Auto-Differentiating Any LLM Workflow: A Farewell to Manual Prompting (arxiv.org)
137 points by meame2010 on Jan 29, 2025 | hide | past | pdf | 32 comments
3798. 3D scene reconstruction in adverse weather conditions via Gaussian splatting (arxiv.org)
53 points by PaulHoule on Jan 28, 2025 | hide | past | pdf | 14 comments
3799. Evolution and the Knightian Blindspot of Machine Learning (arxiv.org)
1 point by jarrattp on Jan 28, 2025 | hide | past | pdf | discuss
3800. ComMer: A Framework for Compressing and Merging User Data for Personalization (arxiv.org)
1 point by PaulHoule on Jan 28, 2025 | hide | past | pdf | discuss
3801. Towards Large Reasoning Models: A Survey of Reinforced Reasoning with LLMs (arxiv.org)
1 point by keepit on Jan 28, 2025 | hide | past | pdf | discuss
3802. DeepSeekMath: Pushing Limits of Mathematical Reasoning in Open Language Models (arxiv.org)
1 point by tosh on Jan 28, 2025 | hide | past | pdf | 1 comment
3803. DeepSeek-V3 Technical Report (arxiv.org)
3 points by doener on Jan 28, 2025 | hide | past | pdf | discuss
3804. Grounding Text-to-Image Models for Controlled High-Quality Image Generation (arxiv.org)
1 point by GenAI-research on Jan 28, 2025 | hide | past | pdf | 1 comment
3805. RL and Transformer = a General-Purpose Problem Solver (arxiv.org)
3 points by keepit on Jan 27, 2025 | hide | past | pdf | discuss
3806. Granger Causality Detection with Kolmogorov-Arnold Networks (arxiv.org)
1 point by Anon84 on Jan 27, 2025 | hide | past | pdf | discuss
3807. Autonomy-of-Experts Models (ArXiv) (arxiv.org)
2 points by erdaltoprak on Jan 27, 2025 | hide | past | pdf | discuss
3808. The Matrix Calculus You Need for Deep Learning (2018) (arxiv.org)
3 points by mpweiher on Jan 26, 2025 | hide | past | pdf | discuss
3809. DeepSeekMath (arxiv.org)
6 points by sonabinu on Jan 26, 2025 | hide | past | pdf | discuss
3810. Kimi K1.5: Scaling Reinforcement Learning with LLMs (arxiv.org)
4 points by anjneymidha on Jan 26, 2025 | hide | past | pdf | discuss