about
4951. An Image Is Worth 32 Tokens for Reconstruction and Generation (arxiv.org)
2 points by doener on Jun 18, 2024 | hide | past | pdf | discuss
4952. Joint Audio and Symbolic Conditioning for Temporally Controlled Text-to-Music (arxiv.org)
2 points by tzury on Jun 18, 2024 | hide | past | pdf | discuss
4953. Refusal in language models is mediated by a single direction (arxiv.org)
209 points by Tomte on Jun 18, 2024 | hide | past | pdf | 44 comments
4954. Compositional Generative Modeling: A Single Model Is Not All You Need (arxiv.org)
2 points by j_maffe on Jun 18, 2024 | hide | past | pdf | discuss
4955. Exploring the Latest LLMs for Leaderboard Extraction (arxiv.org)
1 point by PaulHoule on Jun 18, 2024 | hide | past | pdf | discuss
4956. Transcendence: Generative Models Can Outperform the Experts That Train Them (arxiv.org)
4 points by sanxiyn on Jun 18, 2024 | hide | past | pdf | discuss
4957. LlamaCare: A Large Medical Language Model for Healthcare Knowledge Sharing (arxiv.org)
1 point by PaulHoule on Jun 18, 2024 | hide | past | pdf | discuss
4958. Depth Anything V2 (arxiv.org)
2 points by smusamashah on Jun 18, 2024 | hide | past | pdf | discuss
4959. The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse (arxiv.org)
1 point by PaulHoule on Jun 17, 2024 | hide | past | pdf | discuss
4960. Render and Diffuse: Aligning Image and Action for Behaviour Cloning (arxiv.org)
1 point by geox on Jun 17, 2024 | hide | past | pdf | discuss
4961. Large Language Models for Forecasting and Anomaly Detection (arxiv.org)
2 points by rntn on Jun 17, 2024 | hide | past | pdf | discuss
4962. Large Language Model Confidence Estimation via Black-Box Access (arxiv.org)
1 point by PaulHoule on Jun 17, 2024 | hide | past | pdf | discuss
4963. Neural Thermodynamic Integration: Free Energies from Diffusion Models (arxiv.org)
2 points by jasondavies on Jun 17, 2024 | hide | past | pdf | discuss
4964. An Image Is Worth 32 Tokens for Reconstruction and Generation [pdf] (arxiv.org)
3 points by croes on Jun 17, 2024 | hide | past | pdf | discuss
4965. AlphaMath Almost Zero: process Supervision without process (arxiv.org)
2 points by hardmaru on Jun 17, 2024 | hide | past | pdf | discuss
4966. Creativity has left the chat: The price of debiasing language models (arxiv.org)
174 points by hardmaru on Jun 17, 2024 | hide | past | pdf | 225 comments
4967. Progress Towards Decoding Visual Imagery via FNIRS (arxiv.org)
3 points by mpweiher on Jun 16, 2024 | hide | past | pdf | discuss
4968. RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots (arxiv.org)
1 point by belter on Jun 16, 2024 | hide | past | pdf | discuss
4969. Depth Anything V2 (arxiv.org)
3 points by skilled on Jun 16, 2024 | hide | past | pdf | 1 comment
4970. TextGrad: Automatic "Differentiation" via Text (arxiv.org)
6 points by Protostome on Jun 16, 2024 | hide | past | pdf | discuss
4971. Step-by-Step Diffusion Models: An Elementary Tutorial (arxiv.org)
3 points by Anon84 on Jun 15, 2024 | hide | past | pdf | discuss
4972. What Makes Language Models Good-Enough? (arxiv.org)
3 points by PaulHoule on Jun 15, 2024 | hide | past | pdf | discuss
4973. Scalable, Programmable Look-Up Table Based Neural Acceleration (arxiv.org)
3 points by PaulHoule on Jun 15, 2024 | hide | past | pdf | discuss
4974. Can language models serve as text-based world simulators? (arxiv.org)
95 points by mpweiher on Jun 15, 2024 | hide | past | pdf | 65 comments
4975. Understanding Hallucinations in Diffusion Models Through Mode Interpolation (arxiv.org)
3 points by jasondavies on Jun 15, 2024 | hide | past | pdf | discuss
4976. The Measure of Intelligence (arxiv.org)
16 points by max_ on Jun 15, 2024 | hide | past | pdf | 1 comment
4977. Converting In-Context Learning to Weights in Linearized-Attention Transformers (arxiv.org)
4 points by PaulHoule on Jun 15, 2024 | hide | past | pdf | 1 comment
4978. Discovering Optimization Algorithms With And For Large Language Models (arxiv.org)
2 points by optimalsolver on Jun 14, 2024 | hide | past | pdf | discuss
4979. Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation (arxiv.org)
2 points by reqo on Jun 14, 2024 | hide | past | pdf | discuss
4980. Step-by-Step Diffusion: An Elementary Tutorial (arxiv.org)
3 points by sebg on Jun 14, 2024 | hide | past | pdf | discuss