| 4951. |
An Image Is Worth 32 Tokens for Reconstruction and Generation (arxiv.org) |
|
2 points by doener on Jun 18, 2024 | hide | past | pdf | discuss
|
| 4952. |
Joint Audio and Symbolic Conditioning for Temporally Controlled Text-to-Music (arxiv.org) |
|
2 points by tzury on Jun 18, 2024 | hide | past | pdf | discuss
|
| 4953. |
Refusal in language models is mediated by a single direction (arxiv.org) |
|
209 points by Tomte on Jun 18, 2024 | hide | past | pdf | 44 comments
|
| 4954. |
Compositional Generative Modeling: A Single Model Is Not All You Need (arxiv.org) |
|
2 points by j_maffe on Jun 18, 2024 | hide | past | pdf | discuss
|
| 4955. |
Exploring the Latest LLMs for Leaderboard Extraction (arxiv.org) |
|
1 point by PaulHoule on Jun 18, 2024 | hide | past | pdf | discuss
|
| 4956. |
Transcendence: Generative Models Can Outperform the Experts That Train Them (arxiv.org) |
|
4 points by sanxiyn on Jun 18, 2024 | hide | past | pdf | discuss
|
| 4957. |
LlamaCare: A Large Medical Language Model for Healthcare Knowledge Sharing (arxiv.org) |
|
1 point by PaulHoule on Jun 18, 2024 | hide | past | pdf | discuss
|
| 4958. |
Depth Anything V2 (arxiv.org) |
|
2 points by smusamashah on Jun 18, 2024 | hide | past | pdf | discuss
|
| 4959. |
The Factorization Curse: Which Tokens You Predict Underlie the Reversal Curse (arxiv.org) |
|
1 point by PaulHoule on Jun 17, 2024 | hide | past | pdf | discuss
|
| 4960. |
Render and Diffuse: Aligning Image and Action for Behaviour Cloning (arxiv.org) |
|
1 point by geox on Jun 17, 2024 | hide | past | pdf | discuss
|
| 4961. |
Large Language Models for Forecasting and Anomaly Detection (arxiv.org) |
|
2 points by rntn on Jun 17, 2024 | hide | past | pdf | discuss
|
| 4962. |
Large Language Model Confidence Estimation via Black-Box Access (arxiv.org) |
|
1 point by PaulHoule on Jun 17, 2024 | hide | past | pdf | discuss
|
| 4963. |
Neural Thermodynamic Integration: Free Energies from Diffusion Models (arxiv.org) |
|
2 points by jasondavies on Jun 17, 2024 | hide | past | pdf | discuss
|
| 4964. |
An Image Is Worth 32 Tokens for Reconstruction and Generation [pdf] (arxiv.org) |
|
3 points by croes on Jun 17, 2024 | hide | past | pdf | discuss
|
| 4965. |
AlphaMath Almost Zero: process Supervision without process (arxiv.org) |
|
2 points by hardmaru on Jun 17, 2024 | hide | past | pdf | discuss
|
| 4966. |
Creativity has left the chat: The price of debiasing language models (arxiv.org) |
|
174 points by hardmaru on Jun 17, 2024 | hide | past | pdf | 225 comments
|
| 4967. |
Progress Towards Decoding Visual Imagery via FNIRS (arxiv.org) |
|
3 points by mpweiher on Jun 16, 2024 | hide | past | pdf | discuss
|
| 4968. |
RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots (arxiv.org) |
|
1 point by belter on Jun 16, 2024 | hide | past | pdf | discuss
|
| 4969. |
Depth Anything V2 (arxiv.org) |
|
3 points by skilled on Jun 16, 2024 | hide | past | pdf | 1 comment
|
| 4970. |
TextGrad: Automatic "Differentiation" via Text (arxiv.org) |
|
6 points by Protostome on Jun 16, 2024 | hide | past | pdf | discuss
|
| 4971. |
Step-by-Step Diffusion Models: An Elementary Tutorial (arxiv.org) |
|
3 points by Anon84 on Jun 15, 2024 | hide | past | pdf | discuss
|
| 4972. |
What Makes Language Models Good-Enough? (arxiv.org) |
|
3 points by PaulHoule on Jun 15, 2024 | hide | past | pdf | discuss
|
| 4973. |
Scalable, Programmable Look-Up Table Based Neural Acceleration (arxiv.org) |
|
3 points by PaulHoule on Jun 15, 2024 | hide | past | pdf | discuss
|
| 4974. |
Can language models serve as text-based world simulators? (arxiv.org) |
|
95 points by mpweiher on Jun 15, 2024 | hide | past | pdf | 65 comments
|
| 4975. |
Understanding Hallucinations in Diffusion Models Through Mode Interpolation (arxiv.org) |
|
3 points by jasondavies on Jun 15, 2024 | hide | past | pdf | discuss
|
| 4976. |
The Measure of Intelligence (arxiv.org) |
|
16 points by max_ on Jun 15, 2024 | hide | past | pdf | 1 comment
|
| 4977. |
Converting In-Context Learning to Weights in Linearized-Attention Transformers (arxiv.org) |
|
4 points by PaulHoule on Jun 15, 2024 | hide | past | pdf | 1 comment
|
| 4978. |
Discovering Optimization Algorithms With And For Large Language Models (arxiv.org) |
|
2 points by optimalsolver on Jun 14, 2024 | hide | past | pdf | discuss
|
| 4979. |
Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation (arxiv.org) |
|
2 points by reqo on Jun 14, 2024 | hide | past | pdf | discuss
|
| 4980. |
Step-by-Step Diffusion: An Elementary Tutorial (arxiv.org) |
|
3 points by sebg on Jun 14, 2024 | hide | past | pdf | discuss
|
| More |