| 5281. |
Iterative reasoning preference optimization (arxiv.org) |
|
19 points by Jimmc414 on May 1, 2024 | hide | past | pdf | 4 comments
|
| 5282. |
Building a Large Japanese Web Corpus for Large Language Models (arxiv.org) |
|
76 points by PaulHoule on Apr 30, 2024 | hide | past | pdf | 29 comments
|
| 5283. |
Leveraging Human Speech Processing for Automated Dog Bark Classification (arxiv.org) |
|
1 point by Jimmc414 on Apr 30, 2024 | hide | past | pdf | discuss
|
| 5284. |
Benchmarking Benchmark Leakage in Large Language Models (arxiv.org) |
|
2 points by Jimmc414 on Apr 30, 2024 | hide | past | pdf | discuss
|
| 5285. |
Replacing Judges with Juries: Evaluating LLM Generations with a Panel of Models (arxiv.org) |
|
47 points by Jimmc414 on Apr 30, 2024 | hide | past | pdf | 4 comments
|
| 5286. |
SatBird: Bird Species Distribution Modeling with Remote Sensing and Science Data (arxiv.org) |
|
2 points by Brajeshwar on Apr 30, 2024 | hide | past | pdf | discuss
|
| 5287. |
What do Transformers Know about Government? (arxiv.org) |
|
2 points by PaulHoule on Apr 30, 2024 | hide | past | pdf | discuss
|
| 5288. |
RAGCache: Efficient Knowledge Caching for Retrieval-Augmented Generation (arxiv.org) |
|
33 points by PaulHoule on Apr 30, 2024 | hide | past | pdf | 3 comments
|
| 5289. |
A Survey of Transformers (arxiv.org) |
|
3 points by Anon84 on Apr 30, 2024 | hide | past | pdf | discuss
|
| 5290. |
Long-form music generation with latent diffusion (arxiv.org) |
|
3 points by doodlesdev on Apr 30, 2024 | hide | past | pdf | discuss
|
| 5291. |
Understanding Emergent Abilities of Language Models from the Loss Perspective (arxiv.org) |
|
2 points by veryluckyxyz on Apr 30, 2024 | hide | past | pdf | 1 comment
|
| 5292. |
Large Language Model for Science: A Study on P vs. NP (arxiv.org) |
|
2 points by jonbaer on Apr 30, 2024 | hide | past | pdf | discuss
|
| 5293. |
Pretraining Without Attention (arxiv.org) |
|
2 points by SongofEarth on Apr 30, 2024 | hide | past | pdf | discuss
|
| 5294. |
Benchmarking Mobile Device Control Agents Across Diverse Configurations (arxiv.org) |
|
2 points by Jimmc414 on Apr 29, 2024 | hide | past | pdf | discuss
|
| 5295. |
Tunnel Try-On: Excavating Spatial-Temporal Tunnels for Virtual Try-On in Videos (arxiv.org) |
|
3 points by Jimmc414 on Apr 29, 2024 | hide | past | pdf | discuss
|
| 5296. |
Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs (arxiv.org) |
|
2 points by Jimmc414 on Apr 29, 2024 | hide | past | pdf | discuss
|
| 5297. |
NLP-enabled trajectory map-matching using transformer sequence-to-sequence model (arxiv.org) |
|
1 point by PaulHoule on Apr 29, 2024 | hide | past | pdf | discuss
|
| 5298. |
Fine Tuning LLM for Enterprise: Practical Guidelines and Recommendations (arxiv.org) |
|
2 points by PaulHoule on Apr 29, 2024 | hide | past | pdf | discuss
|
| 5299. |
More Room for Language: Investigating the Effect of Retrieval on Language Models (arxiv.org) |
|
3 points by PaulHoule on Apr 29, 2024 | hide | past | pdf | discuss
|
| 5300. |
Self-Playing Adversarial Language Game Enhances LLM Reasoning (arxiv.org) |
|
4 points by rootforce on Apr 29, 2024 | hide | past | pdf | 1 comment
|
| 5301. |
Understanding Emergent Abilities of Language Models from the Loss Perspective (arxiv.org) |
|
6 points by maccaw on Apr 29, 2024 | hide | past | pdf | 1 comment
|
| 5302. |
Efficiently Serving Large Language Models Through FP6-Centric Algorithm-System (arxiv.org) |
|
5 points by tosh on Apr 28, 2024 | hide | past | pdf | discuss
|
| 5303. |
Make Your LLM Utilize the Context (arxiv.org) |
|
2 points by tosh on Apr 28, 2024 | hide | past | pdf | discuss
|
| 5304. |
The Matrix: A Bayesian Learning Model for LLMs (arxiv.org) |
|
1 point by tosh on Apr 28, 2024 | hide | past | pdf | discuss
|
| 5305. |
Fewer Truncations Improve Language Modeling (arxiv.org) |
|
1 point by PaulHoule on Apr 28, 2024 | hide | past | pdf | discuss
|
| 5306. |
LoRA+: Efficient Low Rank Adaptation of Large Models (arxiv.org) |
|
181 points by veryluckyxyz on Apr 28, 2024 | hide | past | pdf | 47 comments
|
| 5307. |
Step Differences in Instructional Video (arxiv.org) |
|
15 points by zerojames on Apr 28, 2024 | hide | past | pdf | discuss
|
| 5308. |
EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided RL (arxiv.org) |
|
2 points by PaulHoule on Apr 28, 2024 | hide | past | pdf | discuss
|
| 5309. |
Survey on Embedding Models for Knowledge Graph and Its Applications (arxiv.org) |
|
3 points by Anon84 on Apr 28, 2024 | hide | past | pdf | discuss
|
| 5310. |
Let's Think Dot by Dot: Hidden Computation in Transformer Language Models (arxiv.org) |
|
159 points by Jimmc414 on Apr 27, 2024 | hide | past | pdf | 32 comments
|
| More |