about
5281. Iterative reasoning preference optimization (arxiv.org)
19 points by Jimmc414 on May 1, 2024 | hide | past | pdf | 4 comments
5282. Building a Large Japanese Web Corpus for Large Language Models (arxiv.org)
76 points by PaulHoule on Apr 30, 2024 | hide | past | pdf | 29 comments
5283. Leveraging Human Speech Processing for Automated Dog Bark Classification (arxiv.org)
1 point by Jimmc414 on Apr 30, 2024 | hide | past | pdf | discuss
5284. Benchmarking Benchmark Leakage in Large Language Models (arxiv.org)
2 points by Jimmc414 on Apr 30, 2024 | hide | past | pdf | discuss
5285. Replacing Judges with Juries: Evaluating LLM Generations with a Panel of Models (arxiv.org)
47 points by Jimmc414 on Apr 30, 2024 | hide | past | pdf | 4 comments
5286. SatBird: Bird Species Distribution Modeling with Remote Sensing and Science Data (arxiv.org)
2 points by Brajeshwar on Apr 30, 2024 | hide | past | pdf | discuss
5287. What do Transformers Know about Government? (arxiv.org)
2 points by PaulHoule on Apr 30, 2024 | hide | past | pdf | discuss
5288. RAGCache: Efficient Knowledge Caching for Retrieval-Augmented Generation (arxiv.org)
33 points by PaulHoule on Apr 30, 2024 | hide | past | pdf | 3 comments
5289. A Survey of Transformers (arxiv.org)
3 points by Anon84 on Apr 30, 2024 | hide | past | pdf | discuss
5290. Long-form music generation with latent diffusion (arxiv.org)
3 points by doodlesdev on Apr 30, 2024 | hide | past | pdf | discuss
5291. Understanding Emergent Abilities of Language Models from the Loss Perspective (arxiv.org)
2 points by veryluckyxyz on Apr 30, 2024 | hide | past | pdf | 1 comment
5292. Large Language Model for Science: A Study on P vs. NP (arxiv.org)
2 points by jonbaer on Apr 30, 2024 | hide | past | pdf | discuss
5293. Pretraining Without Attention (arxiv.org)
2 points by SongofEarth on Apr 30, 2024 | hide | past | pdf | discuss
5294. Benchmarking Mobile Device Control Agents Across Diverse Configurations (arxiv.org)
2 points by Jimmc414 on Apr 29, 2024 | hide | past | pdf | discuss
5295. Tunnel Try-On: Excavating Spatial-Temporal Tunnels for Virtual Try-On in Videos (arxiv.org)
3 points by Jimmc414 on Apr 29, 2024 | hide | past | pdf | discuss
5296. Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs (arxiv.org)
2 points by Jimmc414 on Apr 29, 2024 | hide | past | pdf | discuss
5297. NLP-enabled trajectory map-matching using transformer sequence-to-sequence model (arxiv.org)
1 point by PaulHoule on Apr 29, 2024 | hide | past | pdf | discuss
5298. Fine Tuning LLM for Enterprise: Practical Guidelines and Recommendations (arxiv.org)
2 points by PaulHoule on Apr 29, 2024 | hide | past | pdf | discuss
5299. More Room for Language: Investigating the Effect of Retrieval on Language Models (arxiv.org)
3 points by PaulHoule on Apr 29, 2024 | hide | past | pdf | discuss
5300. Self-Playing Adversarial Language Game Enhances LLM Reasoning (arxiv.org)
4 points by rootforce on Apr 29, 2024 | hide | past | pdf | 1 comment
5301. Understanding Emergent Abilities of Language Models from the Loss Perspective (arxiv.org)
6 points by maccaw on Apr 29, 2024 | hide | past | pdf | 1 comment
5302. Efficiently Serving Large Language Models Through FP6-Centric Algorithm-System (arxiv.org)
5 points by tosh on Apr 28, 2024 | hide | past | pdf | discuss
5303. Make Your LLM Utilize the Context (arxiv.org)
2 points by tosh on Apr 28, 2024 | hide | past | pdf | discuss
5304. The Matrix: A Bayesian Learning Model for LLMs (arxiv.org)
1 point by tosh on Apr 28, 2024 | hide | past | pdf | discuss
5305. Fewer Truncations Improve Language Modeling (arxiv.org)
1 point by PaulHoule on Apr 28, 2024 | hide | past | pdf | discuss
5306. LoRA+: Efficient Low Rank Adaptation of Large Models (arxiv.org)
181 points by veryluckyxyz on Apr 28, 2024 | hide | past | pdf | 47 comments
5307. Step Differences in Instructional Video (arxiv.org)
15 points by zerojames on Apr 28, 2024 | hide | past | pdf | discuss
5308. EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided RL (arxiv.org)
2 points by PaulHoule on Apr 28, 2024 | hide | past | pdf | discuss
5309. Survey on Embedding Models for Knowledge Graph and Its Applications (arxiv.org)
3 points by Anon84 on Apr 28, 2024 | hide | past | pdf | discuss
5310. Let's Think Dot by Dot: Hidden Computation in Transformer Language Models (arxiv.org)
159 points by Jimmc414 on Apr 27, 2024 | hide | past | pdf | 32 comments