about
5701. Priority Sampling of Large Language Models for Compilers (arxiv.org)
1 point by PaulHoule on Mar 5, 2024 | hide | past | pdf | discuss
5702. Sigmoid Loss for Language Image Pre-Training (2023) (arxiv.org)
32 points by fzliu on Mar 4, 2024 | hide | past | pdf | discuss
5703. Understanding Tree Ensembles as Self-Regularizing Adaptive Smoothers (arxiv.org)
1 point by Anon84 on Mar 4, 2024 | hide | past | pdf | discuss
5704. Synthetic Data Almost from Scratch (arxiv.org)
2 points by milliondreams on Mar 3, 2024 | hide | past | pdf | 1 comment
5705. EMO: Emote Portrait Alive – Generating Expressive Portrait Videos (arxiv.org)
2 points by rolph on Mar 3, 2024 | hide | past | pdf | discuss
5706. Watermarking Makes Language Models Radioactive (arxiv.org)
2 points by PaulHoule on Mar 3, 2024 | hide | past | pdf | discuss
5707. Comparing Inferential Strategies of Humans and LLMs in Deductive Reasoning (arxiv.org)
2 points by PaulHoule on Mar 3, 2024 | hide | past | pdf | discuss
5708. Photonics for Sustainable Computing (arxiv.org)
1 point by CrypticShift on Mar 3, 2024 | hide | past | pdf | discuss
5709. MegaScale: Scaling Large Language Model Training to More Than 10k GPUs (arxiv.org)
2 points by PaulHoule on Mar 2, 2024 | hide | past | pdf | discuss
5710. Fn Benchmarks for Robust Evaluation of Reasoning Performance, and Reasoning Gap (arxiv.org)
2 points by ofou on Mar 2, 2024 | hide | past | pdf | 1 comment
5711. Genie: Generative Interactive Environments (arxiv.org)
2 points by caprock on Mar 2, 2024 | hide | past | pdf | 1 comment
5712. Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs (arxiv.org)
2 points by PaulHoule on Mar 2, 2024 | hide | past | pdf | discuss
5713. ArtPrompt: ASCII Art-Based Jailbreak Attacks Against Aligned LLMs (arxiv.org)
145 points by wut42 on Mar 2, 2024 | hide | past | pdf | 55 comments
5714. The Impact of Input Length on the Reasoning Performance of LLMs (arxiv.org)
2 points by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss
5715. Generating Expressive Portrait Videos with Audio2Video Diffusion Mode (arxiv.org)
2 points by belter on Mar 1, 2024 | hide | past | pdf | discuss
5716. Technical Report on the Checkfor.ai AI-Generated Text Classifier (arxiv.org)
1 point by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss
5717. Approaching Human-Level Forecasting with Language Models (arxiv.org)
2 points by kmdupree on Mar 1, 2024 | hide | past | pdf | discuss
5718. Evaluating Quantized Large Language Models (arxiv.org)
2 points by chuckhend on Mar 1, 2024 | hide | past | pdf | discuss
5719. Text Diffusion with Reinforced Conditioning (arxiv.org)
2 points by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss
5720. LiGNN: Graph Neural Networks at LinkedIn (arxiv.org)
1 point by Anon84 on Mar 1, 2024 | hide | past | pdf | discuss
5721. Evaluating the Performance of ChatGPT for Spam Email Detection (arxiv.org)
1 point by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss
5722. The Unreasonable Effectiveness of Eccentric Automatic Prompts (arxiv.org)
3 points by passwordoops on Mar 1, 2024 | hide | past | pdf | discuss
5723. Language-Based User Profiles for Recommendation (arxiv.org)
1 point by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss
5724. A Usage-Centric Take on Intent Understanding in E-Commerce (arxiv.org)
1 point by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss
5725. StarCoder 2 and The Stack v2: The Next Generation (arxiv.org)
2 points by fspeech on Mar 1, 2024 | hide | past | pdf | discuss
5726. Griffin: Mixing Gated Linear Recurrences with Local Attention for Efficient LMs (arxiv.org)
6 points by vagabund on Mar 1, 2024 | hide | past | pdf | 1 comment
5727. Is the System Message Important to Jailbreaks in Large Language Models? (arxiv.org)
3 points by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss
5728. Quantum Vision Transformers (arxiv.org)
2 points by thatxliner on Mar 1, 2024 | hide | past | pdf | discuss
5729. Chain-of-Thought Unfaithfulness as Disguised Accuracy (arxiv.org)
2 points by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss
5730. Optimizing Sub-Billion Parameter Language Models for On-Device Use Cases (arxiv.org)
2 points by PaulHoule on Mar 1, 2024 | hide | past | pdf | discuss