about
5821. Nvidia Audio Flamingo, Audio LM with Few-Shot Learning and Dialogue Abilities (arxiv.org)
1 point by alok-g on Feb 14, 2024 | hide | past | pdf | discuss
5822. Suppressing Pink Elephants with Direct Principle Feedback (arxiv.org)
5 points by swyx on Feb 13, 2024 | hide | past | pdf | 1 comment
5823. Unifying Point Cloud Perception, Generation and Editing with LLMs (arxiv.org)
2 points by PaulHoule on Feb 13, 2024 | hide | past | pdf | discuss
5824. An Interactive Agent Foundation Model (arxiv.org)
2 points by gerlv on Feb 13, 2024 | hide | past | pdf | discuss
5825. Secret Collusion Among Generative AI Agents [pdf] (arxiv.org)
4 points by bikenaga on Feb 13, 2024 | hide | past | pdf | discuss
5826. Experimenting with Emerging RISC-V Systems for Decentralised Machine Learning (arxiv.org)
3 points by redarguireda on Feb 13, 2024 | hide | past | pdf | discuss
5827. Neural Circuit Diagrams (arxiv.org)
4 points by sva_ on Feb 13, 2024 | hide | past | pdf | discuss
5828. EEG-GPT (arxiv.org)
4 points by docere on Feb 13, 2024 | hide | past | pdf | discuss
5829. Detecting Multimedia Generated by Large AI Models: A Survey (arxiv.org)
2 points by PaulHoule on Feb 12, 2024 | hide | past | pdf | discuss
5830. QuIP#: Even Better LLM Quantization with Hadamard Incoherence, Lattice Codebooks (arxiv.org)
2 points by tosh on Feb 12, 2024 | hide | past | pdf | discuss
5831. Self-Reflective, Hierarchical Agents for Large-Scale API Calls (arxiv.org)
3 points by digitcatphd on Feb 12, 2024 | hide | past | pdf | discuss
5832. EvoMerge: Neuroevolution for Large Language Models (arxiv.org)
3 points by PaulHoule on Feb 11, 2024 | hide | past | pdf | 1 comment
5833. Approximate Nearest Neighbor Search with Window Filters (arxiv.org)
4 points by PaulHoule on Feb 11, 2024 | hide | past | pdf | discuss
5834. Death of MMLU? Revealing the Sensitivity and Instability of LLM Leaderboards (arxiv.org)
2 points by hishamyahya on Feb 10, 2024 | hide | past | pdf | 1 comment
5835. Self-Discover: Large Language Models Self-Compose Reasoning Structures (arxiv.org)
4 points by lukejagg on Feb 10, 2024 | hide | past | pdf | discuss
5836. Buffer Overflow in Mixture of Experts (arxiv.org)
2 points by mfiguiere on Feb 10, 2024 | hide | past | pdf | discuss
5837. Distributed LLM inference over chain of mobile phones (arxiv.org)
3 points by kevinysn on Feb 10, 2024 | hide | past | pdf | 1 comment
5838. Comprehensive Assessment of Jailbreak Attacks Against LLMs (arxiv.org)
1 point by belter on Feb 9, 2024 | hide | past | pdf | 1 comment
5839. Hallucination Is Inevitable: An Innate Limitation of Large Language Models (arxiv.org)
3 points by PerryCox on Feb 9, 2024 | hide | past | pdf | 2 comments
5840. Training LLMs for Reasoning Through Reverse Curriculum Reinforcement Learning (arxiv.org)
1 point by belter on Feb 9, 2024 | hide | past | pdf | discuss
5841. What Algorithms Can Transformers Learn? A Study in Length Generalization (arxiv.org)
2 points by sebg on Feb 9, 2024 | hide | past | pdf | 1 comment
5842. TransTroj: Transferable Backdoor Attacks to Pre-Trained Models (arxiv.org)
22 points by indus on Feb 8, 2024 | hide | past | pdf | discuss
5843. Long Is More for Alignment: A Simple Baseline for Instruction Fine-Tuning (arxiv.org)
2 points by max-andr on Feb 8, 2024 | hide | past | pdf | 1 comment
5844. Dot-product attention learns positional and semantic attention (arxiv.org)
3 points by fzliu on Feb 8, 2024 | hide | past | pdf | discuss
5845. Grandmaster-Level Chess Without Search (arxiv.org)
198 points by jonbaer on Feb 8, 2024 | hide | past | pdf | 130 comments
5846. Direct Language Model Alignment from Online AI Feedback (arxiv.org)
61 points by drcwpl on Feb 8, 2024 | hide | past | pdf | 4 comments
5847. The Case for Co-Designing Model Architectures with Hardware (arxiv.org)
1 point by PaulHoule on Feb 8, 2024 | hide | past | pdf | discuss
5848. ReGAL: Refactoring Programs to Discover Generalizable Abstractions (arxiv.org)
2 points by PaulHoule on Feb 8, 2024 | hide | past | pdf | discuss
5849. Infini-Gram: Scaling Unbounded N-Gram Language Models to a Trillion Tokens (arxiv.org)
3 points by PaulHoule on Feb 7, 2024 | hide | past | pdf | discuss
5850. Escalation Risks from Language Models in Military and Diplomatic Decision-Making (arxiv.org)
52 points by rwmj on Feb 7, 2024 | hide | past | pdf | 12 comments