| 5821. |
Nvidia Audio Flamingo, Audio LM with Few-Shot Learning and Dialogue Abilities (arxiv.org) |
|
1 point by alok-g on Feb 14, 2024 | hide | past | pdf | discuss
|
| 5822. |
Suppressing Pink Elephants with Direct Principle Feedback (arxiv.org) |
|
5 points by swyx on Feb 13, 2024 | hide | past | pdf | 1 comment
|
| 5823. |
Unifying Point Cloud Perception, Generation and Editing with LLMs (arxiv.org) |
|
2 points by PaulHoule on Feb 13, 2024 | hide | past | pdf | discuss
|
| 5824. |
An Interactive Agent Foundation Model (arxiv.org) |
|
2 points by gerlv on Feb 13, 2024 | hide | past | pdf | discuss
|
| 5825. |
Secret Collusion Among Generative AI Agents [pdf] (arxiv.org) |
|
4 points by bikenaga on Feb 13, 2024 | hide | past | pdf | discuss
|
| 5826. |
Experimenting with Emerging RISC-V Systems for Decentralised Machine Learning (arxiv.org) |
|
3 points by redarguireda on Feb 13, 2024 | hide | past | pdf | discuss
|
| 5827. |
Neural Circuit Diagrams (arxiv.org) |
|
4 points by sva_ on Feb 13, 2024 | hide | past | pdf | discuss
|
| 5828. |
EEG-GPT (arxiv.org) |
|
4 points by docere on Feb 13, 2024 | hide | past | pdf | discuss
|
| 5829. |
Detecting Multimedia Generated by Large AI Models: A Survey (arxiv.org) |
|
2 points by PaulHoule on Feb 12, 2024 | hide | past | pdf | discuss
|
| 5830. |
QuIP#: Even Better LLM Quantization with Hadamard Incoherence, Lattice Codebooks (arxiv.org) |
|
2 points by tosh on Feb 12, 2024 | hide | past | pdf | discuss
|
| 5831. |
Self-Reflective, Hierarchical Agents for Large-Scale API Calls (arxiv.org) |
|
3 points by digitcatphd on Feb 12, 2024 | hide | past | pdf | discuss
|
| 5832. |
EvoMerge: Neuroevolution for Large Language Models (arxiv.org) |
|
3 points by PaulHoule on Feb 11, 2024 | hide | past | pdf | 1 comment
|
| 5833. |
Approximate Nearest Neighbor Search with Window Filters (arxiv.org) |
|
4 points by PaulHoule on Feb 11, 2024 | hide | past | pdf | discuss
|
| 5834. |
Death of MMLU? Revealing the Sensitivity and Instability of LLM Leaderboards (arxiv.org) |
|
2 points by hishamyahya on Feb 10, 2024 | hide | past | pdf | 1 comment
|
| 5835. |
Self-Discover: Large Language Models Self-Compose Reasoning Structures (arxiv.org) |
|
4 points by lukejagg on Feb 10, 2024 | hide | past | pdf | discuss
|
| 5836. |
Buffer Overflow in Mixture of Experts (arxiv.org) |
|
2 points by mfiguiere on Feb 10, 2024 | hide | past | pdf | discuss
|
| 5837. |
Distributed LLM inference over chain of mobile phones (arxiv.org) |
|
3 points by kevinysn on Feb 10, 2024 | hide | past | pdf | 1 comment
|
| 5838. |
Comprehensive Assessment of Jailbreak Attacks Against LLMs (arxiv.org) |
|
1 point by belter on Feb 9, 2024 | hide | past | pdf | 1 comment
|
| 5839. |
Hallucination Is Inevitable: An Innate Limitation of Large Language Models (arxiv.org) |
|
3 points by PerryCox on Feb 9, 2024 | hide | past | pdf | 2 comments
|
| 5840. |
Training LLMs for Reasoning Through Reverse Curriculum Reinforcement Learning (arxiv.org) |
|
1 point by belter on Feb 9, 2024 | hide | past | pdf | discuss
|
| 5841. |
What Algorithms Can Transformers Learn? A Study in Length Generalization (arxiv.org) |
|
2 points by sebg on Feb 9, 2024 | hide | past | pdf | 1 comment
|
| 5842. |
TransTroj: Transferable Backdoor Attacks to Pre-Trained Models (arxiv.org) |
|
22 points by indus on Feb 8, 2024 | hide | past | pdf | discuss
|
| 5843. |
Long Is More for Alignment: A Simple Baseline for Instruction Fine-Tuning (arxiv.org) |
|
2 points by max-andr on Feb 8, 2024 | hide | past | pdf | 1 comment
|
| 5844. |
Dot-product attention learns positional and semantic attention (arxiv.org) |
|
3 points by fzliu on Feb 8, 2024 | hide | past | pdf | discuss
|
| 5845. |
Grandmaster-Level Chess Without Search (arxiv.org) |
|
198 points by jonbaer on Feb 8, 2024 | hide | past | pdf | 130 comments
|
| 5846. |
Direct Language Model Alignment from Online AI Feedback (arxiv.org) |
|
61 points by drcwpl on Feb 8, 2024 | hide | past | pdf | 4 comments
|
| 5847. |
The Case for Co-Designing Model Architectures with Hardware (arxiv.org) |
|
1 point by PaulHoule on Feb 8, 2024 | hide | past | pdf | discuss
|
| 5848. |
ReGAL: Refactoring Programs to Discover Generalizable Abstractions (arxiv.org) |
|
2 points by PaulHoule on Feb 8, 2024 | hide | past | pdf | discuss
|
| 5849. |
Infini-Gram: Scaling Unbounded N-Gram Language Models to a Trillion Tokens (arxiv.org) |
|
3 points by PaulHoule on Feb 7, 2024 | hide | past | pdf | discuss
|
| 5850. |
Escalation Risks from Language Models in Military and Diplomatic Decision-Making (arxiv.org) |
|
52 points by rwmj on Feb 7, 2024 | hide | past | pdf | 12 comments
|
| More |