about
5731. Should We Respect LLMs? Studying the Influence of Politeness on LLM Performance (arxiv.org)
1 point by gnicholas on Feb 29, 2024 | hide | past | pdf | 2 comments
5732. StableLM 1.6B Technical Report – includes all data, training, strategy (arxiv.org)
1 point by omnipotent_i on Feb 29, 2024 | hide | past | pdf | 1 comment
5733. Transparent Image Layer Diffusion Using Latent Transparency (arxiv.org)
43 points by lnyan on Feb 29, 2024 | hide | past | pdf | 1 comment
5734. Sora Generates Videos with Stunning Geometrical Consistency (arxiv.org)
2 points by amichail on Feb 28, 2024 | hide | past | pdf | discuss
5735. ModelGPT: Unleashing LLM's Capabilities for Tailored Model Generation (arxiv.org)
2 points by PaulHoule on Feb 28, 2024 | hide | past | pdf | discuss
5736. EyeTrans: Merging Human and Machine Attention for Neural Code Summarization (arxiv.org)
1 point by PaulHoule on Feb 28, 2024 | hide | past | pdf | discuss
5737. Storm: Assisting in Writing Wikipedia-Like Articles from Scratch with LLMs [pdf] (arxiv.org)
2 points by FergusArgyll on Feb 28, 2024 | hide | past | pdf | discuss
5738. The Era of 1-bit LLMs: ternary parameters for cost-effective computing (arxiv.org)
1040 points by fgfm on Feb 28, 2024 | hide | past | pdf | 447 comments
5739. Consensus learning: A novel decentralised ensemble learning paradigm (arxiv.org)
1 point by ganisgan on Feb 28, 2024 | hide | past | pdf | discuss
5740. EMO: Emote Portrait Alive (arxiv.org)
6 points by jonbaer on Feb 28, 2024 | hide | past | pdf | 2 comments
5741. Turn Waste into Worth: Rectifying Top-$K$ Router of Moe (arxiv.org)
1 point by PaulHoule on Feb 27, 2024 | hide | past | pdf | discuss
5742. Genie: Generative Interactive Environments (arxiv.org)
2 points by jonbaer on Feb 27, 2024 | hide | past | pdf | discuss
5743. Nemotron-4 15B large multilingual language model trained on 8T tokens (arxiv.org)
3 points by hack_ml on Feb 27, 2024 | hide | past | pdf | 1 comment
5744. A Survey on Data Selection for Language Models (arxiv.org)
1 point by sebg on Feb 27, 2024 | hide | past | pdf | discuss
5745. ReWOO: Decoupling Reasoning from Observations for Efficient Augmented LMs (2023) (arxiv.org)
2 points by CharlesW on Feb 27, 2024 | hide | past | pdf | discuss
5746. Defending LLMs against Jailbreaking Attacks via Backtranslation (arxiv.org)
67 points by saliagato on Feb 27, 2024 | hide | past | pdf | 48 comments
5747. How Do Humans Write Code? Large Models Do It the Same Way Too (arxiv.org)
1 point by saliagato on Feb 27, 2024 | hide | past | pdf | discuss
5748. Genie: Generative Interactive Environments (arxiv.org)
1 point by reqo on Feb 27, 2024 | hide | past | pdf | discuss
5749. SPML: A DSL for Defending LLMs Against Prompt Attacks (arxiv.org)
6 points by reshabh on Feb 27, 2024 | hide | past | pdf | 2 comments
5750. Genie: Generative Interactive Environments (arxiv.org)
5 points by Anuiran on Feb 26, 2024 | hide | past | pdf | 1 comment
5751. A Survey on Large Language Models for Recommendation (arxiv.org)
2 points by mfiguiere on Feb 26, 2024 | hide | past | pdf | discuss
5752. ScreenAI: A Vision-Language Model for UI and Infographics Understanding (arxiv.org)
1 point by gerlv on Feb 26, 2024 | hide | past | pdf | discuss
5753. Towards Efficient Generative LLM Serving: A Survey from Algorithms to Systems (arxiv.org)
1 point by sebg on Feb 26, 2024 | hide | past | pdf | discuss
5754. The Impact of Reasoning Step Length on Large Language Models (arxiv.org)
2 points by famouswaffles on Feb 26, 2024 | hide | past | pdf | discuss
5755. The Unreasonable Effectiveness of Eccentric Automatic Prompts (arxiv.org)
1 point by BerislavLopac on Feb 26, 2024 | hide | past | pdf | discuss
5756. Genie: Generative Interactive Environments (arxiv.org)
3 points by cma on Feb 26, 2024 | hide | past | pdf | discuss
5757. CLoVe: Encoding Compositional Language in Contrastive Vision-Language Models (arxiv.org)
1 point by ashvardanian on Feb 26, 2024 | hide | past | pdf | discuss
5758. The Unreasonable Effectiveness of Eccentric Automatic Prompts (arxiv.org)
2 points by helsinkiandrew on Feb 26, 2024 | hide | past | pdf | discuss
5759. Learning to Retrieve for Job Matching (arxiv.org)
1 point by PaulHoule on Feb 26, 2024 | hide | past | pdf | discuss
5760. An Empirical Evaluation of LLMs for Solving Offensive Security Challenges (arxiv.org)
1 point by todsacerdoti on Feb 25, 2024 | hide | past | pdf | discuss