about
5521. Model Stock: All we need is just a few fine-tuned models (arxiv.org)
2 points by tosh on Apr 1, 2024 | hide | past | pdf | discuss
5522. Building a Vietnamese Language Model with Advanced Continual Pre-Training (arxiv.org)
1 point by PaulHoule on Apr 1, 2024 | hide | past | pdf | discuss
5523. Born with a Silver Spoon? Socioeconomic Bias in Large Language Models (arxiv.org)
3 points by PaulHoule on Apr 1, 2024 | hide | past | pdf | discuss
5524. Unsolvable Problem Detection: Evaluating Trustworthiness of Vision LMs (arxiv.org)
1 point by villaaston1 on Apr 1, 2024 | hide | past | pdf | discuss
5525. What Was Your Prompt? A Remote Keylogging Attack on AI Assistants (arxiv.org)
2 points by maayank on Apr 1, 2024 | hide | past | pdf | discuss
5526. Editing Models with Task Arithmetic (2022) (arxiv.org)
1 point by tosh on Apr 1, 2024 | hide | past | pdf | discuss
5527. Long-form factuality in large language models (arxiv.org)
2 points by rntn on Apr 1, 2024 | hide | past | pdf | discuss
5528. Instruction-Following Evaluation for Large Language Models (arxiv.org)
1 point by tosh on Apr 1, 2024 | hide | past | pdf | discuss
5529. Topos of Transformer Networks – Thoughts? (arxiv.org)
3 points by cheekyfibonacci on Apr 1, 2024 | hide | past | pdf | 1 comment
5530. Agents Need Not Know Their Purpose (arxiv.org)
2 points by optimalsolver on Apr 1, 2024 | hide | past | pdf | discuss
5531. Are We on the Right Way for Evaluating Large Vision-Language Models? (arxiv.org)
2 points by yhzan on Apr 1, 2024 | hide | past | pdf | discuss
5532. Jamba: A Hybrid Transformer-Mamba Language Model (arxiv.org)
74 points by eitanturok on Apr 1, 2024 | hide | past | pdf | 6 comments
5533. InternLM2 (arxiv.org)
136 points by milliondreams on Mar 31, 2024 | hide | past | pdf | 24 comments
5534. Adaptive RAG – dynamic retrieval methods adjustment (arxiv.org)
126 points by milliondreams on Mar 31, 2024 | hide | past | pdf | 39 comments
5535. Mini-Gemini: Mining the Potential of Multi-Modality Vision Language Models (arxiv.org)
83 points by milliondreams on Mar 31, 2024 | hide | past | pdf | 7 comments
5536. The Attention of Mamba Models (arxiv.org)
2 points by jonbaer on Mar 30, 2024 | hide | past | pdf | discuss
5537. Numerical issues in maximum likelihood parameter estimation for GP interpolation (arxiv.org)
1 point by sieste on Mar 30, 2024 | hide | past | pdf | discuss
5538. SportsNGEN: Sustained Generation of Multi-Player Sports Gameplay (arxiv.org)
3 points by PaulHoule on Mar 29, 2024 | hide | past | pdf | discuss
5539. TnT-LLM: Text Mining at Scale with Large Language Models (arxiv.org)
66 points by PaulHoule on Mar 29, 2024 | hide | past | pdf | 7 comments
5540. Social Intelligence Data Infrastructure (arxiv.org)
1 point by PaulHoule on Mar 29, 2024 | hide | past | pdf | discuss
5541. Can LLMs Separate Instructions from Data? and What Do We Even Mean by That? (arxiv.org)
2 points by rootforce on Mar 29, 2024 | hide | past | pdf | discuss
5542. Out of One, Many: Using Language Models to Simulate Human Samples (2022) (arxiv.org)
2 points by rntn on Mar 29, 2024 | hide | past | pdf | discuss
5543. Long-form factuality in large language models (arxiv.org)
18 points by rootforce on Mar 29, 2024 | hide | past | pdf | 3 comments
5544. LLaMA: Open and Efficient Foundation Language Models (2023) (arxiv.org)
2 points by tosh on Mar 28, 2024 | hide | past | pdf | discuss
5545. Long-Form Factuality in Large Language Models (arxiv.org)
1 point by tosh on Mar 28, 2024 | hide | past | pdf | discuss
5546. Long-form factuality in large language models (arxiv.org)
1 point by TheIronYuppie on Mar 28, 2024 | hide | past | pdf | discuss
5547. Reconstructing Scenes with an Autoregressive Structured Language Model (arxiv.org)
1 point by PaulHoule on Mar 28, 2024 | hide | past | pdf | discuss
5548. EasyEdit: An Easy-to-Use Knowledge Editing Framework for Large Language Models (arxiv.org)
1 point by PaulHoule on Mar 27, 2024 | hide | past | pdf | discuss
5549. LISA: Layerwise Importance Sampling for Memory-Efficient LLM Fine-Tuning (arxiv.org)
3 points by convexstrictly on Mar 27, 2024 | hide | past | pdf | 1 comment
5550. Magic Fixup: Streamlining Photo Editing by Watching Dynamic Videos (arxiv.org)
1 point by PaulHoule on Mar 27, 2024 | hide | past | pdf | discuss