| 5521. |
Model Stock: All we need is just a few fine-tuned models (arxiv.org) |
|
2 points by tosh on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5522. |
Building a Vietnamese Language Model with Advanced Continual Pre-Training (arxiv.org) |
|
1 point by PaulHoule on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5523. |
Born with a Silver Spoon? Socioeconomic Bias in Large Language Models (arxiv.org) |
|
3 points by PaulHoule on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5524. |
Unsolvable Problem Detection: Evaluating Trustworthiness of Vision LMs (arxiv.org) |
|
1 point by villaaston1 on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5525. |
What Was Your Prompt? A Remote Keylogging Attack on AI Assistants (arxiv.org) |
|
2 points by maayank on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5526. |
Editing Models with Task Arithmetic (2022) (arxiv.org) |
|
1 point by tosh on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5527. |
Long-form factuality in large language models (arxiv.org) |
|
2 points by rntn on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5528. |
Instruction-Following Evaluation for Large Language Models (arxiv.org) |
|
1 point by tosh on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5529. |
Topos of Transformer Networks – Thoughts? (arxiv.org) |
|
3 points by cheekyfibonacci on Apr 1, 2024 | hide | past | pdf | 1 comment
|
| 5530. |
Agents Need Not Know Their Purpose (arxiv.org) |
|
2 points by optimalsolver on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5531. |
Are We on the Right Way for Evaluating Large Vision-Language Models? (arxiv.org) |
|
2 points by yhzan on Apr 1, 2024 | hide | past | pdf | discuss
|
| 5532. |
Jamba: A Hybrid Transformer-Mamba Language Model (arxiv.org) |
|
74 points by eitanturok on Apr 1, 2024 | hide | past | pdf | 6 comments
|
| 5533. |
InternLM2 (arxiv.org) |
|
136 points by milliondreams on Mar 31, 2024 | hide | past | pdf | 24 comments
|
| 5534. |
Adaptive RAG – dynamic retrieval methods adjustment (arxiv.org) |
|
126 points by milliondreams on Mar 31, 2024 | hide | past | pdf | 39 comments
|
| 5535. |
Mini-Gemini: Mining the Potential of Multi-Modality Vision Language Models (arxiv.org) |
|
83 points by milliondreams on Mar 31, 2024 | hide | past | pdf | 7 comments
|
| 5536. |
The Attention of Mamba Models (arxiv.org) |
|
2 points by jonbaer on Mar 30, 2024 | hide | past | pdf | discuss
|
| 5537. |
Numerical issues in maximum likelihood parameter estimation for GP interpolation (arxiv.org) |
|
1 point by sieste on Mar 30, 2024 | hide | past | pdf | discuss
|
| 5538. |
SportsNGEN: Sustained Generation of Multi-Player Sports Gameplay (arxiv.org) |
|
3 points by PaulHoule on Mar 29, 2024 | hide | past | pdf | discuss
|
| 5539. |
TnT-LLM: Text Mining at Scale with Large Language Models (arxiv.org) |
|
66 points by PaulHoule on Mar 29, 2024 | hide | past | pdf | 7 comments
|
| 5540. |
Social Intelligence Data Infrastructure (arxiv.org) |
|
1 point by PaulHoule on Mar 29, 2024 | hide | past | pdf | discuss
|
| 5541. |
Can LLMs Separate Instructions from Data? and What Do We Even Mean by That? (arxiv.org) |
|
2 points by rootforce on Mar 29, 2024 | hide | past | pdf | discuss
|
| 5542. |
Out of One, Many: Using Language Models to Simulate Human Samples (2022) (arxiv.org) |
|
2 points by rntn on Mar 29, 2024 | hide | past | pdf | discuss
|
| 5543. |
Long-form factuality in large language models (arxiv.org) |
|
18 points by rootforce on Mar 29, 2024 | hide | past | pdf | 3 comments
|
| 5544. |
LLaMA: Open and Efficient Foundation Language Models (2023) (arxiv.org) |
|
2 points by tosh on Mar 28, 2024 | hide | past | pdf | discuss
|
| 5545. |
Long-Form Factuality in Large Language Models (arxiv.org) |
|
1 point by tosh on Mar 28, 2024 | hide | past | pdf | discuss
|
| 5546. |
Long-form factuality in large language models (arxiv.org) |
|
1 point by TheIronYuppie on Mar 28, 2024 | hide | past | pdf | discuss
|
| 5547. |
Reconstructing Scenes with an Autoregressive Structured Language Model (arxiv.org) |
|
1 point by PaulHoule on Mar 28, 2024 | hide | past | pdf | discuss
|
| 5548. |
EasyEdit: An Easy-to-Use Knowledge Editing Framework for Large Language Models (arxiv.org) |
|
1 point by PaulHoule on Mar 27, 2024 | hide | past | pdf | discuss
|
| 5549. |
LISA: Layerwise Importance Sampling for Memory-Efficient LLM Fine-Tuning (arxiv.org) |
|
3 points by convexstrictly on Mar 27, 2024 | hide | past | pdf | 1 comment
|
| 5550. |
Magic Fixup: Streamlining Photo Editing by Watching Dynamic Videos (arxiv.org) |
|
1 point by PaulHoule on Mar 27, 2024 | hide | past | pdf | discuss
|
| More |