about
5881. A Scalable RISC-V Vector Processor Enabling Efficient Multi-Precision Inference (arxiv.org)
3 points by PaulHoule on Feb 4, 2024 | hide | past | pdf | discuss
5882. CamPro: Camera-Based Anti-Facial Recognition (arxiv.org)
1 point by ulrischa on Feb 4, 2024 | hide | past | pdf | discuss
5883. Forecast Evaluation for Data Scientists: Common Pitfalls and Best Practices (arxiv.org)
3 points by Anon84 on Feb 4, 2024 | hide | past | pdf | discuss
5884. Characterizing and Recovering Information Loss in Text Simplification (arxiv.org)
3 points by PaulHoule on Feb 4, 2024 | hide | past | pdf | discuss
5885. TabLib: A Dataset Of 627M Tables With Context (2023) (arxiv.org)
1 point by optimalsolver on Feb 4, 2024 | hide | past | pdf | discuss
5886. Arrows of Time for Large Language Models (arxiv.org)
6 points by tianlong on Feb 2, 2024 | hide | past | pdf | 3 comments
5887. Formal Mathematics Statement Curriculum Learning (arxiv.org)
1 point by ofou on Feb 2, 2024 | hide | past | pdf | discuss
5888. 10M Context Length LLM Inference | UC Berkeley (arxiv.org)
4 points by hexman on Feb 1, 2024 | hide | past | pdf | discuss
5889. Large-Scale Reinforcement Learning for Diffusion Models (arxiv.org)
1 point by PaulHoule on Feb 1, 2024 | hide | past | pdf | discuss
5890. Infini-Gram: Scaling unbounded n-gram language models to a trillion tokens (arxiv.org)
2 points by nsagent on Feb 1, 2024 | hide | past | pdf | 1 comment
5891. Vision Mamba: Efficient Visual Representation Learning with Bidirectional SSM (arxiv.org)
74 points by andy99 on Feb 1, 2024 | hide | past | pdf | 16 comments
5892. Weaver, a family of LLMs focused on creative writing (arxiv.org)
11 points by vincent_s on Feb 1, 2024 | hide | past | pdf | 1 comment
5893. Matryoshka Representation Learning (arxiv.org)
83 points by fzliu on Feb 1, 2024 | hide | past | pdf | 11 comments
5894. Conformal Prediction Sets Improve Human Decision Making (arxiv.org)
2 points by PaulHoule on Jan 31, 2024 | hide | past | pdf | discuss
5895. Exploring OCR Capabilities of GPT-4V (arxiv.org)
2 points by saliagato on Jan 31, 2024 | hide | past | pdf | discuss
5896. Unique Identification of 50k+ Virtual Reality Users w Head and Hand Motion Data (arxiv.org)
1 point by bookofjoe on Jan 31, 2024 | hide | past | pdf | discuss
5897. Low-Resource Languages Jailbreak GPT-4 (arxiv.org)
1 point by rntn on Jan 31, 2024 | hide | past | pdf | discuss
5898. Moe-LLaVA: Mixture of Experts for Large Vision-Language Models (arxiv.org)
2 points by lukejagg on Jan 31, 2024 | hide | past | pdf | discuss
5899. Image Conditioned Inpainting in Latent Diffusion Models for Virtual Try-All (arxiv.org)
1 point by PaulHoule on Jan 31, 2024 | hide | past | pdf | discuss
5900. Rephrasing the Web: A Recipe for Compute and Data-Efficient Language Modeling (arxiv.org)
3 points by jasondavies on Jan 30, 2024 | hide | past | pdf | 1 comment
5901. Soaring from 4K to 400K: Extending LLM's Context with Activation Beacon (arxiv.org)
4 points by endymi0n on Jan 30, 2024 | hide | past | pdf | discuss
5902. The HSIC Bottleneck: Deep Learning Without Back-Propagation (arxiv.org)
1 point by danny00 on Jan 30, 2024 | hide | past | pdf | discuss
5903. Lumiere – a text-to-video diffusion model (arxiv.org)
1 point by koqoo on Jan 29, 2024 | hide | past | pdf | 1 comment
5904. Can Large Language Models Write Parallel Code? (arxiv.org)
1 point by danielnichols on Jan 29, 2024 | hide | past | pdf | discuss
5905. N-Hits: Neural Hierarchical Interpolation for Time Series Forecasting (arxiv.org)
1 point by rsecora on Jan 29, 2024 | hide | past | pdf | discuss
5906. Are Transformers Effective for Time Series Forecasting? (arxiv.org)
3 points by rsecora on Jan 29, 2024 | hide | past | pdf | discuss
5907. A Large-Scale Analysis of Dataset Cards on Hugging Face (arxiv.org)
2 points by PaulHoule on Jan 29, 2024 | hide | past | pdf | discuss
5908. Single Headed Attention RNN: Stop Thinking with Your Head (2019) (arxiv.org)
1 point by Buttons840 on Jan 29, 2024 | hide | past | pdf | 1 comment
5909. Learning Universal Predictors (arxiv.org)
66 points by jandrewrogers on Jan 29, 2024 | hide | past | pdf | 25 comments
5910. Lumiere: A Space-Time Diffusion Model for Video Generation (arxiv.org)
17 points by ulrischa on Jan 28, 2024 | hide | past | pdf | 1 comment