about
4921. Predicting Neural Network Accuracy from Weights (2021) (arxiv.org)
1 point by johnsutor on Jun 23, 2024 | hide | past | pdf | discuss
4922. Are LLMs Naturally Good at Synthetic Tabular Data Generation? (arxiv.org)
2 points by belter on Jun 23, 2024 | hide | past | pdf | discuss
4923. Delving into ChatGPT usage in academic writing through excess vocabulary (arxiv.org)
164 points by zdw on Jun 22, 2024 | hide | past | pdf | 105 comments
4924. DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models (arxiv.org)
3 points by fintechie on Jun 22, 2024 | hide | past | pdf | discuss
4925. Evaluating the World Model Implicit in a Generative Model (arxiv.org)
3 points by walterbell on Jun 22, 2024 | hide | past | pdf | discuss
4926. Adversarial Perturbations Cannot Reliably Protect Artists from Generative AI (arxiv.org)
4 points by GaggiX on Jun 22, 2024 | hide | past | pdf | 1 comment
4927. Is Programming by Example Solved by LLMs? (arxiv.org)
1 point by PaulHoule on Jun 22, 2024 | hide | past | pdf | discuss
4928. Delving into ChatGPT usage in academic writing through excess vocabulary (arxiv.org)
2 points by taubek on Jun 21, 2024 | hide | past | pdf | discuss
4929. Q*: Improving Multi-Step Reasoning for LLMs with Deliberative Planning (arxiv.org)
34 points by marcelmarais on Jun 21, 2024 | hide | past | pdf | 3 comments
4930. StableSemantics: A Synthetic Dataset of Semantic Representations in Images (arxiv.org)
1 point by afluo on Jun 21, 2024 | hide | past | pdf | discuss
4931. Connecting the Dots: LLMs Can Infer and Verbalize Latent Structure (arxiv.org)
2 points by jasondavies on Jun 21, 2024 | hide | past | pdf | discuss
4932. Adversarial Perturbations Cannot Reliably Protect Artists from Generative AI (arxiv.org)
2 points by idiliv on Jun 21, 2024 | hide | past | pdf | discuss
4933. A Tutorial on Thompson Sampling (arxiv.org)
2 points by sebg on Jun 21, 2024 | hide | past | pdf | discuss
4934. The Paradox of Learning to Reason from Data (2022) (arxiv.org)
1 point by todsacerdoti on Jun 21, 2024 | hide | past | pdf | discuss
4935. RAR-B: Reasoning as Retrieval Benchmark (arxiv.org)
11 points by aminst on Jun 20, 2024 | hide | past | pdf | discuss
4936. A Survey of LLMs for Financial Applications: Progress, Prospects and Challenges (arxiv.org)
2 points by sebg on Jun 20, 2024 | hide | past | pdf | discuss
4937. Multilingual Multimodal Data Hub and Benchmark for Southeast Asian Languages (arxiv.org)
1 point by tellarin on Jun 20, 2024 | hide | past | pdf | discuss
4938. Transcendence: Generative Models Can Outperform the Experts That Train Them (arxiv.org)
2 points by stefankuehnel on Jun 19, 2024 | hide | past | pdf | discuss
4939. CoLoR-Filter: Conditional Loss Reduction Filtering for Targeted LM Pre-Training (arxiv.org)
2 points by pizza on Jun 19, 2024 | hide | past | pdf | discuss
4940. Flash Diffusion: Accelerating Any Conditional Diffusion Model (arxiv.org)
2 points by jasondavies on Jun 19, 2024 | hide | past | pdf | discuss
4941. Llamafuzz: Large Language Model Enhanced Greybox Fuzzing (arxiv.org)
3 points by PaulHoule on Jun 19, 2024 | hide | past | pdf | discuss
4942. Benchmarking the continuous improvement of language agents in deployment (arxiv.org)
2 points by polymorph1sm on Jun 19, 2024 | hide | past | pdf | discuss
4943. Adversarial Perturbations Cannot Reliably Protect Artists from Generative AI (arxiv.org)
5 points by dpaleka on Jun 19, 2024 | hide | past | pdf | discuss
4944. An Evaluation Benchmark for Autoformalization in Lean4 (arxiv.org)
2 points by PaulHoule on Jun 19, 2024 | hide | past | pdf | discuss
4945. Seq1F1B: Efficient Sequence-Level Pipeline Parallelism for LLM Training (arxiv.org)
1 point by wseqyrku on Jun 19, 2024 | hide | past | pdf | discuss
4946. Multilingual Multimodal Data Hub and Benchmark for Southeast Asian Languages (arxiv.org)
1 point by tellarin on Jun 19, 2024 | hide | past | pdf | discuss
4947. Transcendence: Generative Models Can Outperform the Experts That Train Them (arxiv.org)
2 points by jasondavies on Jun 19, 2024 | hide | past | pdf | discuss
4948. Extending the Attention Mechanism in Transformers to Continuum operators (arxiv.org)
1 point by bvsrinivasan on Jun 19, 2024 | hide | past | pdf | discuss
4949. Garak: A Framework for Security Probing Large Language Models (arxiv.org)
2 points by Anon84 on Jun 19, 2024 | hide | past | pdf | discuss
4950. Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning (arxiv.org)
1 point by PaulHoule on Jun 18, 2024 | hide | past | pdf | discuss