about
4621. Choosing the "Brain" for your AI-powered app – My new method, feedback requested (arxiv.org)
2 points by robert_sim on Aug 29, 2024 | hide | past | pdf | 1 comment
4622. Let's Verify Step by Step (arxiv.org)
1 point by sharemywin on Aug 29, 2024 | hide | past | pdf | 2 comments
4623. The Mamba in the Llama (arxiv.org)
2 points by dpstart01 on Aug 29, 2024 | hide | past | pdf | discuss
4624. Chain and Hash, an LLM Fingerprinting Technique (arxiv.org)
2 points by wslh on Aug 29, 2024 | hide | past | pdf | discuss
4625. Diffusion Models Are Real-Time Game Engines (arxiv.org)
2 points by LordNibbler on Aug 29, 2024 | hide | past | pdf | discuss
4626. Implicit Bias Matters for Language Models (arxiv.org)
2 points by fzliu on Aug 28, 2024 | hide | past | pdf | discuss
4627. CRQBench: A Benchmark of Code Reasoning Questions (arxiv.org)
1 point by PaulHoule on Aug 28, 2024 | hide | past | pdf | discuss
4628. Multilingual Multimodal Data Hub and Benchmark for Southeast Asian Languages (arxiv.org)
2 points by tellarin on Aug 28, 2024 | hide | past | pdf | discuss
4629. Sapiens: Foundation for Human Vision Models (arxiv.org)
58 points by soulofmischief on Aug 28, 2024 | hide | past | pdf | 1 comment
4630. Learning to Move Like Professional Counter-Strike Players (arxiv.org)
4 points by SerCe on Aug 28, 2024 | hide | past | pdf | discuss
4631. LongWriter: Unleashing 10k Word Generation from Long Context LLMs (arxiv.org)
7 points by PaulHoule on Aug 28, 2024 | hide | past | pdf | 1 comment
4632. The Geometry of the Set of Equivalent Linear Neural Networks (arxiv.org)
1 point by 082349872349872 on Aug 27, 2024 | hide | past | pdf | discuss
4633. NeuroPapyri: A Deep Attention Embedding Network for Handwritten Papyri Retrieval (arxiv.org)
1 point by PaulHoule on Aug 26, 2024 | hide | past | pdf | discuss
4634. Large Model Strategic Thinking, Small Model Efficiency (arxiv.org)
1 point by Anon84 on Aug 26, 2024 | hide | past | pdf | discuss
4635. A Review of Pseudo-Labeling for Computer Vision (arxiv.org)
1 point by PaulHoule on Aug 26, 2024 | hide | past | pdf | discuss
4636. Realistic Synthetic UGC: A Scaffolding Approach to Generating Online Discussions (arxiv.org)
35 points by PaulHoule on Aug 25, 2024 | hide | past | pdf | 6 comments
4637. GPT-4 is judged more human than humans in displaced and inverted Turing tests (arxiv.org)
3 points by ghita_ on Aug 25, 2024 | hide | past | pdf | 1 comment
4638. Learned Single-Pass Multitasking Perceptual Graphics for Immersive Displays (arxiv.org)
3 points by PaulHoule on Aug 25, 2024 | hide | past | pdf | discuss
4639. Rail-Only: A Low-Cost High-Performance Network for Training LLMs with T Params (arxiv.org)
2 points by edelsohn on Aug 24, 2024 | hide | past | pdf | 1 comment
4640. LLM Pruning and Distillation in Practice: The Minitron Approach (arxiv.org)
1 point by bobismyuncle on Aug 24, 2024 | hide | past | pdf | discuss
4641. The Vizier Gaussian Process Bandit Algorithm (arxiv.org)
1 point by swyx on Aug 23, 2024 | hide | past | pdf | 1 comment
4642. Higher Temperatures and Min_p Sampling (arxiv.org)
1 point by danielhanchen on Aug 23, 2024 | hide | past | pdf | 1 comment
4643. Training independent subnetworks for robust prediction (arxiv.org)
9 points by amrrs on Aug 23, 2024 | hide | past | pdf | discuss
4644. An Evaluation of Deep Learning Models for Stock Market Trend Prediction (arxiv.org)
1 point by beefman on Aug 23, 2024 | hide | past | pdf | 1 comment
4645. When Large Language Model Meets Optimization (arxiv.org)
1 point by nickpsecurity on Aug 23, 2024 | hide | past | pdf | 1 comment
4646. Equivariant Neural Networks and Piecewise Linear Representation Theory (arxiv.org)
1 point by 88888cchhcc on Aug 23, 2024 | hide | past | pdf | discuss
4647. Controlled Decoding from Language Models (arxiv.org)
1 point by vinutheraj on Aug 23, 2024 | hide | past | pdf | discuss
4648. TableBench: A Comprehensive and Complex Benchmark for Table Question Answering (arxiv.org)
2 points by mnoorfawi on Aug 23, 2024 | hide | past | pdf | discuss
4649. StructuredRAG: JSON Response Formatting with Large Language Models (arxiv.org)
38 points by bobvanluijt on Aug 22, 2024 | hide | past | pdf | 4 comments
4650. Why is this a research paper? HybridRAG = VectorRAG context and GraphRAG context (arxiv.org)
4 points by blizzard_knight on Aug 22, 2024 | hide | past | pdf | 1 comment