| 4621. |
Choosing the "Brain" for your AI-powered app – My new method, feedback requested (arxiv.org) |
|
2 points by robert_sim on Aug 29, 2024 | hide | past | pdf | 1 comment
|
| 4622. |
Let's Verify Step by Step (arxiv.org) |
|
1 point by sharemywin on Aug 29, 2024 | hide | past | pdf | 2 comments
|
| 4623. |
The Mamba in the Llama (arxiv.org) |
|
2 points by dpstart01 on Aug 29, 2024 | hide | past | pdf | discuss
|
| 4624. |
Chain and Hash, an LLM Fingerprinting Technique (arxiv.org) |
|
2 points by wslh on Aug 29, 2024 | hide | past | pdf | discuss
|
| 4625. |
Diffusion Models Are Real-Time Game Engines (arxiv.org) |
|
2 points by LordNibbler on Aug 29, 2024 | hide | past | pdf | discuss
|
| 4626. |
Implicit Bias Matters for Language Models (arxiv.org) |
|
2 points by fzliu on Aug 28, 2024 | hide | past | pdf | discuss
|
| 4627. |
CRQBench: A Benchmark of Code Reasoning Questions (arxiv.org) |
|
1 point by PaulHoule on Aug 28, 2024 | hide | past | pdf | discuss
|
| 4628. |
Multilingual Multimodal Data Hub and Benchmark for Southeast Asian Languages (arxiv.org) |
|
2 points by tellarin on Aug 28, 2024 | hide | past | pdf | discuss
|
| 4629. |
Sapiens: Foundation for Human Vision Models (arxiv.org) |
|
58 points by soulofmischief on Aug 28, 2024 | hide | past | pdf | 1 comment
|
| 4630. |
Learning to Move Like Professional Counter-Strike Players (arxiv.org) |
|
4 points by SerCe on Aug 28, 2024 | hide | past | pdf | discuss
|
| 4631. |
LongWriter: Unleashing 10k Word Generation from Long Context LLMs (arxiv.org) |
|
7 points by PaulHoule on Aug 28, 2024 | hide | past | pdf | 1 comment
|
| 4632. |
The Geometry of the Set of Equivalent Linear Neural Networks (arxiv.org) |
|
1 point by 082349872349872 on Aug 27, 2024 | hide | past | pdf | discuss
|
| 4633. |
NeuroPapyri: A Deep Attention Embedding Network for Handwritten Papyri Retrieval (arxiv.org) |
|
1 point by PaulHoule on Aug 26, 2024 | hide | past | pdf | discuss
|
| 4634. |
Large Model Strategic Thinking, Small Model Efficiency (arxiv.org) |
|
1 point by Anon84 on Aug 26, 2024 | hide | past | pdf | discuss
|
| 4635. |
A Review of Pseudo-Labeling for Computer Vision (arxiv.org) |
|
1 point by PaulHoule on Aug 26, 2024 | hide | past | pdf | discuss
|
| 4636. |
Realistic Synthetic UGC: A Scaffolding Approach to Generating Online Discussions (arxiv.org) |
|
35 points by PaulHoule on Aug 25, 2024 | hide | past | pdf | 6 comments
|
| 4637. |
GPT-4 is judged more human than humans in displaced and inverted Turing tests (arxiv.org) |
|
3 points by ghita_ on Aug 25, 2024 | hide | past | pdf | 1 comment
|
| 4638. |
Learned Single-Pass Multitasking Perceptual Graphics for Immersive Displays (arxiv.org) |
|
3 points by PaulHoule on Aug 25, 2024 | hide | past | pdf | discuss
|
| 4639. |
Rail-Only: A Low-Cost High-Performance Network for Training LLMs with T Params (arxiv.org) |
|
2 points by edelsohn on Aug 24, 2024 | hide | past | pdf | 1 comment
|
| 4640. |
LLM Pruning and Distillation in Practice: The Minitron Approach (arxiv.org) |
|
1 point by bobismyuncle on Aug 24, 2024 | hide | past | pdf | discuss
|
| 4641. |
The Vizier Gaussian Process Bandit Algorithm (arxiv.org) |
|
1 point by swyx on Aug 23, 2024 | hide | past | pdf | 1 comment
|
| 4642. |
Higher Temperatures and Min_p Sampling (arxiv.org) |
|
1 point by danielhanchen on Aug 23, 2024 | hide | past | pdf | 1 comment
|
| 4643. |
Training independent subnetworks for robust prediction (arxiv.org) |
|
9 points by amrrs on Aug 23, 2024 | hide | past | pdf | discuss
|
| 4644. |
An Evaluation of Deep Learning Models for Stock Market Trend Prediction (arxiv.org) |
|
1 point by beefman on Aug 23, 2024 | hide | past | pdf | 1 comment
|
| 4645. |
When Large Language Model Meets Optimization (arxiv.org) |
|
1 point by nickpsecurity on Aug 23, 2024 | hide | past | pdf | 1 comment
|
| 4646. |
Equivariant Neural Networks and Piecewise Linear Representation Theory (arxiv.org) |
|
1 point by 88888cchhcc on Aug 23, 2024 | hide | past | pdf | discuss
|
| 4647. |
Controlled Decoding from Language Models (arxiv.org) |
|
1 point by vinutheraj on Aug 23, 2024 | hide | past | pdf | discuss
|
| 4648. |
TableBench: A Comprehensive and Complex Benchmark for Table Question Answering (arxiv.org) |
|
2 points by mnoorfawi on Aug 23, 2024 | hide | past | pdf | discuss
|
| 4649. |
StructuredRAG: JSON Response Formatting with Large Language Models (arxiv.org) |
|
38 points by bobvanluijt on Aug 22, 2024 | hide | past | pdf | 4 comments
|
| 4650. |
Why is this a research paper? HybridRAG = VectorRAG context and GraphRAG context (arxiv.org) |
|
4 points by blizzard_knight on Aug 22, 2024 | hide | past | pdf | 1 comment
|
| More |