about
5341. Phi-3 Technical Report (arxiv.org)
411 points by varunvummadi on Apr 23, 2024 | hide | past | pdf | 130 comments
5342. FPGA Architecture for Deep Learning: Survey and Future Directions (arxiv.org)
128 points by matt_d on Apr 22, 2024 | hide | past | pdf | 52 comments
5343. Analyzing the Performance of Large Language Models on Code Summarization (arxiv.org)
1 point by PaulHoule on Apr 22, 2024 | hide | past | pdf | discuss
5344. Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing (arxiv.org)
2 points by hnhn34 on Apr 22, 2024 | hide | past | pdf | discuss
5345. LLM Agents Can Autonomously Exploit One-Day Vulnerabilities (arxiv.org)
1 point by _____k on Apr 22, 2024 | hide | past | pdf | 1 comment
5346. Who Validates the Validators? Aligning LLM-Assisted Evaluation of LLM Outputs (arxiv.org)
2 points by fatso784 on Apr 22, 2024 | hide | past | pdf | discuss
5347. Survey Study on AI Agent Architectures (2024) (arxiv.org)
77 points by jslampe on Apr 22, 2024 | hide | past | pdf | 16 comments
5348. YaART: Yet Another Art Rendering Technology (arxiv.org)
2 points by avi_chawla on Apr 22, 2024 | hide | past | pdf | discuss
5349. Lossless Acceleration of Long Sequence Generation (arxiv.org)
1 point by eshoyuan on Apr 22, 2024 | hide | past | pdf | discuss
5350. Many-Shot In-Context Learning (arxiv.org)
62 points by Anon84 on Apr 22, 2024 | hide | past | pdf | 1 comment
5351. RecurrentGemma: Moving Past Transformers for Efficient Open Language Models (arxiv.org)
47 points by CharlesW on Apr 22, 2024 | hide | past | pdf | 3 comments
5352. The Illusion of State in State-Space Models (arxiv.org)
2 points by georgehill on Apr 21, 2024 | hide | past | pdf | discuss
5353. Lossless Acceleration of LLM via Adaptive N-Gram Parallel Decoding (arxiv.org)
136 points by PaulHoule on Apr 21, 2024 | hide | past | pdf | 23 comments
5354. A Comprehensive Overview of Large Language Models (arxiv.org)
2 points by rbanffy on Apr 21, 2024 | hide | past | pdf | discuss
5355. From r to Q∗: Your Language Model is a Q-Function (arxiv.org)
2 points by throwaway71271 on Apr 21, 2024 | hide | past | pdf | discuss
5356. Deep Neural Networks via Complex Network Theory: A Perspective (arxiv.org)
1 point by Anon84 on Apr 21, 2024 | hide | past | pdf | discuss
5357. Modeling Boundedly Rational Agents with Latent Inference Budgets (arxiv.org)
2 points by danboarder on Apr 21, 2024 | hide | past | pdf | discuss
5358. Measuring Multitask Language Understanding (arxiv.org)
1 point by tosh on Apr 20, 2024 | hide | past | pdf | discuss
5359. How to avoid machine learning pitfalls (arxiv.org)
7 points by golergka on Apr 20, 2024 | hide | past | pdf | discuss
5360. AgentKit: Flow Engineering with Graphs, Not Coding (arxiv.org)
1 point by kjhughes on Apr 20, 2024 | hide | past | pdf | discuss
5361. Training-Free Long-Context Scaling of Large Language Models (arxiv.org)
1 point by tosh on Apr 20, 2024 | hide | past | pdf | discuss
5362. Trillion-Parameter Sequential Transducers for Generative Recommendations (arxiv.org)
2 points by jonbaer on Apr 20, 2024 | hide | past | pdf | discuss
5363. MABC: MultiAgent Blockchain Collab for RC analysis in microservices architecture (arxiv.org)
2 points by panqueca on Apr 19, 2024 | hide | past | pdf | discuss
5364. JetMoE: Reaching Llama2 Performance with 0.1M Dollars (arxiv.org)
1 point by PaulHoule on Apr 19, 2024 | hide | past | pdf | discuss
5365. Optimizing the Deployment of Tiny Transformers on Low-Power MCUs (arxiv.org)
1 point by PaulHoule on Apr 19, 2024 | hide | past | pdf | discuss
5366. Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing (arxiv.org)
3 points by milliondreams on Apr 19, 2024 | hide | past | pdf | discuss
5367. InstantMesh: Efficient 3D Mesh Generation from a Single Image (arxiv.org)
2 points by rrampage on Apr 19, 2024 | hide | past | pdf | discuss
5368. LLM Agents Can Autonomously Exploit One-Day Vulnerabilities with 87% Success (arxiv.org)
2 points by namanyayg on Apr 19, 2024 | hide | past | pdf | discuss
5369. Demystifying RCE Vulnerabilities in LLM-Integrated Apps (arxiv.org)
1 point by dytir on Apr 19, 2024 | hide | past | pdf | discuss
5370. Forecasting the Future: Advancements in Large Meteorological Models (arxiv.org)
2 points by PaulHoule on Apr 18, 2024 | hide | past | pdf | discuss