about
4021. 2022 Recommendation System paper from ByteDance – the parent company of TikTo (arxiv.org)
1 point by alexcos on Dec 16, 2024 | hide | past | pdf | 1 comment
4022. GenEx: Generating an Explorable World (arxiv.org)
1 point by pr337h4m on Dec 16, 2024 | hide | past | pdf | discuss
4023. Boundless Socratic Learning with Language Games (arxiv.org)
1 point by enoch2090 on Dec 16, 2024 | hide | past | pdf | discuss
4024. Competitive debate evidence dataset for LLM persuasion (arxiv.org)
2 points by Der_Einzige on Dec 16, 2024 | hide | past | pdf | discuss
4025. Gaze-LLE: Gaze Target Estimation via Large-Scale Learned Encoders (arxiv.org)
2 points by asah on Dec 15, 2024 | hide | past | pdf | discuss
4026. Evaluating Emerging AI/ML Accelerators: IPU, RDU and Nvidia/AMD GPUs (arxiv.org)
1 point by teleforce on Dec 14, 2024 | hide | past | pdf | 1 comment
4027. Improving training time and GPU utilization in geo-distributed LLM training (arxiv.org)
1 point by PaulHoule on Dec 14, 2024 | hide | past | pdf | discuss
4028. SlimLM: An Efficient Small Language Model for On-Device Document Assistance (arxiv.org)
3 points by PaulHoule on Dec 14, 2024 | hide | past | pdf | discuss
4029. Specifications: The missing link to make development of LLM an eng discipline (arxiv.org)
2 points by shishirpatil on Dec 14, 2024 | hide | past | pdf | discuss
4030. Fast-Splat: Fast, Ambiguity-Free Semantics Transfer in Gaussian Splatting (arxiv.org)
2 points by PaulHoule on Dec 14, 2024 | hide | past | pdf | discuss
4031. Towards Reasoning in Large Language Models: A Survey (2023) (arxiv.org)
1 point by rntn on Dec 14, 2024 | hide | past | pdf | discuss
4032. A Knowledge Graph and LLM-Driven Approach for Conversational Recommendation (arxiv.org)
1 point by PaulHoule on Dec 14, 2024 | hide | past | pdf | discuss
4033. Best-of-N Jailbreaking (arxiv.org)
68 points by flyingpumba on Dec 14, 2024 | hide | past | pdf | 15 comments
4034. Dank Learning: Generating Memes Using Deep Neural Networks (2018) (arxiv.org)
1 point by benatkin on Dec 14, 2024 | hide | past | pdf | discuss
4035. Phi-4 Technical Report (arxiv.org)
1 point by SerCe on Dec 14, 2024 | hide | past | pdf | discuss
4036. Planning-Driven Programming: A Large Language Model Programming Workflow (arxiv.org)
1 point by PaulHoule on Dec 13, 2024 | hide | past | pdf | discuss
4037. Best-of-N Jailbreaking (arxiv.org)
17 points by tzury on Dec 13, 2024 | hide | past | pdf | 1 comment
4038. Phi-4 Technical Report (arxiv.org)
3 points by tosh on Dec 13, 2024 | hide | past | pdf | discuss
4039. Beyond Gaussians: Fast and High-Fidelity 3D Splatting with Linear Kernels (arxiv.org)
2 points by PaulHoule on Dec 13, 2024 | hide | past | pdf | discuss
4040. Frontier Models are Capable of In-context Scheming (arxiv.org)
10 points by trott on Dec 12, 2024 | hide | past | pdf | 1 comment
4041. Forking Paths in Neural Text Generation (arxiv.org)
2 points by pizza on Dec 12, 2024 | hide | past | pdf | discuss
4042. LLM-Take: Theme-Aware Keyword Extraction Using Large Language Models (arxiv.org)
1 point by klaussilveira on Dec 12, 2024 | hide | past | pdf | discuss
4043. Antelope: Potent and Concealed Jailbreak Attack Strategy (arxiv.org)
1 point by belter on Dec 12, 2024 | hide | past | pdf | discuss
4044. Understanding Hidden Computations in Chain-of-Thought Reasoning (arxiv.org)
1 point by riemann77 on Dec 12, 2024 | hide | past | pdf | 1 comment
4045. Exceeding Conventional FPGA Roofline Limit by LUT-Based Efficient Multiplication (arxiv.org)
1 point by PaulHoule on Dec 11, 2024 | hide | past | pdf | 1 comment
4046. Hymba: A Hybrid-Head Architecture for Small Language Models (arxiv.org)
2 points by PaulHoule on Dec 11, 2024 | hide | past | pdf | discuss
4047. Memory-Aware Stream Processing for Attention Acceleration on Edge Devices (arxiv.org)
3 points by PaulHoule on Dec 11, 2024 | hide | past | pdf | discuss
4048. From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step (arxiv.org)
2 points by calebkaiser on Dec 11, 2024 | hide | past | pdf | discuss
4049. Asynchronous LLM Function Calling (arxiv.org)
2 points by omarsar on Dec 11, 2024 | hide | past | pdf | discuss
4050. Granite Guardian Models (From IBM) (arxiv.org)
1 point by omarsar on Dec 11, 2024 | hide | past | pdf | discuss