about
Stories from April 6, 2026 (UTC)
Go back a day, month, or year. Go forward a day.
1. Optimizing Time, Cost, and Generalization in Distributed Large-Batch Training (arxiv.org)
2 points by PaulHoule 180 days ago | hide | past | pdf | discuss
2. Model2Kernel: Model-Aware Symbolic Execution for Safe CUDA Kernels (arxiv.org)
2 points by PaulHoule 180 days ago | hide | past | pdf | discuss
3. Show HN: WebGPU LLM inference comprehensive benchmark (arxiv.org)
2 points by yu3zhou4 181 days ago | hide | past | pdf | 2 comments
4. Test-Time Scaling Makes Overtraining Compute-Optimal (arxiv.org)
1 point by matt_d 180 days ago | hide | past | pdf | discuss
5. Analyzing Reverse Address Translation Overheads in Multi-GPU Scale-Up Pods (arxiv.org)
1 point by matt_d 180 days ago | hide | past | pdf | discuss
6. Provider-Dependent Energy Effects of Prompt Compression (arxiv.org)
1 point by PaulHoule 180 days ago | hide | past | pdf | discuss
7. Benchmark-Dependent Output Dynamics in LLM Prompt Compression (arxiv.org)
1 point by PaulHoule 181 days ago | hide | past | pdf | discuss