about
211. Accelerating LLM Inference with Lossless Speculative Decoding Algorithms (2025) (arxiv.org)
Speculative decoding has a small model guess several upcoming words that a big model then checks in one pass. These methods let the two models use different word lists, keep output identical to the big model alone, and sped generation up to 2.8 times.
1 point by wslh 29 days ago | hide | past | pdf | discuss
212. StoryScope: Investigating Idiosyncrasies in AI Fiction (arxiv.org)
A tool pulls out story choices—how much say characters have, how jumbled the timeline is—to tell human from AI without style. These clues spotted AI stories 93.2%, nearly matching style-based checks; AI stories cluster in tidy single-track plots while human ones vary more.
1 point by efavdb 29 days ago | hide | past | pdf | discuss