In plain words: Simply writing the same prompt more than once before asking for an answer makes popular chat models answer better when they are not thinking step by step. It beats the usual single prompt without generating extra words or taking longer.
Abstract · Prompt Repetition Improves Non-Reasoning LLMs
When not using reasoning, repeating the input prompt improves performance for popular models (Gemini, GPT, Claude, and Deepseek) without increasing the number of generated tokens or latency.
Yaniv Leviathan, Matan Kalman, Yossi Matias
arXiv:2512.14982 · cs.LG, cs.AI, cs.CL · submitted Dec 17, 2025
abstract · pdf · html