1 comment

[ 2.7 ms ] story [ 15.6 ms ] thread
"When not using reasoning, repeating the input prompt improves performance for popular models (Gemini, GPT, Claude, and Deepseek) without increasing the number of generated tokens or latency."