
Gemini 3.6 Flash wrote 40 rounds of fiction. Three of the four chains ended in a loop
Gemini 3.6 Flash shipped July 21, 2026 with an efficiency pitch: 17% fewer output tokens and a lower price. Within 48 hours we ran it through the same protocol as our Qwen 3.8 test: two Chinese novels, 20 consecutive continuation rounds each, paired double-blind against its predecessor 3.5 Flash. The fantasy epic went 9-3 for the new model; the palace novel collapsed to 1-10 with the late window at 0-11. Genre preference does not explain the split: three of the four chains fell into plot loops within 20 rounds, and the two generations break differently. 3.5 replays whole rounds verbatim (repetition detection hits 100%), 3.6 on the palace novel decays into a verbatim dead loop, and 3.6 on the fantasy epic shows a new failure variant in our records: semantic-level rewinding that replays the same story beat in fresh wording six or seven times while every mechanical metric stays green. The structural statistics sided with the loser on both books, and the prompt-wording fix that once rescued gemini-3.1-pro does not reach this disease. Measured: about 6 seconds per round vs 12.5, $3.19 for the whole experiment, all pinned to a 2026-07-23 aggregator endpoint.
Read the post →







































































