Artificial Analysis on X: "With high reasoning, Gemini 3.8 Flash averages 48k output tokens per task, a 30% increase compared to 3.7 Flash. On low reasoning, Gemini 3.8 Flash outputs an average of 14k tokens per task, just under GP-5.6 Sol with max reasoning (17k)"