Cache-heavy agent loop cost
200K cached + 20K fresh input + 10K output tokens
Gemini 3.5 Flash
Gemini 3.5 Flash has the lower estimated token cost for this stated workload. GPT-4.1 has no published cached-input rate, so cached tokens use its listed input rate.