One of the cheapest 1M-context models — built for cheap, high-volume calls.
Gemini 3.1 Flash Lite is Google's budget model: extremely cheap, fast, and still carrying a 1M-token context. Ideal for the high-volume, low-complexity work in an agent loop where a flagship would be wasteful.
Cheap, high-volume steps that still need long context.
The model picks the moves; the agent runs the loop, the tools, and the guardrails. Once you've chosen a model, see which agent gets the most out of it.
Every model on this page is included with Standard Compute