Last reviewed
Correct answer: C. A faster, lighter variant
Explanation
Google's line-up is built around this trade-off. In August 2026 the Gemini API ships Flash models like Gemini 3.7 Flash, lighter Flash-Lite models like Gemini 3.5 Flash-Lite, and a Pro tier, Gemini 3.1 Pro. Google calls 3.5 Flash-Lite its fastest, most cost-effective 3.5 model for high-throughput execution, and prices it at $0.30 per million input tokens against $1.50 for 3.7 Flash once introductory pricing ends. For a short classification, or a million routine calls a day, that gap is the argument.
Lighter is not the same as worse. Moving up buys deeper reasoning, stronger coding and larger thinking budgets; Google pitches Gemini 3.1 Pro at advanced intelligence and complex problem-solving. You pay for that in latency and money on every call, including the ones that never needed it.
The other options fail on their own terms. Picking the slowest variant on purpose spends time and budget for nothing. Every general Gemini model handles text, so a variant with no text capability describes nothing in the line-up. Refusing to use a model at all misreads why the light tiers exist. Check the models page before committing; these names turn over quickly.
Sources
“Our fastest, most cost-effective 3.5 model for high-throughput execution.”
“Our most cost-efficient GA model, optimized for high-volume agentic tasks, translation, and simple data processing.”
“Our most capable Flash model, built for complex coding, agentic workflows, and reliable multi-step execution.”
Practise 2 questions on this topic
Take Gemini Models — Timed Test (2 questions) — scored instantly, explanation for every question, no login.