Google envia Gemini 3.6 Flash com 17% menos tokens, preço de $7.50; 3.5 Pro ainda atrasado
Google lançou Gemini 3.6 Flash em 21 de julho de 2026 como seu modelo workhorse para agentic coding, trabalho conhecimento, e tarefas multimodal. Preço token output caiu de $9.00 para $7.50 por milhão (entrada mantida em $1.50), e o modelo consome 17% menos output tokens por tarefa que Gemini 3.5 Flash no Índice Artificial Analysis, alcançando ganhos eficiência 65% em benchmarks agentic específicos como DeepSWE. O corte preço adesivo combinado mais eficiência token se compõe para redução custo efetivo 31% por tarefa completada, com até 71% poupança em workloads agentic coding.
Google também enviou Gemini 3.5 Flash-Lite em $0.30/$2.50 por milhão tokens para tarefas alto-throughput e anunciou Gemini 3.5 Flash Cyber, modelo tuned-segurança restrito acesso governo e parceiros. Corte conhecimento Mço 2026 representa avanço 14 meses sobre 3.5 Flash janeiro 2025 data. Em benchmarks agentic aplicados, 3.6 Flash ganhou terreno: DeepSWE melhorou 37% para 49%, MLE Bench 49.7% para 63.9%, OSWorld computer-use 78.4% para 83.0%.
Contudo, Índice Artificial Analysis Intelligence independente, Gemini 3.6 Flash pontua exatamente 50—inalterado 3.5 Flash. Isto eficiência release, não salto capacidade. O modelo negocia razão bruta poder eficiência token e latência menor, otimizando econômica específica workflows agentic em vez razão profundidade fronteira. Artificial Analysis independentemente mediu tempo completão tarefa média caindo 2.7 minutos 1.3 minutos e custo médio tarefa $0.59 para $0.50.
Notably ausente: Gemini 3.5 Pro permanece teste de parceiro sem data disponição pública, tendo perdido promessa I/O maio 2026 e alvo posterior junho. Enquanto Google envia três modelos, flagship 3.5 geração é travado. Pré-treino Gemini 4 iniciado e enviará depois. Para equipes executando inferência agentic alto-volume, preço 3.6 Flash e ganhos eficiência importam; para aqueles avaliando capacidade flagship, linha entrega Google continua atraso Anthropic OpenAI.
Fontes
- Primary source
- 9to5google.com
“Gemini 3.6 Flash consumes 17% fewer output tokens compared to 3.5 Flash, while taking fewer reasoning steps and tool calls to accomplish multi-step workflows, and is priced lower at $1.50/1M input tokens and $7.50/1M output tokens”
- trilogyai.substack.com
“Gemini 3.6 Flash pricing is $1.50/$7.50 per million tokens — but token efficiency gains mean the effective cost per completed task drops ~31%, and up to ~71% on agentic coding workloads”
- felloai.com
“On the Artificial Analysis Intelligence Index it scores around 50 which is above the field average, but roughly flat against 3.5 Flash”
- 9to5google.com
“Gemini 3.6 Flash scores 49 percent on DeepSWE benchmark, a notable increase from the 37 percent achieved by version 3.5, and pushes machine learning engineering performance higher, scoring 63.9 percent on MLE-Bench compared to 49.7 percent previously”
- techcrunch.com
“Google teased the release of Pro as part of the 3.5 Flash release in May, saying the Pro version was already being used internally, and we look forward to rolling it out next month. Last week, Bloomberg reported that Google was facing internal delays in launching the 3.5 Pro”