New AI University AI Topics
← AI News

Google's Gemini 3.6 Flash model cuts AI agent token costs by up to 65% on long horizon engineering tasks —and 3.5 Pro is on the way

Source: VentureBeat AI

Summary

  • Google's Gemini 3.6 Flash model is the latest addition to the Gemini family, which aims to make AI agents faster, smarter, and cheaper.
  • The model is priced at $1.50 per million input tokens and $7.50 per million output tokens.
  • This represents a significant decrease in costs compared to previous models.
  • Additionally, Google has released Gemini 3.5 Flash-Lite and Gemini 3.5 Flash Cyber, which are also designed to be more efficient and cost-effective.
  • The Gemini 3.5 Flash-Lite model is priced at $0.30 per million input tokens and $2.50 per million output tokens.
  • Google has also announced that Gemini 3.5 Pro is on the way, which will likely be an even more advanced model.
  • The release of these new models and pricing plans is part of Google's efforts to make AI more accessible and affordable for businesses and developers.
  • Gemini 3.6 Flash model is a significant improvement over its predecessor, Gemini 3.1 Pro Preview, which was priced at $2 per million input tokens and $12 per million output tokens.
  • This represents a decrease of up to 65% in costs, making it a more attractive option for businesses and developers.

Why It Matters

  • The development of more efficient and cost-effective AI models like Gemini 3.6 Flash has significant implications for businesses and developers.
  • With these models, companies can create more advanced AI agents without breaking the bank.
  • This could lead to increased adoption of AI technology across various industries, from customer service to healthcare.
  • Additionally, the decreased costs of AI models could lead to increased competition in the market, driving innovation and pushing companies to create even more efficient and effective AI solutions.
  • The release of Gemini 3.5 Pro is also significant, as it will likely be an even more advanced model that pushes the boundaries of what is possible with AI.
  • This could lead to even more exciting applications and use cases for AI technology.

GenAI EXPLAINED

Token Costs: Token costs refer to the amount of money charged by AI models for processing and generating text. In the case of Gemini 3.6 Flash, the token cost is $1.50 per million input tokens and $7.50 per million output tokens. This means that for every million characters of text that the model processes or generates, the company has to pay $1.50 and $7.50, respectively.

Input Tokens: Input tokens refer to the text that is fed into the AI model for processing. In the case of Gemini 3.6 Flash, the input token cost is $1.50 per million characters.

Output Tokens: Output tokens refer to the text that is generated by the AI model. In the case of Gemini 3.6 Flash, the output token cost is $7.50 per million characters.