Google's Gemini 3.6 Flash model cuts AI agent token costs by up to 65% on long horizon engineering tasks —and 3.5 Pro is on the way
Source: VentureBeat AI
Summary
- Google's Gemini 3.6 Flash model is the latest addition to the Gemini family, which aims to make AI agents faster, smarter, and cheaper.
- The model is priced at $1.50 per million input tokens and $7.50 per million output tokens.
- This represents a significant decrease in costs compared to previous models.
- Additionally, Google has released Gemini 3.5 Flash-Lite and Gemini 3.5 Flash Cyber, which are also designed to be more efficient and cost-effective.
- The Gemini 3.5 Flash-Lite model is priced at $0.30 per million input tokens and $2.50 per million output tokens.
- Google has also announced that Gemini 3.5 Pro is on the way, which will likely be an even more advanced model.
- The release of these new models and pricing plans is part of Google's efforts to make AI more accessible and affordable for businesses and developers.
- Gemini 3.6 Flash model is a significant improvement over its predecessor, Gemini 3.1 Pro Preview, which was priced at $2 per million input tokens and $12 per million output tokens.
- This represents a decrease of up to 65% in costs, making it a more attractive option for businesses and developers.
Why It Matters
- The development of more efficient and cost-effective AI models like Gemini 3.6 Flash has significant implications for businesses and developers.
- With these models, companies can create more advanced AI agents without breaking the bank.
- This could lead to increased adoption of AI technology across various industries, from customer service to healthcare.
- Additionally, the decreased costs of AI models could lead to increased competition in the market, driving innovation and pushing companies to create even more efficient and effective AI solutions.
- The release of Gemini 3.5 Pro is also significant, as it will likely be an even more advanced model that pushes the boundaries of what is possible with AI.
- This could lead to even more exciting applications and use cases for AI technology.
GenAI EXPLAINED
Token Costs: Token costs refer to the amount of money charged by AI models for processing and generating text. In the case of Gemini 3.6 Flash, the token cost is $1.50 per million input tokens and $7.50 per million output tokens. This means that for every million characters of text that the model processes or generates, the company has to pay $1.50 and $7.50, respectively.
Input Tokens: Input tokens refer to the text that is fed into the AI model for processing. In the case of Gemini 3.6 Flash, the input token cost is $1.50 per million characters.
Output Tokens: Output tokens refer to the text that is generated by the AI model. In the case of Gemini 3.6 Flash, the output token cost is $7.50 per million characters.
MORE FROM THIS EDITION