gemini-3.1-flash-lite is currently Google's latest and most cost-effective model, optimized for large-scale agent-based tasks, translation, and simple data processing.
gemini-3.1-flash-lite is currently Google's latest and most cost-effective model, optimized for large-scale agent-based tasks, translation, and simple data processing.
gemini-3.1-flash-lite-nothink has a 1,048,576 token context window. It supports up to 65,536 output tokens.
On AIHubMix, gemini-3.1-flash-lite-nothink costs $0.25 per million input tokens and $1.5 per million output tokens.
gemini-3.1-flash-lite-nothink accepts text, image, video, audio and PDF input.
gemini-3.1-flash-lite-nothink supports thinking, tool calling, function calling, structured outputs, web search, deep search and long context. Per-protocol parameter support is listed in the capability table on this page.
gemini-3.1-flash-lite-nothink is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to gemini-3.1-flash-lite-nothink — no other code changes needed.
gemini-3.1-flash-lite-nothink is developed by Google. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
Gemini 3.8 Flash is Google's most intelligent Flash-series model, designed for…
Gemini 3.8 Flash free version: Free model resources are limited and provided only for…
Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web…
Gemini 3.7 Flash free version: Free model resources are limited and provided only for…
Gemini 3.6 Flash provides sustained frontier-level intelligence optimized for real-world…
Google's newest, most compact, and most cost-effective image generation and editing…
Use gemini-3.1-flash-lite-nothink via the AIHubMix unified API — one interface for every major LLM.