Bring your own key: Google Gemini's API rates, no markup. Keys and chats stay in your browser. Run Google Gemini in parallel with other models, then compare and merge the answers.
Google gives Gemini a genuinely free tier, a million-token context window, and native image generation. Big-AGI runs all of it on your own key with the controls the Gemini app keeps to itself. Set the thinking level, ground answers in Google Search with a recency window, tune media resolution, or generate images at a chosen aspect ratio and size. No card to begin.
It reaches the newest previews too: Deep Research agents that return long-form reports, and Omni Flash, which turns a prompt into a short video with sound. Then beam Gemini's long-context read against Claude and GPT and keep what they agree on.
Nano Banana 2 Lite
NEWGemini 3.1 Flash Lite Image.
131K
$0.25
$1.5
Jun 2026
Gemini Omni Flash Preview (video)
NEWGemini Omni Flash Preview
197K
$1.5
$17.5
Jun 2026
Nano Banana 2
Gemini 3.1 Flash Image.
131K
$0.5
$3
May 2026
Nano Banana Pro
Gemini 3 Pro Image
164K
$2
$12
May 2026
Gemini 3.5 Flash
Gemini 3.5 Flash
1.1M
$1.5
$9
May 2026
Antigravity Agent Preview (2026-05)
Preview release of Antigravity Agent (05-2026)
197K
$1.5
$9
May 2026
Gemini 3.1 Flash-Lite
Gemini 3.1 Flash Lite
1.1M
$0.25
$1.5
May 2026
Deep Research Max Preview (2026-04)
Preview release (April 21st, 2026) of Deep Research Max
197K
$1.25
$10
Apr 2026
Deep Research Preview (2026-04)
Preview release (April 21th, 2026) of Deep Research
197K
$1.25
$10
Apr 2026
Gemini 3.1 Flash TTS Preview
Gemini 3.1 Flash TTS Preview
25K
$1
-
Apr 2026
Gemini Robotics-ER 1.6 Preview
Gemini Robotics-ER 1.6 Preview
197K
$1
$5
Apr 2026
Gemma 4 26B A4B IT
Gemma 4 26B A4B IT
295K
-
-
Apr 2026
Gemma 4 31B IT
Gemma 4 31B IT
295K
-
-
Apr 2026
Gemini 3.1 Flash-Lite Preview
Gemini 3.1 Flash Lite Preview
1.1M
$0.25
$1.5
Mar 2026
Nano Banana 2 Preview
Gemini 3.1 Flash Image Preview.
131K
$0.5
$3
Feb 2026
Gemini 3.1 Pro Preview
Gemini 3.1 Pro Preview
1.1M
$2
$12
Feb 2026
Gemini 3.1 Pro Preview (Custom Tools)
Gemini 3.1 Pro Preview optimized for custom tool usage
1.1M
$2
$12
Feb 2026
Gemini 3 Flash Preview
Gemini 3 Flash Preview
1.1M
$0.5
$3
Dec 2025
Deep Research Pro Preview
Preview release (December 12th, 2025) of Deep Research Pro
197K
$1.25
$10
Dec 2025
Nano Banana Pro
Gemini 3 Pro Image Preview
164K
$2
$12
Nov 2025
Nano Banana Pro Preview
Gemini 3 Pro Image Preview
164K
$2
$12
Nov 2025
Gemini 2.5 Computer Use Preview 10-2025
Gemini 2.5 Computer Use Preview 10-2025
197K
$1.25
$10
Oct 2025
Nano Banana
Gemini 2.5 Flash Preview Image
66K
$0.3
-
Oct 2025
Gemini 2.5 Flash-Lite
Stable version of Gemini 2.5 Flash-Lite, released in July of 2025
1.1M
$0.1
$0.4
Jul 2025
Gemini 2.5 Flash
Stable version of Gemini 2.5 Flash, our mid-size multimodal model that supports up to 1 million tokens, released in June of 2025.
1.1M
$0.3
$2.5
Jun 2025
Gemini 2.5 Pro
Stable release (June 17th, 2025) of Gemini 2.5 Pro
1.1M
$1.25
$10
Jun 2025
Gemini 2.5 Flash Preview TTS
Gemini 2.5 Flash Preview TTS
25K
$0.5
-
May 2025
Gemini 2.5 Pro Preview TTS
Gemini 2.5 Pro Preview TTS
25K
$1
-
May 2025
Gemini 2.0 Flash 001
Stable version of Gemini 2.0 Flash, our fast and versatile multimodal model for scaling across diverse tasks, released in January of 2025.
1.1M
$0.1
$0.4
Feb 2025
Nano Banana 2 Lite
NEWGemini 3.1 Flash Lite Image.
Gemini Omni Flash Preview (video)
NEWGemini Omni Flash Preview
Nano Banana 2
Gemini 3.1 Flash Image.
Nano Banana Pro
Gemini 3 Pro Image
Gemini 3.5 Flash
Gemini 3.5 Flash
Antigravity Agent Preview (2026-05)
Preview release of Antigravity Agent (05-2026)
Gemini 3.1 Flash-Lite
Gemini 3.1 Flash Lite
Deep Research Max Preview (2026-04)
Preview release (April 21st, 2026) of Deep Research Max
Deep Research Preview (2026-04)
Preview release (April 21th, 2026) of Deep Research
Gemini 3.1 Flash TTS Preview
Gemini 3.1 Flash TTS Preview
Gemini Robotics-ER 1.6 Preview
Gemini Robotics-ER 1.6 Preview
Gemma 4 26B A4B IT
Gemma 4 26B A4B IT
Gemma 4 31B IT
Gemma 4 31B IT
Gemini 3.1 Flash-Lite Preview
Gemini 3.1 Flash Lite Preview
Nano Banana 2 Preview
Gemini 3.1 Flash Image Preview.
Gemini 3.1 Pro Preview
Gemini 3.1 Pro Preview
Gemini 3.1 Pro Preview (Custom Tools)
Gemini 3.1 Pro Preview optimized for custom tool usage
Gemini 3 Flash Preview
Gemini 3 Flash Preview
Deep Research Pro Preview
Preview release (December 12th, 2025) of Deep Research Pro
Nano Banana Pro
Gemini 3 Pro Image Preview
Nano Banana Pro Preview
Gemini 3 Pro Image Preview
Gemini 2.5 Computer Use Preview 10-2025
Gemini 2.5 Computer Use Preview 10-2025
Nano Banana
Gemini 2.5 Flash Preview Image
Gemini 2.5 Flash-Lite
Stable version of Gemini 2.5 Flash-Lite, released in July of 2025
Gemini 2.5 Flash
Stable version of Gemini 2.5 Flash, our mid-size multimodal model that supports up to 1 million tokens, released in June of 2025.
Gemini 2.5 Pro
Stable release (June 17th, 2025) of Gemini 2.5 Pro
Gemini 2.5 Flash Preview TTS
Gemini 2.5 Flash Preview TTS
Gemini 2.5 Pro Preview TTS
Gemini 2.5 Pro Preview TTS
Gemini 2.0 Flash 001
Stable version of Gemini 2.0 Flash, our fast and versatile multimodal model for scaling across diverse tasks, released in January of 2025.
1
Create an API key at the Google Gemini console.
2
Paste it into Big-AGI's model settings.
3
Start chatting, or Beam it against other models and fuse the answers.
Add your Google AI Studio key and use Gemini models at Google's own API rates. Big-AGI isn't a reseller: no markup and no intermediary, so Google bills you directly.
The Gemini app is Google's model in Google's box. Big-AGI runs Gemini next to GPT, Claude, Grok, and anything else you've connected, side by side in one conversation, which the Gemini app isn't built to do.
It also surfaces controls the consumer app keeps hidden: Gemini's API lets you set a thinking budget and temperature per request, real parameters Google ships, not settings we made up. Your keys stay in your browser instead of your Google account.
Billing shifts from monthly to per-token too. A light user can spend less in a month than a subscription costs, and a heavy user never runs into a usage cap, only a bill for tokens actually used. Beam parallel runs do use more tokens than a single reply, that part isn't free.
Turn on Direct Connection under Google's Advanced settings and the browser calls Google directly, bypassing the Big-AGI server, whenever your key is client-side and Google's API allows it. Your keys stay in your browser. Chats are stored locally first and sync only if you turn it on. The AI Inspector shows the exact request before it goes out: model, parameters, token counts, and a cost estimate.
Beam runs Gemini against GPT, Claude, or any other model you've connected, in parallel, then hands you Fusions: several strategies that combine, cross-check, and synthesize the parallel answers, which beats just picking the best one. Parallel runs use more tokens than a single chat.
Your key, your data, your choice of model. Big-AGI is open source and self-hostable, so you can check exactly how Google Gemini is called.
BIG-AGI
Resources
© 2026 Token Fabrics·Built with passion in San Diego