Use DeepSeek Models in Big-AGI.

Bring your own key: DeepSeek's API rates, no markup. Keys and chats stay in your browser. Run DeepSeek in parallel with other models, then compare and merge the answers.

DeepSeek V4 Pro
DeepSeek V4 Flash

All supported DeepSeek models

ModelContextInputOutputReleased

DeepSeek V4 Pro

ReasoningTools / functions

Premium reasoning model with 1M context. Supports extended thinking modes, JSON output, and function calling.

1M

$0.44

$0.87

Apr 2026

DeepSeek V4 Flash

ReasoningTools / functions

Fast general-purpose model with 1M context. Supports extended thinking modes, JSON output, and function calling.

1M

$0.14

$0.28

Apr 2026

DeepSeek V4 Pro

Apr 2026

Premium reasoning model with 1M context. Supports extended thinking modes, JSON output, and function calling.

ReasoningTools / functions
1M · in $0.44 · out $0.87

DeepSeek V4 Flash

Apr 2026

Fast general-purpose model with 1M context. Supports extended thinking modes, JSON output, and function calling.

ReasoningTools / functions
1M · in $0.14 · out $0.28
2 models · sorted by release date · prices in USD per 1M tokens · refreshed every 30 minutesCompare every model across vendors →

Get started in 3 steps

1

Create an API key at the DeepSeek console.

2

Paste it into Big-AGI's model settings.

3

Start chatting, or Beam it against other models and fuse the answers.

Running DeepSeek in Big-AGI

Add your DeepSeek API key and use DeepSeek's models at DeepSeek's own API rates. Big-AGI adds no markup and no intermediary: billing runs directly between you and DeepSeek, and your keys stay in your browser.

  • Your key, your billing. Usage is billed by DeepSeek to your account. Big-AGI does not meter or charge for model usage.
  • Reasoning, visible. Big-AGI renders DeepSeek's reasoning output in full, so you can follow how a model reached its answer, not just read the final line.
  • Old models don't go stale. When DeepSeek retires a model, Big-AGI tracks the alias through to its successor and surfaces the retirement date, so you're never left calling a dead model.

Why not just the DeepSeek app?

The DeepSeek app is free and fine for a quick question, but it only ever shows you DeepSeek's answer. Beam sends the same prompt to DeepSeek and to GPT, Claude, or Gemini at once, so agreement or disagreement across labs becomes a signal you can act on, not a guess. You also get parameters the app hides (temperature, system prompt, per-turn model swaps), and your own key stays in your browser instead of sitting on DeepSeek's servers.

Your keys and your data

Turn on Direct Connection and the browser calls DeepSeek directly, bypassing the Big-AGI server, whenever your key is client-side and DeepSeek allows it. Your keys stay in your browser. Chats are stored locally first and sync only if you turn it on. The AI Inspector opens on any message to show the exact request sent to DeepSeek, the token counts, and a cost estimate for that call.

DeepSeek in Beam

Run DeepSeek in parallel with GPT, Claude, and Gemini on the same prompt, then reach for Fusions: several strategies that combine, cross-check, and synthesize the parallel answers, which beats just picking the single best one. Parallel runs use more tokens than a single chat.

Bring your DeepSeek key. Keep control.

Your key, your data, your choice of model. Big-AGI is open source and self-hostable, so you can check exactly how DeepSeek is called.

© 2026 Token Fabrics·Built with passion in San Diego