Use Meta Models in Big-AGI.

Muse Spark on your own key: a million-token context, reasoning effort from minimal to xhigh, web search and Muse Image - with the Contributor models that train on your prompts hidden until you opt in.

All supported Meta models

ModelContextInputOutputReleased

Muse Spark 1.3

NEW
VisionReasoningTools / functionsWeb search

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of informati…

1M

$1.25

$4.25

Sep 2026

Muse Spark 1.3 Contributor

NEW
VisionReasoningTools / functionsWeb search

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic,…

1M

$0.10

$0.20

Sep 2026

Muse Spark 1.2 Contributor

NEW
VisionReasoningTools / functionsWeb search

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully chea…

1M

$0.10

$0.20

Aug 2026

Muse Glimmer 30B

NEW
VisionReasoningTools / functionsWeb search

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on co…

131K

$0.30

$1.10

Aug 2026

Muse Spark 1.2

NEW
VisionReasoningTools / functionsWeb search

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and…

1M

$1.25

$4.25

Aug 2026

Muse Spark 1.1

VisionReasoningTools / functionsWeb search

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, w…

1M

$1.25

$4.25

Jul 2026

Llama 4 Maverick 17B 128E Instruct Nvfp4

Vision

Meta chat model. https://huggingface.co/api/models/RedHatAI/Llama-4-Maverick-17B-128E-Instruct-NVFP4

1M

-

-

Jun 2026

Llama 4 Scout 17B 16E Instruct Fp8 Lora

Vision

Meta chat model.

10M

-

-

May 2026

Llama 3.3 70B Instruct FP8 Lora

Meta chat model.

131K

-

-

May 2026

Llama 4 Scout (17Bx16E)

Vision

Meta chat model. https://huggingface.co/api/models/meta-llama/Llama-4-Scout-17B-16E

262K

-

-

Jun 2025

Llama Guard 4 12B

VisionWeb search

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be use…

164K

$0.18

$0.18

Apr 2025

Llama 3.1 405B

Meta chat model. https://huggingface.co/api/models/meta-llama/Llama-3.1-405B

131K

-

-

Apr 2025

Llama 3.2 1B

Meta chat model. https://huggingface.co/api/models/meta-llama/Llama-3.2-1B

131K

-

-

Apr 2025

Llama 3.1 70B

Meta chat model. https://huggingface.co/api/models/meta-llama/Llama-3.1-70B

131K

-

-

Apr 2025

Llama 4 Scout

VisionTools / functionsWeb search

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It su…

1.3M

$0.10

$0.30

Apr 2025

Llama 4 Maverick

VisionTools / functionsWeb search

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts…

1M

$0.20

$0.70

Apr 2025

Llama 4 Scout Instruct (17Bx16E)

deprecated

Meta chat model. https://huggingface.co/meta-llama/Llama-4-Scout-17B-16E-Instruct

1M

$0.18

$0.59

Apr 2025

meta-llama/Llama-2-7b-chat-hf

Meta chat model. https://huggingface.co/meta-llama/Llama-2-7b-chat-hf

4K

-

-

Apr 2025

Meta Llama 3.1 8B Instruct Turbo

deprecated

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3.1-8B-Instruct

131K

$0.18

$0.18

Mar 2025

Llama 3.3 70B Instruct

Tools / functionsWeb search

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 inst…

131K

$0.10

$0.32

Dec 2024

Meta Llama 3.1 405B Instruct

deprecated

Meta chat model. https://huggingface.co/meta-llama/Llama-3.1-405B-Instruct

4K

$3.50

$3.50

Dec 2024

Meta Llama 3.3 70B Instruct Turbo

Meta chat model. https://huggingface.co/meta-llama/Llama-3.3-70B-Instruct

131K

$1.04

$1.04

Dec 2024

Llama 3.2 90B Vision

Vision

Llama large vision model. NVIDIA serves a reduced 32K context (native: 128K).

33K

Free

Free

Sep 2024

Llama 3.2 11B Vision

Vision

Llama vision model for image understanding.

131K

Free

Free

Sep 2024

Llama 3.2 1B

Tools / functions

Tiny Llama for edge-class tasks.

131K

Free

Free

Sep 2024

Llama 3.2 3B

Tools / functions

Small Llama for lightweight tasks.

131K

Free

Free

Sep 2024

Llama 3.1 8B

Tools / functions

Small fast Llama for utility tasks, with tool calling. Retires on NVIDIA 2026-08-25.

131K

Free

Free

Jul 2024

Llama 3.1 70B

Tools / functions

Meta Llama 3.1 70B instruction-tuned (superseded by Llama 3.3 70B).

131K

Free

Free

Jul 2024

Meta Llama 3.1 70B Instruct Turbo

deprecated

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3.1-70B-Instruct

131K

$0.88

$0.88

Jul 2024

Meta Llama 3 8B Instruct Reference

deprecated

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct

8K

$0.20

$0.20

Apr 2024

Llama 3 8B Instruct

deprecated
Web search

Meta's latest class of model (Llama 3) launched with a variety of sizes & flavors. This 8B instruct-tuned version was optimized for high quality dialogue useca…

8K

$0.14

$0.14

Apr 2024

Meta Llama 3 8B Instruct

deprecated

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct

8K

$0.20

$0.20

Apr 2024

Meta Llama 3 70B Instruct Turbo

deprecated

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3-70B-Instruct

8K

$0.88

$0.88

-

Meta Llama 3 8B Instruct Lite

deprecated

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct

8K

$0.14

$0.14

-

Muse Spark 1.3

NEW
Sep 2026

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of informati…

VisionReasoningTools / functionsWeb search
1M · in $1.25 · out $4.25

Muse Spark 1.3 Contributor

NEW
Sep 2026

Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic,…

VisionReasoningTools / functionsWeb search
1M · in $0.10 · out $0.20

Muse Spark 1.2 Contributor

NEW
Aug 2026

Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully chea…

VisionReasoningTools / functionsWeb search
1M · in $0.10 · out $0.20

Muse Glimmer 30B

NEW
Aug 2026

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on co…

VisionReasoningTools / functionsWeb search
131K · in $0.30 · out $1.10

Muse Spark 1.2

NEW
Aug 2026

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and…

VisionReasoningTools / functionsWeb search
1M · in $1.25 · out $4.25

Muse Spark 1.1

Jul 2026

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, w…

VisionReasoningTools / functionsWeb search
1M · in $1.25 · out $4.25

Llama 4 Maverick 17B 128E Instruct Nvfp4

Jun 2026

Meta chat model. https://huggingface.co/api/models/RedHatAI/Llama-4-Maverick-17B-128E-Instruct-NVFP4

Vision
1M · in - · out -

Llama 4 Scout 17B 16E Instruct Fp8 Lora

May 2026

Meta chat model.

Vision
10M · in - · out -

Llama 3.3 70B Instruct FP8 Lora

May 2026

Meta chat model.

131K · in - · out -

Llama 4 Scout (17Bx16E)

Jun 2025

Meta chat model. https://huggingface.co/api/models/meta-llama/Llama-4-Scout-17B-16E

Vision
262K · in - · out -

Llama Guard 4 12B

Apr 2025

Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be use…

VisionWeb search
164K · in $0.18 · out $0.18

Llama 3.1 405B

Apr 2025

Meta chat model. https://huggingface.co/api/models/meta-llama/Llama-3.1-405B

131K · in - · out -

Llama 3.2 1B

Apr 2025

Meta chat model. https://huggingface.co/api/models/meta-llama/Llama-3.2-1B

131K · in - · out -

Llama 3.1 70B

Apr 2025

Meta chat model. https://huggingface.co/api/models/meta-llama/Llama-3.1-70B

131K · in - · out -

Llama 4 Scout

Apr 2025

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It su…

VisionTools / functionsWeb search
1.3M · in $0.10 · out $0.30

Llama 4 Maverick

Apr 2025

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts…

VisionTools / functionsWeb search
1M · in $0.20 · out $0.70

Llama 4 Scout Instruct (17Bx16E)

deprecated
Apr 2025

Meta chat model. https://huggingface.co/meta-llama/Llama-4-Scout-17B-16E-Instruct

1M · in $0.18 · out $0.59

meta-llama/Llama-2-7b-chat-hf

Apr 2025

Meta chat model. https://huggingface.co/meta-llama/Llama-2-7b-chat-hf

4K · in - · out -

Meta Llama 3.1 8B Instruct Turbo

deprecated
Mar 2025

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3.1-8B-Instruct

131K · in $0.18 · out $0.18

Llama 3.3 70B Instruct

Dec 2024

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 inst…

Tools / functionsWeb search
131K · in $0.10 · out $0.32

Meta Llama 3.1 405B Instruct

deprecated
Dec 2024

Meta chat model. https://huggingface.co/meta-llama/Llama-3.1-405B-Instruct

4K · in $3.50 · out $3.50

Meta Llama 3.3 70B Instruct Turbo

Dec 2024

Meta chat model. https://huggingface.co/meta-llama/Llama-3.3-70B-Instruct

131K · in $1.04 · out $1.04

Llama 3.2 90B Vision

Sep 2024

Llama large vision model. NVIDIA serves a reduced 32K context (native: 128K).

Vision
33K · in Free · out Free

Llama 3.2 11B Vision

Sep 2024

Llama vision model for image understanding.

Vision
131K · in Free · out Free

Llama 3.2 1B

Sep 2024

Tiny Llama for edge-class tasks.

Tools / functions
131K · in Free · out Free

Llama 3.2 3B

Sep 2024

Small Llama for lightweight tasks.

Tools / functions
131K · in Free · out Free

Llama 3.1 8B

Jul 2024

Small fast Llama for utility tasks, with tool calling. Retires on NVIDIA 2026-08-25.

Tools / functions
131K · in Free · out Free

Llama 3.1 70B

Jul 2024

Meta Llama 3.1 70B instruction-tuned (superseded by Llama 3.3 70B).

Tools / functions
131K · in Free · out Free

Meta Llama 3.1 70B Instruct Turbo

deprecated
Jul 2024

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3.1-70B-Instruct

131K · in $0.88 · out $0.88

Meta Llama 3 8B Instruct Reference

deprecated
Apr 2024

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct

8K · in $0.20 · out $0.20

Llama 3 8B Instruct

deprecated
Apr 2024

Meta's latest class of model (Llama 3) launched with a variety of sizes & flavors. This 8B instruct-tuned version was optimized for high quality dialogue useca…

Web search
8K · in $0.14 · out $0.14

Meta Llama 3 8B Instruct

deprecated
Apr 2024

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct

8K · in $0.20 · out $0.20

Meta Llama 3 70B Instruct Turbo

deprecated
-

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3-70B-Instruct

8K · in $0.88 · out $0.88

Meta Llama 3 8B Instruct Lite

deprecated
-

Meta chat model. https://huggingface.co/meta-llama/Meta-Llama-3-8B-Instruct

8K · in $0.14 · out $0.14
34 models · sorted by release date · prices in USD per 1M tokens · refreshed hourlyCompare every model across vendors ->

Get started in 3 steps

1

Create an API key at the Meta console.

2

Paste it into Big-AGI's model settings.

3

Start chatting, or Beam it against other models and fuse the answers.

Meta Model API in Big-AGI

Add your Meta Model API key and use the Muse models at Meta's own rates: Big-AGI adds no markup, and Meta bills you directly. Every Muse Spark version is served on two tiers. Standard does not use your prompts for training; Contributor is cheaper in exchange for Meta training on your prompts and completions. Big-AGI hides the Contributor models until you unhide them.

  • Reasoning, always on. Muse Spark reasons before every answer. The effort control runs from minimal to xhigh, set per request; the default is Meta's, high.
  • Web search and images. Web search is a per-model switch, billed by Meta per search. Muse Image generates and edits images from text and reference images, at a flat per-image price.
  • Direct Connection. Meta's API accepts browser-direct requests, so the toggle is available under Advanced once your key is in the browser.

Your key and your data

Your key stays in your browser and is sent only with your requests. Chats are stored on your device first and sync only if you turn sync on. The AI Inspector shows every request, the token counts and the cost estimate, including the reasoning tokens Muse Spark spends before answering.

Meta in Beam

Beam runs Muse Spark next to Claude, GPT, Gemini and anything else you connected, in parallel, then fuses the answers. Parallel runs use more tokens than a single chat, and Muse Spark's reasoning tokens count too.

Meta questions

FAQ

What is the difference between Muse Spark Standard and Contributor?

They are the same model. Standard does not use your prompts and completions for training and costs $1.25 per million input tokens and $4.25 per million output tokens as of September 2, 2026. Contributor costs $0.10 and $0.20 in exchange for Meta training future models on your prompts and completions, with lower rate limits. Big-AGI hides the Contributor models until you unhide them in the model list.

Why does my Meta Model API key return Unauthorized?

Meta returns the same Unauthorized error for a wrong key, a revoked key, and a missing key. Recreate the key at dev.meta.ai/api-keys and paste it into the Meta Model API Key field; keys start with LLM_ and are shown once at creation.

Why does Muse Spark pause before answering?

Muse Spark reasons before every reply, and Meta's default reasoning effort is high. In our September 2, 2026 check a two-word reply spent about 340 reasoning tokens and 9 seconds, with the first token after 2.2 seconds. Lower the Reasoning Effort control - minimal to xhigh, there is no off - to trade depth for speed.

Can Big-AGI generate images with Muse Image?

Yes. Muse Image 1.0 appears in the model list next to Muse Spark and generates or edits images from text and reference images, at a flat $0.01 per image billed by Meta. It returns one WebP image per request and runs without streaming.

Bring your Meta key. Keep control.

Your key, your data, your choice of model. Big-AGI's Open branch is open source and self-hostable, so you can check exactly how Meta is called.

Launch Big-AGI

© 2026 Token Fabrics·Built with passion in San Diego