<- All Models

Llama 3.3 70B Instruct

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

131K context·$0.71 in·$0.71 out·Dec 2024 release·Tools / functions, Web search

Providers

Where Llama 3.3 70B Instruct runs in Big-AGI, with each service's own model ID and USD-per-1M-token rates. The summary above prefers the creator's claim; rates below are each service's own.

ServiceModel IDContextInputOutputCache read

meta-llama/llama-3.3-70b-instruct

131K

$0.71

$0.71

$0.71

OpenRouter

delisted

meta-llama/llama-3.3-70b-instruct:free

131K

Free

Free

-

NVIDIA

delisted

meta/llama-3.3-70b-instruct

131K

Free

Free

-

meta-llama/Llama-3.3-70B-Instruct

131K

-

-

-

nim/meta/llama-3.3-70b-instruct

16K

-

-

-

Usage mix in Big-AGI: OpenRouter 100%Specs as reported by each service · usage over the last 2 days · refreshed hourly

Run Llama 3.3 70B Instruct in Big-AGI.

Connect your own key on any of the services above, at their rates, no markup - and run it side by side with every other model.

Launch Big-AGI

© 2026 Token Fabrics·Built with passion in San Diego