<- All Models
Meta Llama 3.1 (8B params). Fast, cost-effective for high-volume tasks. 131K context and max output. ~560 t/s on Groq. Retiring 2026-08-16 -> GPT-OSS 20B.
Available on
Where [Meta] Llama 3.1 · 8B Instant runs in Big-AGI, with each service's own model ID and USD-per-1M-token rates. The summary above prefers the creator's claim; rates below are each service's own.
Connect your own key on any of the services above, at their rates, no markup - and run it side by side with every other model.
Launch Big-AGIBIG-AGI
Resources
© 2026 Token Fabrics·Built with passion in San Diego