RivenGet started

Llama-3.1-8B-Instant vs Llama-3.3-70B-Versatile: API Pricing Compared

Side-by-side API pricing and capabilities for Llama-3.1-8B-Instant and Llama-3.3-70B-Versatile. Both are available through Riven behind one OpenAI-compatible endpoint — at vendor list price with transparent per-token pricing, or a flat $3 per million tokens on pay-as-you-go.

Llama-3.1-8B-InstantLlama-3.3-70B-Versatile
Input price ($/1M tokens)$0.05$0.59
Output price ($/1M tokens)$0.08$0.79
Vendorproviderprovider
Capabilitieschat, fastchat

Model pages: Llama-3.1-8B-Instant · Llama-3.3-70B-Versatile

FAQ

Which is cheaper, Llama-3.1-8B-Instant or Llama-3.3-70B-Versatile?

On Riven both are billed at vendor list price with transparent per-token pricing: Llama-3.1-8B-Instant is $0.05/1M input and $0.08/1M output; Llama-3.3-70B-Versatile is $0.59/1M input and $0.79/1M output. On the pay-as-you-go plan, every model is a flat $3 per million tokens.

Can I use Llama-3.1-8B-Instant and Llama-3.3-70B-Versatile with one API key?

Yes. One Riven key reaches both models through the same OpenAI-compatible endpoint, so you can A/B them by changing only the model id.

Is there a markup on these prices?

No. Riven lists these models at vendor list price with transparent per-token pricing. Pay-as-you-go users instead pay a flat $3 per million tokens across every model.