RivenGet started

DeepSeek-V4-Flash vs Llama-4-Maverick-17B-128E-Instruct-FP8: API Pricing Compared

Side-by-side API pricing and capabilities for DeepSeek-V4-Flash and Llama-4-Maverick-17B-128E-Instruct-FP8. Both are available through Riven behind one OpenAI-compatible endpoint — at vendor list price with transparent per-token pricing, or a flat $3 per million tokens on pay-as-you-go.

DeepSeek-V4-FlashLlama-4-Maverick-17B-128E-Instruct-FP8
Input price ($/1M tokens)$0.07$0.25
Output price ($/1M tokens)$0.17$1.00
Vendordeepseekmeta-llama
Capabilitieschat, reasoning, coding, agenticchat

Model pages: DeepSeek-V4-Flash · Llama-4-Maverick-17B-128E-Instruct-FP8

FAQ

Which is cheaper, DeepSeek-V4-Flash or Llama-4-Maverick-17B-128E-Instruct-FP8?

On Riven both are billed at vendor list price with transparent per-token pricing: DeepSeek-V4-Flash is $0.07/1M input and $0.17/1M output; Llama-4-Maverick-17B-128E-Instruct-FP8 is $0.25/1M input and $1.00/1M output. On the pay-as-you-go plan, every model is a flat $3 per million tokens.

Can I use DeepSeek-V4-Flash and Llama-4-Maverick-17B-128E-Instruct-FP8 with one API key?

Yes. One Riven key reaches both models through the same OpenAI-compatible endpoint, so you can A/B them by changing only the model id.

Is there a markup on these prices?

No. Riven lists these models at vendor list price with transparent per-token pricing. Pay-as-you-go users instead pay a flat $3 per million tokens across every model.