llama-3.1-8b-instant

A Groq model. Every figure below was read off Groq's own pricing page and stamped with the date it was read — last pass 2026-08-04. Nothing here is estimated.

Input / 1M tokens$0.05
Output / 1M tokens$0.08
Context window128k
Vendor-quoted tokens/sec840

Price is the easy half. What it costs you also depends on how fast the first token arrives and what survives concurrency — that part is measured. See every Groq model on its provider page.