the model
llama-3.1-8b-instant
Served by Groq. Point a task at it and the audit prices the ceiling before a token is spent.
Vendored from the released binary · supernovae-st/nika@5d61b7e13f1f (v0.107.0) · digest-verified, re-derived at build, gated in CI
Point a task at it
tasks:
ask:
infer:
model: "groq/llama-8b"
prompt: "…"The seats · 1
- groq/llama-8b128k context · 8k out
Groq · the provider room
The price
- groq$0.05 in · $0.08 out · open weights
per million tokens · 131k window