AGPL-3.0-or-later · forever.

the model

llama-3.1-8b-instant

Served by Groq. Point a task at it and the audit prices the ceiling before a token is spent.

Vendored from the released binary · supernovae-st/nika@5d61b7e13f1f (v0.107.0) · digest-verified, re-derived at build, gated in CI

Point a task at it

tasks:
  ask:
    infer:
      model: "groq/llama-8b"
      prompt: "…"

The seats · 1

  1. groq/llama-8b128k context · 8k out

    Groq · the provider room

The price

  1. groq$0.05 in · $0.08 out · open weights

    per million tokens · 131k window