AGPL-3.0-or-later · forever.

the model

meta/llama-3.1-70b-instruct

Served by NVIDIA. Point a task at it and the audit prices the ceiling before a token is spent.

Vendored from the released binary · supernovae-st/nika@5d61b7e13f1f (v0.107.0) · digest-verified, re-derived at build, gated in CI

Point a task at it

tasks:
  ask:
    infer:
      model: "nvidia/llama-70b"
      prompt: "…"

The seats · 1

  1. nvidia/llama-70b128k context · 4k out

    NVIDIA · the provider room

The price

  1. nvidia$0 in · $0 out · open weights

    per million tokens · 128k window