the model
meta/llama-3.1-70b-instruct
Served by NVIDIA. Point a task at it and the audit prices the ceiling before a token is spent.
Vendored from the released binary · supernovae-st/nika@5d61b7e13f1f (v0.107.0) · digest-verified, re-derived at build, gated in CI
Point a task at it
tasks:
ask:
infer:
model: "nvidia/llama-70b"
prompt: "…"The seats · 1
- nvidia/llama-70b128k context · 4k out
NVIDIA · the provider room
The price
- nvidia$0 in · $0 out · open weights
per million tokens · 128k window