406B parameters. Writes text from a prompt. This is the kind of model people mean by "an LLM".
Packaged as FP8. FP8 needs a GPU that supports it — Hopper or newer.
llama3.1 — the lab's own terms rather than a standard licence. Worth reading before you ship.