Llama 3.1 Nemotron Nano 8B V1 llama-3.1-nemotron-nano-8b-v1

nvidia · nvidia

unranked — no quality signals yet

params
8B
arch
dense
license
other open weights

# architecture

attn shared GQA 32q/8kv feed-forward (dense) every parameter, every token out ×32 layers 8B — no routing
layers
32
d_model
4096
heads
32q / 8kv
head dim
128
vocab
128256
family
llama

signal path of one layer, generated from the registry's structured fields — dimension lines quote the model's real numbers; an MoE trace forks at the router, a dense trace runs straight through. Geometry fields come from the repo's config.json.

# why ranked

Overall: provisional — score 3.7 ● low — a single signal; treat with caution

provisional: too few independent signals for a numbered position — the score is shown, but this model sorts after every ranked model.

signalweightinput (0–100)
aa_intelligence 0.70
bench_composite 0.30 3.7

benchmark panel evidence:

complete panelscore (0–100)
aime-2025-v13.7

missing signals are dropped and the remaining weights renormalized — never imputed.

full methodology

# trend

OGM score · last 36 days
OGM score over 36 days

# benchmarks

benchmarkscoresourcedate
AIME 2025 0.5 / 1 LLM Stats
Bfcl-v2 0.6 / 1 LLM Stats
Math-500 1.0 / 1 LLM Stats

# where to run

no known API providers — download the weights from Hugging Face and run them yourself.

source aliases
llmstats
llama-3.1-nemotron-nano-8b-v1