MDL
Llama 3.2 Instruct 1B
Meta
Verified against Artificial Analysis · Jul 27, 2026
Capability
16
percentile index
Price /M
$0 / $0
in / out
Context
—
Avg score
9.6
11 benchmarks
AA Intelligence
1.1
index
Weights
Open
Released
Sep 25, 2024
About
Llama 3.2 Instruct 1B is a open-weights model in the Llama family from Meta. Benchmarked on 11 evals, averaging 9.6.
Benchmark scores · 11
avg 9.6 — higher is betterThe research behind this model
all papers mentioning it →Information-Aware KV Cache Compression for Long Reasoning▲ 11 on HF · Jun 25, 2026Off-the-Shelf LLMs as Process Scorers: Training-Free Alternative to PRMs for Mathematical Reasoning▲ 7 on HF · Jun 1, 2026BioRefusalAudit: Auditing Biosecurity Refusal Depth Using General and Domain-Fine-Tuned Sparse AutoencodersMay 28, 2026Analyzing Quality-Latency-Resource Trade-offs in a Technical Documentation RAG Assistant Using LoRA AdaptationMay 27, 2026NestedKV: Nested Memory Routing for Long-Context KV Cache CompressionMay 26, 2026
Mentions matched by name in title/abstract — from the arXiv + HF daily corpus.
History
full ledger →Capability over time
GPQA Diamond17.7 → 20.8
yesterday
MATH26.6 → 27.7
yesterday
IFBench22.7 → 23.5
yesterday
GPQA Diamond19.6 → 17.7
8d ago
MATH14 → 26.6
8d ago
IFBench22.8 → 22.7
8d ago
Append-only ledger — every observed change to this model's numbers.
API providers
Artificial Analysis$0/M