WUMBOLABS / EVALUATIONS CANONICAL MODEL EVALUATION

/evaluations/granite-4.2-8b

Granite 4.2 8B

Granite 4.2 8B — the current WumboLabs evidence state on one page: tested profiles, validated context, and the full chronological testing history. Each value is attributed to the profile and event that measured it.

NOT READY / latest evidence 2026-09-24

WumboLabs tests Granite 4.2 8B on real consumer hardware. This is the canonical model page: current state first, then every tested profile and every evidence event. Values are attributed to the profile and event that measured them; historical findings remain the evidence of their tested stack and are never silently replaced.

Current state

Tested profiles

llama.cpp official Q4_K_M (reasoning-on, deployment sampler, 32K) — CURRENT

Profile identity: granite-4.2-8b-q4-k-m-llamacpp-b10999-rtx5070-deployment.

FieldValue
Runtimenot recorded
Artifactnot recorded
Precisionnot recorded

Status: current canonical/recommended tested surface.

Canonical Evidence Profile Metadata

Events on this profile:

Testing history

Newest first. Each event is one immutable testing/publication event; the exact scientific report lives in the canonical evidence repository linked at the top of each event.

2026-09-24 — First published evaluation via linked completion of the methodology-repair predecessor (official Q4_K_M, upstream llama.cpp b10999)

WELP Initial Evaluation — IBM Granite 4.2 8B (reasoning-on) · profile: llama.cpp official Q4_K_M (reasoning-on, deployment sampler, 32K) · maturity: CURRENT_WELP · status: NOT_READY

Canonical Evidence / Full Report View on GitHub

Identity
FieldValue
ModelGranite 4.2 8B
Producernot recorded
Tested artifactnot recorded
Precisionnot recorded
Artifact SHA-25616a9369d0805f80b7377d25d87f937a90c05dc04ad79173a52001e42c9aab311
Campaigngranite-4.2-8b-rtx5070-welp-context-outcome-completion-2026-09-24
Record date2026-09-24
Runtime and hardware
FieldValue
Enginenot recorded
Runtime versionnot recorded
HardwareNVIDIA GeForce RTX 5070 12 GB
WELP outcome
  • Outcome: COMPLETE_PASS
  • Classification: not recorded

Publication state: published — canonical evidence: https://github.com/WumboLabs/evaluations/blob/adcfa178b6dd9b52eede2c433ffc691dbcd26941/models/granite-4.2-8b/events/granite-42-8b-rtx5070-welp-context-completion-20260924/REPORT.md

Context profile
FieldValue
Practical defaultnot recorded tokens
Guarded contextnot recorded tokens
Native model-card maximumnot recorded tokens
Model-card envelope completenot recorded
Native maximum dispositionnot recorded
Quality and capabilities
  • Constrained result: not recorded
Guardrails and limitations

Reliability: not recorded

LocalMaxxing
FieldValue
StatusSUBMITTED
Canonical contextnot recorded tokens
tok/s outnot recorded
TTFTnot recorded
Submission referencenot recorded
verifiedRunnull (not claimed)

SUBMITTED/NEW, service record cmug9qe0i0d5wlq01lep45uua (APPROVED, verifiedRun false)

Canonical evidence

Canonical public evidence: https://github.com/WumboLabs/evaluations/blob/adcfa178b6dd9b52eede2c433ffc691dbcd26941/models/granite-4.2-8b/events/granite-42-8b-rtx5070-welp-context-completion-20260924/REPORT.md

This event section is a human-readable derivative of the accepted local WELP campaign evidence named above; the campaign's REPORT.md is the authoritative scientific source. Results are bounded by the tested artifact, runtime, hardware, configuration, and protocol snapshot, and are not universal model rankings.

Canonical evidence

All canonical public evidence lives in WumboLabs/evaluations. Each event links an immutable full-commit/path citation; each profile remains a distinct scientific identity, not a separate repository.