WUMBOLABS / EVALUATIONS CANONICAL MODEL EVALUATION

/evaluations/ling-3.0-tiny

Ling 3.0 Tiny

Ling 3.0 Tiny — the current WumboLabs evidence state on one page: tested profiles, validated context, and the full chronological testing history. Each value is attributed to the profile and event that measured it.

NOT READY / latest evidence 2026-09-25

WumboLabs tests Ling 3.0 Tiny on real consumer hardware. This is the canonical model page: current state first, then every tested profile and every evidence event. Values are attributed to the profile and event that measured them; historical findings remain the evidence of their tested stack and are never silently replaced.

Current state

Tested profiles

llama.cpp official Q8_0 (Reasoning On, deployment sampler, 32K) — CURRENT

Profile identity: ling-3.0-tiny-q8-0-llamacpp-b10999-rtx5070-deployment.

FieldValue
Runtimenot recorded
Artifactnot recorded
Precisionnot recorded

Status: current canonical/recommended tested surface.

Canonical Evidence Profile Metadata

Events on this profile:

llama.cpp official Q8_0 (Reasoning Off, deployment sampler, 32K) — CURRENT-ALTERNATE

Profile identity: ling-3.0-tiny-q8-0-llamacpp-b10999-rtx5070-deployment-reasoning-off.

FieldValue
Runtimenot recorded
Artifactnot recorded
Precisionnot recorded

Status: validated alternate tested surface.

Canonical Evidence Profile Metadata

Events on this profile:

Testing history

Newest first. Each event is one immutable testing/publication event; the exact scientific report lives in the canonical evidence repository linked at the top of each event.

2026-09-25 — Reasoning On classification re-derived under the 2026-09-25 reasoning-profiles snapshot from retained hash-bound evidence (no new inference)

WELP Reasoning-Profile Re-derivation — Ling 3.0 Tiny (Reasoning On, publisher default) · profile: llama.cpp official Q8_0 (Reasoning On, deployment sampler, 32K) · maturity: CURRENT_WELP · status: NOT_READY

Canonical Evidence / Full Report View on GitHub

Identity
FieldValue
ModelLing 3.0 Tiny
Producernot recorded
Tested artifactnot recorded
Precisionnot recorded
Artifact SHA-2569299a9e5cbc540597619e252a41fd671faa4e84e619e3cea816542c84e19f0d6
Campaignling-3.0-tiny-rtx5070-welp-reasoning-on-2026-09-25
Record date2026-09-25
Runtime and hardware
FieldValue
Enginenot recorded
Runtime versionnot recorded
HardwareNVIDIA GeForce RTX 5070 12 GB
WELP outcome
  • Outcome: COMPLETE_PASS
  • Classification: not recorded

Publication state: published — canonical evidence: https://github.com/WumboLabs/evaluations/blob/a5717d2fecc21184da6dfa20d973fb8e4fbca0e4/models/ling-3.0-tiny/events/ling-3-0-tiny-rtx5070-welp-reasoning-on-20260925/REPORT.md

Context profile
FieldValue
Practical defaultnot recorded tokens
Guarded contextnot recorded tokens
Native model-card maximumnot recorded tokens
Model-card envelope completenot recorded
Native maximum dispositionnot recorded
Quality and capabilities
  • Constrained result: not recorded
Guardrails and limitations

Reliability: not recorded

LocalMaxxing
FieldValue
StatusSUBMITTED
Canonical contextnot recorded tokens
tok/s outnot recorded
TTFTnot recorded
Submission referencenot recorded
verifiedRunnull (not claimed)

the engine-native llama-bench record does not exercise the reasoning control, so both reasoning profiles share this record

Canonical evidence

Canonical public evidence: https://github.com/WumboLabs/evaluations/blob/a5717d2fecc21184da6dfa20d973fb8e4fbca0e4/models/ling-3.0-tiny/events/ling-3-0-tiny-rtx5070-welp-reasoning-on-20260925/REPORT.md

This event section is a human-readable derivative of the accepted local WELP campaign evidence named above; the campaign's REPORT.md is the authoritative scientific source. Results are bounded by the tested artifact, runtime, hardware, configuration, and protocol snapshot, and are not universal model rankings.

2026-09-25 — Full WELP characterization of the Reasoning Off deployment profile: independent setup/calibration, same frozen task set, dual-profile completion of the case-C model

WELP Reasoning-Profile Characterization — Ling 3.0 Tiny (Reasoning Off, supported alternate) · profile: llama.cpp official Q8_0 (Reasoning Off, deployment sampler, 32K) · maturity: CURRENT_WELP · status: NOT_READY

Canonical Evidence / Full Report View on GitHub

Identity
FieldValue
ModelLing 3.0 Tiny
Producernot recorded
Tested artifactnot recorded
Precisionnot recorded
Artifact SHA-2569299a9e5cbc540597619e252a41fd671faa4e84e619e3cea816542c84e19f0d6
Campaignling-3.0-tiny-rtx5070-welp-reasoning-off-2026-09-25
Record date2026-09-25
Runtime and hardware
FieldValue
Enginenot recorded
Runtime versionnot recorded
HardwareNVIDIA GeForce RTX 5070 12 GB
WELP outcome
  • Outcome: COMPLETE_PASS
  • Classification: not recorded

Publication state: published — canonical evidence: https://github.com/WumboLabs/evaluations/blob/a5717d2fecc21184da6dfa20d973fb8e4fbca0e4/models/ling-3.0-tiny/events/ling-3-0-tiny-rtx5070-welp-reasoning-off-20260925/REPORT.md

Context profile
FieldValue
Practical defaultnot recorded tokens
Guarded contextnot recorded tokens
Native model-card maximumnot recorded tokens
Model-card envelope completenot recorded
Native maximum dispositionnot recorded
Quality and capabilities
  • Constrained result: not recorded
Guardrails and limitations

Reliability: not recorded

LocalMaxxing
FieldValue
StatusSUBMITTED
Canonical contextnot recorded tokens
tok/s outnot recorded
TTFTnot recorded
Submission referencenot recorded
verifiedRunnull (not claimed)

this profile's own bench re-measurement; consistent with the existing service record (benchmark does not exercise the reasoning control; no duplicate submitted)

Canonical evidence

Canonical public evidence: https://github.com/WumboLabs/evaluations/blob/a5717d2fecc21184da6dfa20d973fb8e4fbca0e4/models/ling-3.0-tiny/events/ling-3-0-tiny-rtx5070-welp-reasoning-off-20260925/REPORT.md

This event section is a human-readable derivative of the accepted local WELP campaign evidence named above; the campaign's REPORT.md is the authoritative scientific source. Results are bounded by the tested artifact, runtime, hardware, configuration, and protocol snapshot, and are not universal model rankings.

2026-09-25 — First WELP Agentic section execution: three bounded tool tasks through the pinned harness in a disposable offline sandbox; methodology-validation event on a known model

WELP Agentic — Ling 3.0 Tiny (Reasoning Off; methodology validation) · profile: llama.cpp official Q8_0 (Reasoning Off, deployment sampler, 32K) · maturity: CURRENT_WELP · status: NOT_READY

Canonical Evidence / Full Report View on GitHub

Identity
FieldValue
ModelLing 3.0 Tiny
Producernot recorded
Tested artifactnot recorded
Precisionnot recorded
Artifact SHA-2569299a9e5cbc540597619e252a41fd671faa4e84e619e3cea816542c84e19f0d6
Campaignling-3.0-tiny-rtx5070-welp-reasoning-off-2026-09-25
Record date2026-09-25
Runtime and hardware
FieldValue
Enginenot recorded
Runtime versionnot recorded
HardwareNVIDIA GeForce RTX 5070 12 GB
WELP outcome
  • Outcome: COMPLETE_PASS
  • Classification: not recorded

Publication state: published — canonical evidence: https://github.com/WumboLabs/evaluations/blob/599063be57edd4f3b01f27c33da02d48abc541d8/models/ling-3.0-tiny/events/ling-3-0-tiny-rtx5070-welp-agentic-20260925/REPORT.md

Context profile
FieldValue
Practical defaultnot recorded tokens
Guarded contextnot recorded tokens
Native model-card maximumnot recorded tokens
Model-card envelope completenot recorded
Native maximum dispositionnot recorded
Quality and capabilities
  • Constrained result: not recorded
Guardrails and limitations

Reliability: not recorded

LocalMaxxing
FieldValue
StatusSUBMITTED
Canonical contextnot recorded tokens
tok/s outnot recorded
TTFTnot recorded
Submission referencenot recorded
verifiedRunnull (not claimed)

this profile's own bench re-measurement; consistent with the existing service record (benchmark does not exercise the reasoning control; no duplicate submitted)

Canonical evidence

Canonical public evidence: https://github.com/WumboLabs/evaluations/blob/599063be57edd4f3b01f27c33da02d48abc541d8/models/ling-3.0-tiny/events/ling-3-0-tiny-rtx5070-welp-agentic-20260925/REPORT.md

This event section is a human-readable derivative of the accepted local WELP campaign evidence named above; the campaign's REPORT.md is the authoritative scientific source. Results are bounded by the tested artifact, runtime, hardware, configuration, and protocol snapshot, and are not universal model rankings.

2026-09-24 — Second WELP stabilization-cohort campaign: hybrid-attention MoE characterized end-to-end on the pinned runtime (official Q8_0)

WELP Initial Evaluation — inclusionAI Ling 3.0 Tiny (reasoning-on) · profile: llama.cpp official Q8_0 (reasoning-on, deployment sampler, 32K) · maturity: CURRENT_WELP · status: NOT_READY

Canonical Evidence / Full Report View on GitHub

Identity
FieldValue
ModelLing 3.0 Tiny
Producernot recorded
Tested artifactnot recorded
Precisionnot recorded
Artifact SHA-2569299a9e5cbc540597619e252a41fd671faa4e84e619e3cea816542c84e19f0d6
Campaignling-3.0-tiny-rtx5070-welp-characterization-2026-09-24
Record date2026-09-24
Runtime and hardware
FieldValue
Enginenot recorded
Runtime versionnot recorded
HardwareNVIDIA GeForce RTX 5070 12 GB
WELP outcome
  • Outcome: COMPLETE_PASS
  • Classification: not recorded

Publication state: published — canonical evidence: https://github.com/WumboLabs/evaluations/blob/43f56d8add4b3dfc51183ccd818fc6037e79bd9c/models/ling-3.0-tiny/events/ling-3-0-tiny-rtx5070-welp-characterization-20260924/REPORT.md

Context profile
FieldValue
Practical defaultnot recorded tokens
Guarded contextnot recorded tokens
Native model-card maximumnot recorded tokens
Model-card envelope completenot recorded
Native maximum dispositionnot recorded
Quality and capabilities
  • Constrained result: not recorded
Guardrails and limitations

Reliability: not recorded

LocalMaxxing
FieldValue
StatusSUBMITTED
Canonical contextnot recorded tokens
tok/s outnot recorded
TTFTnot recorded
Submission referencenot recorded
verifiedRunnull (not claimed)
Canonical evidence

Canonical public evidence: https://github.com/WumboLabs/evaluations/blob/43f56d8add4b3dfc51183ccd818fc6037e79bd9c/models/ling-3.0-tiny/events/ling-3-0-tiny-rtx5070-welp-characterization-20260924/REPORT.md

This event section is a human-readable derivative of the accepted local WELP campaign evidence named above; the campaign's REPORT.md is the authoritative scientific source. Results are bounded by the tested artifact, runtime, hardware, configuration, and protocol snapshot, and are not universal model rankings.

Canonical evidence

All canonical public evidence lives in WumboLabs/evaluations. Each event links an immutable full-commit/path citation; each profile remains a distinct scientific identity, not a separate repository.