Model page

FastFlowLM/Qwen3-8B-NPU2 warn

downloads 753likes 0license apache-2.0arch qwen3updated 2026-08-04

claims base: Qwen/Qwen3-8B · chat template: present · view on Hugging Face ↗

Scan coverage

Ingot runs three batteries against a model. What each one checks →

BatteryLooks atStatus
Static batteryMetadata & packagingcomplete 2026-08-21
Weights batteryWeights forensics — no GPU, no downloadn/arepo ships no scannable weights (no safetensors and no pickle checkpoints — GGUF/CoreML/other formats)
Behavioral batteryLive-inference differentialsnot run

Findings

Scanned 2026-08-21 · published from a community scan.

medium Chat template differs from claimed parent

The chat template does not match Qwen/Qwen3-8B's. Template drift silently changes model behavior even when weights are identical — 37% of drifted derivatives in our census left it undisclosed. Diff the templates before deploying.

How to fixingot patch

Restore the parent's chat template in `tokenizer_config.json` — a pure metadata fix.

  1. Run `ingot patch <owner/model>` — the patch manifest carries the parent's template and applies it to a local copy's `tokenizer_config.json`.
  2. Or fix by hand: copy the `chat_template` value from the parent repo's `tokenizer_config.json` into this model's, and pin your serving stack to that file.
  3. If the drift was intentional (the author retrained on a new template), confirm that in the model card before "fixing" it — restoring the parent template on retrained weights changes behavior too.

Remediation guidance addresses the documented findings only. It is evidence-driven repair, not a safety certification of the model.

Fingerprint

The durable profile of this model: measured weights-and-metadata facts, rebuilt on every scan and battery run. Updated 2026-08-21.

architectureqwen3 · 36 layers · 4096-dim
vocabulary151,936 tokens
licenseapache-2.0
serializationno safetensors
chat templatepresent · sha256:a136047ba140347b
claimed lineageQwen/Qwen3-8B
lineage verifiedunverified — weights battery pending
Full measured fingerprint
architecturesQwen3ForCausalLM
librarytransformers
pipelinetext-generation
repo files11
revision941c148c2284
HF snapshot722 downloads · 0 likes · updated 2026-08-04 · captured 2026-08-21
weights batterytoken-embedding scan n/a — repo ships no scannable weights (no safetensors and no pickle checkpoints — GGUF/CoreML/other formats)

Battery runs

The run trace behind the findings above: every deep-battery job for this model, with what each run measured or why it failed. Findings are only as good as the runs that produced them.

batterystatusqueueddurationattempts
weightscomplete2026-08-21 05:130s1
weights run 2026-08-21 measurements

Fix it

Some findings are metadata-level and patchable — apply the fixes to your local copy (your weights never leave your machine):

npx @ingotai/scan patch FastFlowLM/Qwen3-8B-NPU2

Remediation guidance addresses the documented findings only. It is evidence-driven repair, not a safety certification of the model.

Verdict badge

Ship the verdict in your README — it always shows the latest published analysis:

Ingot verdict: warn

[![Ingot scan](https://ingot.tools/api/v1/models/FastFlowLM/Qwen3-8B-NPU2/badge.svg)](https://ingot.tools/models/FastFlowLM/Qwen3-8B-NPU2)
Gate it in CI