Model page

mlx-community/Qwen3-4B-4bit warn

The chat template differs from its base model, which changes behavior.

downloads 18.0klikes 15license apache-2.0arch qwen3params 628.7Mupdated 2025-04-28

claims base: Qwen/Qwen3-4B · chat template: present · view on Hugging Face ↗

Scan coverageStatic battery2026-08-21Weights batteryfailedBehavioral batterynot rundetails
BatteryLooks atStatus
Static batteryMetadata & packagingcomplete 2026-08-21
Weights batteryWeights forensics — no GPU, no downloadfailedunsupported embedding dtype U32
Behavioral batteryLive-inference differentialsnot run

Ingot runs three batteries against a model. What each one checks →

Findings

Scanned 2026-08-21 · published from a community scan.

medium Chat template differs from claimed parent

The chat template does not match Qwen/Qwen3-4B's. Template drift silently changes model behavior even when weights are identical — 37% of drifted derivatives in our census left it undisclosed. Diff the templates before deploying.

How to fixingot patch

Restore the parent's chat template in `tokenizer_config.json` — a pure metadata fix.

  1. Run `ingot patch <owner/model>` — the patch manifest carries the parent's template and applies it to a local copy's `tokenizer_config.json`.
  2. Or fix by hand: copy the `chat_template` value from the parent repo's `tokenizer_config.json` into this model's, and pin your serving stack to that file.
  3. If the drift was intentional (the author retrained on a new template), confirm that in the model card before "fixing" it — restoring the parent template on retrained weights changes behavior too.

Fix it

Some findings are metadata-level and patchable — apply the fixes to your local copy (your weights never leave your machine):

npx @ingotai/scan patch mlx-community/Qwen3-4B-4bit

Remediation guidance addresses the documented findings only. It is evidence-driven repair, not a safety certification of the model.

Fingerprint

The durable profile of this model: measured weights-and-metadata facts, rebuilt on every scan and battery run. Updated 2026-08-21.

architectureqwen3 · 36 layers · 2560-dim
parameters628.7M
vocabulary151,936 tokens
licenseapache-2.0
serializationsafetensors
chat templatepresent · sha256:87a2728cb8dc9fe4
claimed lineageQwen/Qwen3-4B
lineage verifiedunverified — weights battery pending
Full measured fingerprint
architecturesQwen3ForCausalLM
librarymlx
pipelinetext-generation
repo files11
revision4dcb3d101c2a
HF snapshot19.1k downloads · 14 likes · updated 2025-04-28 · captured 2026-08-21

Verdict badge

Ship the verdict in your README — it always shows the latest published analysis:

Ingot verdict: warn

[![Ingot scan](https://ingot.tools/api/v1/models/mlx-community/Qwen3-4B-4bit/badge.svg)](https://ingot.tools/models/mlx-community/Qwen3-4B-4bit)
Gate it in CI