Qwen2.5-Coder-1.5B-Instruct (q4f16_1, WebLLM)

models.skillsafe.ai/qwen2.5-coder-1.5b-instruct-q4f16@30184a4f713a/

Vetted New WebGPU

Text generation model for web-llm (q4f16_1). Runs on WebGPU — 833 MB downloaded once from models.skillsafe.ai, then cached for every SkillSafe app that uses it. Inference happens on your device; nothing you enter is uploaded to load it, and it never costs a credit. Weights from Hugging Face · skillsafe-ai/qwen2.5-coder-1.5b-instruct-q4f16, pinned at e99732de48ec — also loadable straight from Hugging Face outside SkillSafe (how).

Share

Details

Catalogue idqwen2.5-coder-1.5b-instruct-q4f16@30184a4f713a
Runtimeweb-llm ≥ 0.2.80
DeviceWebGPU
Variantq4f16_1
Download833 MB · 31 files
LicenceApache-2.0 · notice
Pinned ate99732de48ec21158e5119a7758cc45e464cfc8d
ApprovedSep 22, 2026
Statusactive — every file is a live vetted hash

Files

Each file is served at an immutable URL; the canonical form is the SHA-256 itself. Tokenizer and config JSON are not here by design — they ship in your app bundle.

PathFormatSizeSHA-256
params_shard_0.bin bin 111 MB 3c312ff8880a…178e66
params_shard_1.bin bin 21.3 MB e9746bfae164…4b8cf0
params_shard_10.bin bin 25.1 MB c2de8982284a…d4ecda
params_shard_11.bin bin 25.1 MB fe3fb13462cb…012227
params_shard_12.bin bin 25.1 MB 6e2428c621f8…671dcb
params_shard_13.bin bin 25.1 MB c351eb26223c…7ca476
params_shard_14.bin bin 25.1 MB 32466b2c68c1…5cb134
params_shard_15.bin bin 25.1 MB 11fbd058b080…93b5e2
params_shard_16.bin bin 25.1 MB 88ca4f3c3b5f…9c134b
params_shard_17.bin bin 25.1 MB 27cb36c5b409…a9f42a
params_shard_18.bin bin 25.1 MB 4dd21bf71bb1…8b9f7d
params_shard_19.bin bin 25.1 MB 63595695c796…d9c4aa
params_shard_2.bin bin 25.1 MB 7123276d20fc…f52b3a
params_shard_20.bin bin 25.1 MB e74ac3965a66…35d05b
params_shard_21.bin bin 25.1 MB 8cdac59722ba…3b19b5
params_shard_22.bin bin 25.1 MB 7355b4a63af1…980320
params_shard_23.bin bin 25.1 MB 6880182ff408…618a68
params_shard_24.bin bin 25.1 MB 5147a10cfb11…d569b3
params_shard_25.bin bin 25.1 MB a9f8b1eb8f83…b7b131
params_shard_26.bin bin 25.1 MB bbafb40e746d…1659d9
params_shard_27.bin bin 25.1 MB 274552455f7e…3a5e1c
params_shard_28.bin bin 25.1 MB 84fb8349cac7…057777
params_shard_29.bin bin 17.7 MB cc09bf175d86…1bd8fc
params_shard_3.bin bin 25.1 MB 113c72600912…384672
params_shard_4.bin bin 25.1 MB 384e14ea7e1f…280044
params_shard_5.bin bin 25.1 MB 8830daa29ffc…a866b5
params_shard_6.bin bin 25.1 MB 3693fc3c4103…c22ea0
params_shard_7.bin bin 25.1 MB ae42f18931f0…ff0541
params_shard_8.bin bin 25.1 MB 408bb892e85d…392a84
params_shard_9.bin bin 25.1 MB 5ce5e702e4b9…daa741
lib/Qwen2-1.5B-Instruct-q4f16_1-ctx4k_cs1k-webgpu.wasm wasm 5.1 MB 6cd9f132ad25…cec538
Use it in an app declaration · SDK loader · web-llm · URLs · Hugging Face — generated from this entry

Add to the body of POST /v1/apps/{slug}/releases (or a release session). An unknown or withdrawn model is refused with a 400 naming it; the app page then shows "downloads 833 MB · runs on your device" and /models.txt carries the attribution.

{
  "models": [
    {
      "id": "qwen2.5-coder-1.5b-instruct-q4f16",
      "revision": "30184a4f713a"
    }
  ]
}

Evaluation

shard sizes + md5 vs ndarray-cache.json; implied parameter count vs the checkpoint; parameter table vs model_lib embedded metadata — MLC's published weights, checked against the WebGPU library. Evaluated Sep 22, 2026.

Library fit

params_matched: 311 · symbolic_dims_resolved: [] · model_type: qwen2 · quantization: q4f16_1 · effective: {"context_window_size":4096,"prefill_chunk_size":1024}

Toolchain: python 3.12.13 · platform Darwin 25.6.0 arm64 · torch 2.10.0 · onnx 1.23.0 · onnxruntime 1.30.0. Recipe models/recipes/qwen2.5-coder-1.5b-instruct-q4f16.yaml (856bf330211f). Full manifest.json

Licence & attribution

Apache-2.0 · licence text · notice

Qwen2.5-Coder-1.5B-Instruct: Copyright 2024 Alibaba Cloud. Apache License 2.0. https://huggingface.co/Qwen/Qwen2.5-Coder-1.5B-Instruct — q4f16_1 weights published by MLC (https://huggingface.co/mlc-ai/Qwen2.5-Coder-1.5B-Instruct-q4f16_1-MLC).

Apps that declare this model get this text in their generated /models.txt, so a licence that requires a notice always carries one.

Revisions

RevisionVariantSizeApprovedStatusApps
qwen2.5-coder-1.5b-instruct-q4f16@30184a4f713a q4f16_1 833 MB Sep 22, 2026 active 0
qwen2.5-coder-1.5b-instruct-q4f16@30184a4f q4f16_1 828 MB Sep 22, 2026 withdrawn 0

Every file here was approved by exact SHA-256 after a structural audit of the graph, fetched from a content-pinned source, and is served credential-free at an immutable URL. The SDK re-verifies the hash on your device before it caches or returns anything. Missing a variant? Request it — or read how the registry works.