CATALOGUE RECORD / HUB-DERIVED

cyankiwi

Qwen3-Coder-30B-A3B-Instruct-AWQ-4bit

Revision 4bd30395b72ea6045edd04806c4fea448d4467b3

—TOTAL PARAMETERS
—CHECKPOINT BYTES
128 / 8EXPERTS / PER TOKEN
974KDOWNLOADS

INDEX INCONSISTENCY

This repository publishes two different parameter counts.

The safetensors index declares 5,306,567,040 parameters in its total, while its own per-dtype map sums to 31,466,441,088. Those two fields describe the same tensors, so one of them is wrong.

No parameter count is published for this record. Picking the more plausible of two contradictory figures would be a guess presented as a fact. The tensor byte total below is computed from the per-dtype map alone and is cross-checked against the repository’s stored bytes.

TENSOR ACCOUNTING

Where the bytes are.

Summed from the safetensors index, one row per dtype. A parameter count alone cannot produce this figure, because a checkpoint may mix widths.

DtypeParametersBytes eachBytesShare
I3229,896,998,9124119.6 GB—
BF161,569,404,92823.1 GB—
I6437,2488297984 B—

EVERY FIELD, WITH ITS ORIGIN

Sourced or undetermined. Never assumed.

Each value below names the exact API field it was computed from. Where the Hub does not establish a value, the reason is shown instead of a plausible default.

Architecture
qwen3_moeconfig.model_type
Model classes
Qwen3MoeForCausalLMconfig.architectures
Routed experts
128config.num_experts
Experts per token
8config.num_experts_per_tok
Shared experts
UndeterminedThe published config declares no always-on shared experts.
Routing sparsity
16.0× (1 of every 16.0 experts)config.num_experts / config.num_experts_per_tok
Total parameters
UndeterminedThe published index is internally inconsistent: safetensors.total declares 5,306,567,040 parameters while the per-dtype map sums to 31,466,441,088. Neither figure can be treated as the parameter count.
Checkpoint bytes
UndeterminedPer-dtype tensor bytes (122727103488) exceed the repository's stored bytes (18105930006), so the dtype map cannot be read literally for this checkpoint.
Ships below 16-bit
UndeterminedPrecision cannot be characterised from an inconsistent dtype map.
Quantisation method
compressed-tensorsconfig.quantization_config.quant_method
Trained context
UndeterminedThe config summary omits max_position_embeddings. Trained context is a model-card claim, not a derivable fact.
Declared licence
apache-2.0cardData.license
Base model
Qwen/Qwen3-Coder-30B-A3B-InstructcardData.base_model
Library
transformerslibrary_name
Files in repository
18siblings
Last modified
2026-07-21lastModified

LICENCE POSTURE

Permissive

The repository declares a licence that is generally read as permitting commercial use. Read the licence file in the repository before relying on that.

This is a reading of a metadata field, not legal advice, and it does not account for the licences of upstream models or training data.

PROVENANCE

Derived from https://huggingface.co/api/models/cyankiwi/Qwen3-Coder-30B-A3B-Instruct-AWQ-4bit in the snapshot generated 2026-09-05. Evidence class hub_derived: computed mechanically from publisher metadata, never measured by this project.