CATALOGUE RECORD / HUB-DERIVED
Qwen
Qwen3-30B-A3B-GPTQ-Int4
Revision 9b534e4318b7ebc3c961a839f13eb18b1833f441
30.5BTOTAL PARAMETERS
—CHECKPOINT BYTES
128 / 8EXPERTS / PER TOKEN
403KDOWNLOADS
TENSOR ACCOUNTING
Where the bytes are.
Summed from the safetensors index, one row per dtype. A parameter count alone cannot produce this figure, because a checkpoint may mix widths.
| Dtype | Parameters | Bytes each | Bytes | Share |
|---|---|---|---|---|
I32 | 29,896,998,912 | 4 | 119.6 GB | — |
F16 | 635,123,712 | 2 | 1.3 GB | — |
EVERY FIELD, WITH ITS ORIGIN
Sourced or undetermined. Never assumed.
Each value below names the exact API field it was computed from. Where the Hub does not establish a value, the reason is shown instead of a plausible default.
- Architecture
- qwen3_moe
config.model_type - Model classes
- Qwen3MoeForCausalLM
config.architectures - Routed experts
- 128
config.num_experts - Experts per token
- 8
config.num_experts_per_tok - Shared experts
- UndeterminedThe published config declares no always-on shared experts.
- Routing sparsity
- 16.0× (1 of every 16.0 experts)
config.num_experts / config.num_experts_per_tok - Total parameters
- 30,532,122,624
safetensors.total - Checkpoint bytes
- UndeterminedThis checkpoint packs 4-bit weights into wider integer containers, so element count multiplied by container width is not the checkpoint size. Use repository storage, which is measured on disk.
- Ships below 16-bit
- Yes
config.quantization_config.bits - Quantisation method
- gptq
config.quantization_config.quant_method - Trained context
- UndeterminedThe config summary omits max_position_embeddings. Trained context is a model-card claim, not a derivable fact.
- Declared licence
- apache-2.0
cardData.license - Base model
- Qwen/Qwen3-30B-A3B
cardData.base_model - Library
- transformers
library_name - Files in repository
- 11
siblings - Last modified
- 2025-05-21
lastModified
LICENCE POSTURE
Permissive
The repository declares a licence that is generally read as permitting commercial use. Read the licence file in the repository before relying on that.
This is a reading of a metadata field, not legal advice, and it does not account for the licences of upstream models or training data.