CATALOGUE RECORD / HUB-DERIVED
nvidia
Qwen3.5-122B-A10B-NVFP4
Revision 98915d837c4e7c87ac8296d02e89de19b3207e6d
INDEX INCONSISTENCY
This repository publishes two different parameter counts.
The safetensors index declares 64,580,759,280 parameters in its total, while its own per-dtype map sums to 74,352,195,824. Those two fields describe the same tensors, so one of them is wrong.
No parameter count is published for this record. Picking the more plausible of two contradictory figures would be a guess presented as a fact. The tensor byte total below is computed from the per-dtype map alone and is cross-checked against the repository’s stored bytes.
TENSOR ACCOUNTING
Where the bytes are.
Summed from the safetensors index, one row per dtype. A parameter count alone cannot produce this figure, because a checkpoint may mix widths.
| Dtype | Parameters | Bytes each | Bytes | Share |
|---|---|---|---|---|
U8 | 57,982,058,496 | 1 | 58.0 GB | 69.5% |
BF16 | 9,122,380,016 | 2 | 18.2 GB | 21.9% |
F8_E4M3 | 7,247,757,312 | 1 | 7.2 GB | 8.7% |
EVERY FIELD, WITH ITS ORIGIN
Sourced or undetermined. Never assumed.
Each value below names the exact API field it was computed from. Where the Hub does not establish a value, the reason is shown instead of a plausible default.
- Architecture
- qwen3_5_moe
config.model_type - Model classes
- Qwen3_5MoeForConditionalGeneration
config.architectures - Routed experts
- UndeterminedThe published config exposes no routed-expert count. The Hub's config summary omits fields some architectures place only in the full config.json.
- Experts per token
- UndeterminedThe published config exposes no per-token expert count.
- Shared experts
- UndeterminedThe published config declares no always-on shared experts.
- Routing sparsity
- UndeterminedRouting sparsity requires both a routed-expert and a per-token expert count.
- Total parameters
- UndeterminedThe published index is internally inconsistent: safetensors.total declares 64,580,759,280 parameters while the per-dtype map sums to 74,352,195,824. Neither figure can be treated as the parameter count.
- Checkpoint bytes
- 83,474,575,840 (83.5 GB)
safetensors.parameters - Ships below 16-bit
- Yes
safetensors.parameters - Quantisation method
- modelopt
config.quantization_config.quant_method - Trained context
- UndeterminedThe config summary omits max_position_embeddings. Trained context is a model-card claim, not a derivable fact.
- Declared licence
- apache-2.0
cardData.license - Base model
- Qwen/Qwen3.5-122B-A10B
cardData.base_model - Library
- Model Optimizer
library_name - Files in repository
- 21
siblings - Last modified
- 2026-06-02
lastModified
LICENCE POSTURE
Permissive
The repository declares a licence that is generally read as permitting commercial use. Read the licence file in the repository before relying on that.
This is a reading of a metadata field, not legal advice, and it does not account for the licences of upstream models or training data.
WEIGHT RESIDENCY FLOOR
The count below which it cannot fit.
Ceiling of checkpoint bytes over advertised accelerator memory. This is a lower bound on accelerator count for weights alone — KV cache, activations and runtime overhead all sit on top, so a real deployment needs more.