Language
Qwen 3.5 35B-A3B (MoE) (MTP)
Qwen 3.5 35B-A3B (MoE) with built-in Multi-Token-Prediction head. Single-file MTP GGUF — no separate drafter needed. Unsloth measures ~1.5-2× speedup over the non-MTP baseline.
Apache 2.0Source verified
Specification
Model ID
qwen3.5-35b-a3b-mtpFamily
qwen3.5
Modality
Language
Parameters
35B (MoE, 3B active)
Context
131,072 tokens
Quantization
UD-Q4_K_XL
Weights
21.0 GB
Minimum RAM
28 GB
Source
Weights and provenance.
Repository
Registry license
Apache 2.0
Source license
apache-2.0
Access
Open — weights fetch without accepting additional terms.
Serve it
# Pull the weights onto a node
tenzro model download qwen3.5-35b-a3b-mtp
# Serve it
tenzro model serve qwen3.5-35b-a3b-mtp
# Or query the registry entry over JSON-RPC
curl https://rpc.tenzro.xyz \
-X POST -H "Content-Type: application/json" \
-d '{"jsonrpc":"2.0","id":1,
"method":"tenzro_getModel",
"params":[{"model_id":"qwen3.5-35b-a3b-mtp"}]}'Same family
Qwen 3.5 0.8B
0.8B
Qwen 3.5 0.8B (MTP)
0.8B
Qwen 3.5 122B-A10B (MoE)
122B (MoE, 10B active)
Qwen 3.5 122B-A10B (MoE) (MTP)
122B (MoE, 10B active)
Qwen 3.5 27B
27B
Qwen 3.5 27B (MTP)
27B
Qwen 3.5 2B
2B
Qwen 3.5 2B (MTP)
2B
Qwen 3.5 35B-A3B (MoE)
35B (MoE, 3B active)
Qwen 3.5 397B-A17B (MoE)
397B (MoE, 17B active)
Qwen 3.5 397B-A17B (MoE) (MTP)
397B (MoE, 17B active)
Qwen 3.5 4B
4B