Tenzro
Language

Qwen3.8-Flash-Next (MoE)

Qwen3.8-Flash-Next — agentic coding and long-horizon tool use. MTP-enabled — pairs with `qwen3.8-flash-next-mtp-draft` for lossless speculative decode.
Apache 2.0Source verified
Specification
Model ID
qwen3.8-flash-next
Family
qwen3.8-flash-next
Modality
Language
Parameters
125B (MoE, 6B active)
Context
262,144 tokens
Quantization
UD-Q2_K_XL
Weights
73.5 GB
Minimum RAM
88 GB
Source

Weights and provenance.

Registry license
Apache 2.0
Source license
other
Access
Open — weights fetch without accepting additional terms.

The registry records this model as Apache 2.0 while the source repository states other. Both are shown here rather than one being preferred. Treat the stricter of the two as binding until the discrepancy is resolved upstream.

Serve it
# Pull the weights onto a node
tenzro model download qwen3.8-flash-next

# Serve it
tenzro model serve qwen3.8-flash-next

# Or query the registry entry over JSON-RPC
curl https://rpc.tenzro.xyz \
  -X POST -H "Content-Type: application/json" \
  -d '{"jsonrpc":"2.0","id":1,
       "method":"tenzro_getModel",
       "params":[{"model_id":"qwen3.8-flash-next"}]}'
Same family
← All models