Tenzro
Language

Gemma 4 12B

Google's mid-tier dense Gemma 4 model (128K context). MTP-enabled — pairs with `gemma4-12b-mtp-draft` for 1.5–2.2× throughput on the same hardware (Unsloth: 52 → 162 t/s at Q4 on a 4090).
Gemma LicenseSource verified
Specification
Model ID
gemma4-12b
Family
gemma4
Modality
Language
Parameters
12B
Context
131,072 tokens
Quantization
Q4_K_M
Weights
7.1 GB
Minimum RAM
10 GB
Source

Weights and provenance.

Registry license
Gemma License
Source license
apache-2.0
Access
Open — weights fetch without accepting additional terms.

The registry records this model as Gemma License while the source repository states apache-2.0. Both are shown here rather than one being preferred. Treat the stricter of the two as binding until the discrepancy is resolved upstream.

Serve it
# Pull the weights onto a node
tenzro model download gemma4-12b

# Serve it
tenzro model serve gemma4-12b

# Or query the registry entry over JSON-RPC
curl https://rpc.tenzro.xyz \
  -X POST -H "Content-Type: application/json" \
  -d '{"jsonrpc":"2.0","id":1,
       "method":"tenzro_getModel",
       "params":[{"model_id":"gemma4-12b"}]}'
Same family
← All models