a51/teutonic-q3-4b-5hgmbeef-t5138
Parameters
4.0B
Total size
7.5 GB
Files
16
Quantization
BF16
No README on this version
Push a README.md with the model files to see it rendered here.
Model architecture
config.json- Architecture
- Qwen3ForCausalLM
- Model type
- qwen3
- Hidden size
- 2,560
- Layers
- 36
- Attention heads
- 32 (8 kv)
- FFN size
- 9,728
- Vocab size
- 151,936
- Context window
- 40K
Files
16 itemsmodel-00008-of-00009.safetensors
c87cdfd4f076
942.6 MB
safetensors
model-00003-of-00009.safetensors
ec3e93ce8cc1
942.6 MB
safetensors
model-00007-of-00009.safetensors
653c42dcfd45
937.6 MB
safetensors
model-00002-of-00009.safetensors
e8b9ca1b939e
937.6 MB
safetensors
model-00001-of-00009.safetensors
36da6c52c25f
934.4 MB
safetensors
model-00006-of-00009.safetensors
7ac754a745a8
915.1 MB
safetensors
model-00005-of-00009.safetensors
ec3cdf6d08c0
915.1 MB
safetensors
model-00004-of-00009.safetensors
667c795140ca
910.1 MB
safetensors
model-00009-of-00009.safetensors
49b283d2663d
237.5 MB
safetensors
tokenizer.json
be75606093db
10.9 MB
model.safetensors.index.json
549e24ee5dfe
32.1 KB
config.json
cba674823863
1.6 KB
q3_train_state.json
e6baeec4e8c5
701 B
tokenizer_config.json
2e3bd3cb5055
693 B
teutonic_model_store.json
f10a4d522d19
548 B
generation_config.json
b77461537176
213 B