mainsafetensorsbf16Public
cuddle-ord/teutonic-5ey6gmawko-top380
sha256:d74dcfe6970e01871a1a69a90cb15625dfd78c7cfac01df97947b26b96830fd5·Indexed 10d ago
Parameters
8.6B
Total size
16.0 GB
Files
11
Quantization
BF16
No README on this version
Push a README.md with the model files to see it rendered here.
Model architecture
config.json- Architecture
- QuasarForCausalLM
- Model type
- quasar_text
- Hidden size
- 4,096
- Layers
- 32
- Attention heads
- 16 (4 kv)
- FFN size
- 12,288
- Vocab size
- 248,320
- Context window
- 2048K
Files
11 itemsmodel-00003-of-00004.safetensors
02e776ffffa7
4.7 GB
safetensors
model-00002-of-00004.safetensors
708438bbe41f
4.7 GB
safetensors
model-00001-of-00004.safetensors
73677971ecba
4.6 GB
safetensors
model-00004-of-00004.safetensors
cfa7191843b7
2.1 GB
safetensors
tokenizer.json
87a7830d63fc
19.1 MB
modeling_qwen3_5.py
a9e25ab03df1
104.7 KB
model.safetensors.index.json
62d1e320ad1a
37.2 KB
configuration_qwen3_5.py
6e5d4d9ef7b2
14.4 KB
config.json
4429106ed8a2
2.1 KB
tokenizer_config.json
eb5735429072
1.1 KB
generation_config.json
75521f2ad5f9
214 B