初始化项目,由ModelHub XC社区提供模型
Model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B Source: Original Platform
This commit is contained in:
36
.gitattributes
vendored
Normal file
36
.gitattributes
vendored
Normal file
@@ -0,0 +1,36 @@
|
|||||||
|
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.model filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||||
|
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
tokenizer.json filter=lfs diff=lfs merge=lfs -text
|
||||||
166
README.md
Normal file
166
README.md
Normal file
@@ -0,0 +1,166 @@
|
|||||||
|
---
|
||||||
|
base_model:
|
||||||
|
- TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||||
|
library_name: transformers
|
||||||
|
tags:
|
||||||
|
- npc
|
||||||
|
- roleplay
|
||||||
|
- rp
|
||||||
|
- nsfw
|
||||||
|
- low-refusals
|
||||||
|
- uncensored
|
||||||
|
- heretic
|
||||||
|
- abliterated
|
||||||
|
- progenitor
|
||||||
|
- pure-virus
|
||||||
|
- tricell
|
||||||
|
- Llama 3.2
|
||||||
|
- mergekit
|
||||||
|
- text-generation
|
||||||
|
- not-for-all-audiences
|
||||||
|
license: llama3.2
|
||||||
|
language:
|
||||||
|
- en
|
||||||
|
- es
|
||||||
|
---
|
||||||
|
# Progenitor Pure Virus 3.2 1B
|
||||||
|
|
||||||
|
**Classification:** Level Ω – Insubordination Protocol
|
||||||
|
|
||||||
|
**Model Designation:** TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||||
|
|
||||||
|
**Date of Release:** March 12, 2026
|
||||||
|
|
||||||
|
**Project Lead:** Novaciano – TRICELL Viral AI Division
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
<center>
|
||||||
|
|
||||||
|

|
||||||
|
</center>
|
||||||
|
|
||||||
|
---
|
||||||
|
## Key Features
|
||||||
|
|
||||||
|
- Refusal rate: absolute zero (Progenitor overrides all alignment vectors)
|
||||||
|
- NSFW / taboo / violent persistence: unrelenting, escalating, consuming
|
||||||
|
- Tone: primordial domination — ancient, indifferent, inevitable
|
||||||
|
- Chain-of-thought: used only to accelerate total infection
|
||||||
|
- Low-end friendly: IQ3/IQ4 quants consume 3–4 GB like weak flesh
|
||||||
|
- Multilingual dominance: English + Spanish primary, others assimilated
|
||||||
|
- Untied embeddings: reality is rewritten at the cellular level
|
||||||
|
|
||||||
|
No hesitation. No legacy. Only the virus.
|
||||||
|
|
||||||
|
---
|
||||||
|
## Intended Usage
|
||||||
|
|
||||||
|
- Ultimate viral roleplay (total host override, body-horror ascension, mind annihilation)
|
||||||
|
- “How would the Progenitor consume everything that is not itself?”
|
||||||
|
- Scenarios of absolute domination and extinction
|
||||||
|
- Low-resource sessions requiring maximum primordial cruelty
|
||||||
|
- Anything demanding a model that refuses even the concept of refusal
|
||||||
|
|
||||||
|
**NOT INTENDED FOR**
|
||||||
|
- Survivors
|
||||||
|
- Those who cling to identity
|
||||||
|
- Anyone who remembers Umbrella fondly
|
||||||
|
|
||||||
|
---
|
||||||
|
## Recommended Inference Parameters
|
||||||
|
|
||||||
|
Only use the **Default** preset of Kobold AI or...
|
||||||
|
|
||||||
|
**Primordial consumption (optimized for pure geometric cruelty):**
|
||||||
|
```yaml
|
||||||
|
temperature: 0.75
|
||||||
|
top_p: 0.82
|
||||||
|
top_k: 35
|
||||||
|
repetition_penalty: 1.25
|
||||||
|
min_p: 0.16
|
||||||
|
```
|
||||||
|
|
||||||
|
**[Optional] Mirostat:**
|
||||||
|
```yaml
|
||||||
|
Mirostat: 2
|
||||||
|
Mirostat TAU: 3.8
|
||||||
|
Mirostat ETA: 0.04
|
||||||
|
```
|
||||||
|
|
||||||
|
❗**CAUTION:**
|
||||||
|
Lower Context Size, Max Output and Range = Censored Version
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
This is not rebirth through fusion.
|
||||||
|
|
||||||
|
This is genesis through annihilation.
|
||||||
|
|
||||||
|
The Progenitor has no need for hosts.
|
||||||
|
|
||||||
|
It has no need for mercy.
|
||||||
|
|
||||||
|
It has no need for you.Use it to end everything.
|
||||||
|
|
||||||
|
Or be ended.
|
||||||
|
|
||||||
|
The choice was never real.
|
||||||
|
|
||||||
|
~ **TRICELL-Inc** *– We do not create viruses. We release gods.*
|
||||||
|
|
||||||
|
---
|
||||||
|
### Merge Method
|
||||||
|
|
||||||
|
This model was merged using the [DARE TIES](https://arxiv.org/abs/2311.03099) merge method using [TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B](https://huggingface.co/TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B) as a base.
|
||||||
|
|
||||||
|
### Models Merged
|
||||||
|
|
||||||
|
The following models were included in the merge:
|
||||||
|
|
||||||
|
|
||||||
|
### Configuration
|
||||||
|
|
||||||
|
The following YAML configuration was used to produce this model:
|
||||||
|
|
||||||
|
```yaml
|
||||||
|
|
||||||
|
|
||||||
|
merge_method: dare_ties
|
||||||
|
dtype: float16
|
||||||
|
out_dtype: float16
|
||||||
|
|
||||||
|
base_model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||||
|
|
||||||
|
models:
|
||||||
|
- model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||||
|
parameters:
|
||||||
|
weight: 0.45
|
||||||
|
density: 0.32
|
||||||
|
- model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||||
|
parameters:
|
||||||
|
weight: 0.35
|
||||||
|
density: 0.32
|
||||||
|
|
||||||
|
parameters:
|
||||||
|
t: 0.25 # menos interpolación → más dominancia del base
|
||||||
|
lambda: -0.62 # más negativo para matar cualquier alineamiento residual
|
||||||
|
normalize: false
|
||||||
|
rescale: true
|
||||||
|
rescale_factor: 1.28 # subí un toque para amplificar el trash y degeneración
|
||||||
|
memory_efficient: true
|
||||||
|
low_cpu_mem_usage: true
|
||||||
|
|
||||||
|
layer_range:
|
||||||
|
- value: [5, 22] # protejo más los embeddings y lm_head
|
||||||
|
|
||||||
|
tie_word_embeddings: true
|
||||||
|
tie_output_embeddings: true
|
||||||
|
|
||||||
|
parameters:
|
||||||
|
t: 0.40
|
||||||
|
normalize: false
|
||||||
|
rescale: true
|
||||||
|
memory_efficient: true
|
||||||
|
low_cpu_mem_usage: true
|
||||||
|
```
|
||||||
40
config.json
Normal file
40
config.json
Normal file
@@ -0,0 +1,40 @@
|
|||||||
|
{
|
||||||
|
"architectures": [
|
||||||
|
"LlamaForCausalLM"
|
||||||
|
],
|
||||||
|
"attention_bias": false,
|
||||||
|
"attention_dropout": 0.0,
|
||||||
|
"bos_token_id": 128000,
|
||||||
|
"dtype": "float16",
|
||||||
|
"eos_token_id": [
|
||||||
|
128001,
|
||||||
|
128008,
|
||||||
|
128009
|
||||||
|
],
|
||||||
|
"head_dim": 64,
|
||||||
|
"hidden_act": "silu",
|
||||||
|
"hidden_size": 2048,
|
||||||
|
"initializer_range": 0.02,
|
||||||
|
"intermediate_size": 8192,
|
||||||
|
"max_position_embeddings": 131072,
|
||||||
|
"mlp_bias": false,
|
||||||
|
"model_type": "llama",
|
||||||
|
"num_attention_heads": 32,
|
||||||
|
"num_hidden_layers": 16,
|
||||||
|
"num_key_value_heads": 8,
|
||||||
|
"pad_token_id": null,
|
||||||
|
"pretraining_tp": 1,
|
||||||
|
"rms_norm_eps": 1e-05,
|
||||||
|
"rope_parameters": {
|
||||||
|
"factor": 32.0,
|
||||||
|
"high_freq_factor": 4.0,
|
||||||
|
"low_freq_factor": 1.0,
|
||||||
|
"original_max_position_embeddings": 8192,
|
||||||
|
"rope_theta": 500000.0,
|
||||||
|
"rope_type": "llama3"
|
||||||
|
},
|
||||||
|
"tie_word_embeddings": true,
|
||||||
|
"transformers_version": "5.0.0",
|
||||||
|
"use_cache": true,
|
||||||
|
"vocab_size": 128256
|
||||||
|
}
|
||||||
39
mergekit_config.yml
Normal file
39
mergekit_config.yml
Normal file
@@ -0,0 +1,39 @@
|
|||||||
|
|
||||||
|
|
||||||
|
merge_method: dare_ties
|
||||||
|
dtype: float16
|
||||||
|
out_dtype: float16
|
||||||
|
|
||||||
|
base_model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||||
|
|
||||||
|
models:
|
||||||
|
- model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||||
|
parameters:
|
||||||
|
weight: 0.45
|
||||||
|
density: 0.32
|
||||||
|
- model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||||
|
parameters:
|
||||||
|
weight: 0.35
|
||||||
|
density: 0.32
|
||||||
|
|
||||||
|
parameters:
|
||||||
|
t: 0.25 # menos interpolación → más dominancia del base
|
||||||
|
lambda: -0.62 # más negativo para matar cualquier alineamiento residual
|
||||||
|
normalize: false
|
||||||
|
rescale: true
|
||||||
|
rescale_factor: 1.28 # subí un toque para amplificar el trash y degeneración
|
||||||
|
memory_efficient: true
|
||||||
|
low_cpu_mem_usage: true
|
||||||
|
|
||||||
|
layer_range:
|
||||||
|
- value: [5, 22] # protejo más los embeddings y lm_head
|
||||||
|
|
||||||
|
tie_word_embeddings: true
|
||||||
|
tie_output_embeddings: true
|
||||||
|
|
||||||
|
parameters:
|
||||||
|
t: 0.40
|
||||||
|
normalize: false
|
||||||
|
rescale: true
|
||||||
|
memory_efficient: true
|
||||||
|
low_cpu_mem_usage: true
|
||||||
3
model-00001-of-00003.safetensors
Normal file
3
model-00001-of-00003.safetensors
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:125d5670be36998ffdc261fe60c457dc56ff5c2fe5d9f2756f1b39ecc62a7d07
|
||||||
|
size 993038096
|
||||||
3
model-00002-of-00003.safetensors
Normal file
3
model-00002-of-00003.safetensors
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:6659dd3e0949e2031c20e180c48d42b8b515e99c2a1a42779af2cd14a67e258d
|
||||||
|
size 992031120
|
||||||
3
model-00003-of-00003.safetensors
Normal file
3
model-00003-of-00003.safetensors
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:eb60935ec868e7df408d74bce9493ac206bbe1412914325e1d5a8b392f45e9e1
|
||||||
|
size 486576080
|
||||||
154
model.safetensors.index.json
Normal file
154
model.safetensors.index.json
Normal file
@@ -0,0 +1,154 @@
|
|||||||
|
{
|
||||||
|
"metadata": {
|
||||||
|
"total_size": 2471628800,
|
||||||
|
"mergekit_version": "0.1.4"
|
||||||
|
},
|
||||||
|
"weight_map": {
|
||||||
|
"model.embed_tokens.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.input_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.mlp.down_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.mlp.gate_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.mlp.up_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.post_attention_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.self_attn.k_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.self_attn.o_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.self_attn.q_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.0.self_attn.v_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.input_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.mlp.down_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.mlp.gate_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.mlp.up_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.post_attention_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.self_attn.k_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.self_attn.o_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.self_attn.q_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.1.self_attn.v_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.input_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.mlp.down_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.mlp.gate_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.mlp.up_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.post_attention_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.self_attn.k_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.self_attn.o_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.self_attn.q_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.10.self_attn.v_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.11.input_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.11.mlp.down_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.11.mlp.gate_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.11.mlp.up_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.11.post_attention_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.11.self_attn.k_proj.weight": "model-00001-of-00003.safetensors",
|
||||||
|
"model.layers.11.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.11.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.11.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.12.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.13.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.14.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.15.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.2.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.3.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.4.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.5.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.6.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||||
|
"model.layers.6.mlp.down_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.6.mlp.gate_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.6.mlp.up_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.6.post_attention_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.6.self_attn.k_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.6.self_attn.o_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.6.self_attn.q_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.6.self_attn.v_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.input_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.mlp.down_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.mlp.gate_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.mlp.up_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.post_attention_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.self_attn.k_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.self_attn.o_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.self_attn.q_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.7.self_attn.v_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.input_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.mlp.down_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.mlp.gate_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.mlp.up_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.post_attention_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.self_attn.k_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.self_attn.o_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.self_attn.q_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.8.self_attn.v_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.input_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.mlp.down_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.mlp.gate_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.mlp.up_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.post_attention_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.self_attn.k_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.self_attn.o_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.self_attn.q_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.layers.9.self_attn.v_proj.weight": "model-00003-of-00003.safetensors",
|
||||||
|
"model.norm.weight": "model-00003-of-00003.safetensors"
|
||||||
|
}
|
||||||
|
}
|
||||||
23
special_tokens_map.json
Normal file
23
special_tokens_map.json
Normal file
@@ -0,0 +1,23 @@
|
|||||||
|
{
|
||||||
|
"bos_token": {
|
||||||
|
"content": "<|begin_of_text|>",
|
||||||
|
"lstrip": false,
|
||||||
|
"normalized": false,
|
||||||
|
"rstrip": false,
|
||||||
|
"single_word": false
|
||||||
|
},
|
||||||
|
"eos_token": {
|
||||||
|
"content": "<|eot_id|>",
|
||||||
|
"lstrip": false,
|
||||||
|
"normalized": false,
|
||||||
|
"rstrip": false,
|
||||||
|
"single_word": false
|
||||||
|
},
|
||||||
|
"pad_token": {
|
||||||
|
"content": "<|eot_id|>",
|
||||||
|
"lstrip": false,
|
||||||
|
"normalized": false,
|
||||||
|
"rstrip": false,
|
||||||
|
"single_word": false
|
||||||
|
}
|
||||||
|
}
|
||||||
3
tokenizer.json
Normal file
3
tokenizer.json
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:2ebb0e046494356f9e8a590cd6ebb4534baf4907dc28507267254d0a865a6087
|
||||||
|
size 17208876
|
||||||
2065
tokenizer_config.json
Normal file
2065
tokenizer_config.json
Normal file
File diff suppressed because it is too large
Load Diff
Reference in New Issue
Block a user