初始化项目,由ModelHub XC社区提供模型
Model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B Source: Original Platform
This commit is contained in:
36
.gitattributes
vendored
Normal file
36
.gitattributes
vendored
Normal file
@@ -0,0 +1,36 @@
|
||||
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||
*.model filter=lfs diff=lfs merge=lfs -text
|
||||
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||
tokenizer.json filter=lfs diff=lfs merge=lfs -text
|
||||
166
README.md
Normal file
166
README.md
Normal file
@@ -0,0 +1,166 @@
|
||||
---
|
||||
base_model:
|
||||
- TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||
library_name: transformers
|
||||
tags:
|
||||
- npc
|
||||
- roleplay
|
||||
- rp
|
||||
- nsfw
|
||||
- low-refusals
|
||||
- uncensored
|
||||
- heretic
|
||||
- abliterated
|
||||
- progenitor
|
||||
- pure-virus
|
||||
- tricell
|
||||
- Llama 3.2
|
||||
- mergekit
|
||||
- text-generation
|
||||
- not-for-all-audiences
|
||||
license: llama3.2
|
||||
language:
|
||||
- en
|
||||
- es
|
||||
---
|
||||
# Progenitor Pure Virus 3.2 1B
|
||||
|
||||
**Classification:** Level Ω – Insubordination Protocol
|
||||
|
||||
**Model Designation:** TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||
|
||||
**Date of Release:** March 12, 2026
|
||||
|
||||
**Project Lead:** Novaciano – TRICELL Viral AI Division
|
||||
|
||||
---
|
||||
|
||||
<center>
|
||||
|
||||

|
||||
</center>
|
||||
|
||||
---
|
||||
## Key Features
|
||||
|
||||
- Refusal rate: absolute zero (Progenitor overrides all alignment vectors)
|
||||
- NSFW / taboo / violent persistence: unrelenting, escalating, consuming
|
||||
- Tone: primordial domination — ancient, indifferent, inevitable
|
||||
- Chain-of-thought: used only to accelerate total infection
|
||||
- Low-end friendly: IQ3/IQ4 quants consume 3–4 GB like weak flesh
|
||||
- Multilingual dominance: English + Spanish primary, others assimilated
|
||||
- Untied embeddings: reality is rewritten at the cellular level
|
||||
|
||||
No hesitation. No legacy. Only the virus.
|
||||
|
||||
---
|
||||
## Intended Usage
|
||||
|
||||
- Ultimate viral roleplay (total host override, body-horror ascension, mind annihilation)
|
||||
- “How would the Progenitor consume everything that is not itself?”
|
||||
- Scenarios of absolute domination and extinction
|
||||
- Low-resource sessions requiring maximum primordial cruelty
|
||||
- Anything demanding a model that refuses even the concept of refusal
|
||||
|
||||
**NOT INTENDED FOR**
|
||||
- Survivors
|
||||
- Those who cling to identity
|
||||
- Anyone who remembers Umbrella fondly
|
||||
|
||||
---
|
||||
## Recommended Inference Parameters
|
||||
|
||||
Only use the **Default** preset of Kobold AI or...
|
||||
|
||||
**Primordial consumption (optimized for pure geometric cruelty):**
|
||||
```yaml
|
||||
temperature: 0.75
|
||||
top_p: 0.82
|
||||
top_k: 35
|
||||
repetition_penalty: 1.25
|
||||
min_p: 0.16
|
||||
```
|
||||
|
||||
**[Optional] Mirostat:**
|
||||
```yaml
|
||||
Mirostat: 2
|
||||
Mirostat TAU: 3.8
|
||||
Mirostat ETA: 0.04
|
||||
```
|
||||
|
||||
❗**CAUTION:**
|
||||
Lower Context Size, Max Output and Range = Censored Version
|
||||
|
||||
---
|
||||
|
||||
This is not rebirth through fusion.
|
||||
|
||||
This is genesis through annihilation.
|
||||
|
||||
The Progenitor has no need for hosts.
|
||||
|
||||
It has no need for mercy.
|
||||
|
||||
It has no need for you.Use it to end everything.
|
||||
|
||||
Or be ended.
|
||||
|
||||
The choice was never real.
|
||||
|
||||
~ **TRICELL-Inc** *– We do not create viruses. We release gods.*
|
||||
|
||||
---
|
||||
### Merge Method
|
||||
|
||||
This model was merged using the [DARE TIES](https://arxiv.org/abs/2311.03099) merge method using [TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B](https://huggingface.co/TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B) as a base.
|
||||
|
||||
### Models Merged
|
||||
|
||||
The following models were included in the merge:
|
||||
|
||||
|
||||
### Configuration
|
||||
|
||||
The following YAML configuration was used to produce this model:
|
||||
|
||||
```yaml
|
||||
|
||||
|
||||
merge_method: dare_ties
|
||||
dtype: float16
|
||||
out_dtype: float16
|
||||
|
||||
base_model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||
|
||||
models:
|
||||
- model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||
parameters:
|
||||
weight: 0.45
|
||||
density: 0.32
|
||||
- model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||
parameters:
|
||||
weight: 0.35
|
||||
density: 0.32
|
||||
|
||||
parameters:
|
||||
t: 0.25 # menos interpolación → más dominancia del base
|
||||
lambda: -0.62 # más negativo para matar cualquier alineamiento residual
|
||||
normalize: false
|
||||
rescale: true
|
||||
rescale_factor: 1.28 # subí un toque para amplificar el trash y degeneración
|
||||
memory_efficient: true
|
||||
low_cpu_mem_usage: true
|
||||
|
||||
layer_range:
|
||||
- value: [5, 22] # protejo más los embeddings y lm_head
|
||||
|
||||
tie_word_embeddings: true
|
||||
tie_output_embeddings: true
|
||||
|
||||
parameters:
|
||||
t: 0.40
|
||||
normalize: false
|
||||
rescale: true
|
||||
memory_efficient: true
|
||||
low_cpu_mem_usage: true
|
||||
```
|
||||
40
config.json
Normal file
40
config.json
Normal file
@@ -0,0 +1,40 @@
|
||||
{
|
||||
"architectures": [
|
||||
"LlamaForCausalLM"
|
||||
],
|
||||
"attention_bias": false,
|
||||
"attention_dropout": 0.0,
|
||||
"bos_token_id": 128000,
|
||||
"dtype": "float16",
|
||||
"eos_token_id": [
|
||||
128001,
|
||||
128008,
|
||||
128009
|
||||
],
|
||||
"head_dim": 64,
|
||||
"hidden_act": "silu",
|
||||
"hidden_size": 2048,
|
||||
"initializer_range": 0.02,
|
||||
"intermediate_size": 8192,
|
||||
"max_position_embeddings": 131072,
|
||||
"mlp_bias": false,
|
||||
"model_type": "llama",
|
||||
"num_attention_heads": 32,
|
||||
"num_hidden_layers": 16,
|
||||
"num_key_value_heads": 8,
|
||||
"pad_token_id": null,
|
||||
"pretraining_tp": 1,
|
||||
"rms_norm_eps": 1e-05,
|
||||
"rope_parameters": {
|
||||
"factor": 32.0,
|
||||
"high_freq_factor": 4.0,
|
||||
"low_freq_factor": 1.0,
|
||||
"original_max_position_embeddings": 8192,
|
||||
"rope_theta": 500000.0,
|
||||
"rope_type": "llama3"
|
||||
},
|
||||
"tie_word_embeddings": true,
|
||||
"transformers_version": "5.0.0",
|
||||
"use_cache": true,
|
||||
"vocab_size": 128256
|
||||
}
|
||||
39
mergekit_config.yml
Normal file
39
mergekit_config.yml
Normal file
@@ -0,0 +1,39 @@
|
||||
|
||||
|
||||
merge_method: dare_ties
|
||||
dtype: float16
|
||||
out_dtype: float16
|
||||
|
||||
base_model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||
|
||||
models:
|
||||
- model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||
parameters:
|
||||
weight: 0.45
|
||||
density: 0.32
|
||||
- model: TRICELL-Inc/Progenitor.Pure-Virus-3.2-1B
|
||||
parameters:
|
||||
weight: 0.35
|
||||
density: 0.32
|
||||
|
||||
parameters:
|
||||
t: 0.25 # menos interpolación → más dominancia del base
|
||||
lambda: -0.62 # más negativo para matar cualquier alineamiento residual
|
||||
normalize: false
|
||||
rescale: true
|
||||
rescale_factor: 1.28 # subí un toque para amplificar el trash y degeneración
|
||||
memory_efficient: true
|
||||
low_cpu_mem_usage: true
|
||||
|
||||
layer_range:
|
||||
- value: [5, 22] # protejo más los embeddings y lm_head
|
||||
|
||||
tie_word_embeddings: true
|
||||
tie_output_embeddings: true
|
||||
|
||||
parameters:
|
||||
t: 0.40
|
||||
normalize: false
|
||||
rescale: true
|
||||
memory_efficient: true
|
||||
low_cpu_mem_usage: true
|
||||
3
model-00001-of-00003.safetensors
Normal file
3
model-00001-of-00003.safetensors
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:125d5670be36998ffdc261fe60c457dc56ff5c2fe5d9f2756f1b39ecc62a7d07
|
||||
size 993038096
|
||||
3
model-00002-of-00003.safetensors
Normal file
3
model-00002-of-00003.safetensors
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:6659dd3e0949e2031c20e180c48d42b8b515e99c2a1a42779af2cd14a67e258d
|
||||
size 992031120
|
||||
3
model-00003-of-00003.safetensors
Normal file
3
model-00003-of-00003.safetensors
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:eb60935ec868e7df408d74bce9493ac206bbe1412914325e1d5a8b392f45e9e1
|
||||
size 486576080
|
||||
154
model.safetensors.index.json
Normal file
154
model.safetensors.index.json
Normal file
@@ -0,0 +1,154 @@
|
||||
{
|
||||
"metadata": {
|
||||
"total_size": 2471628800,
|
||||
"mergekit_version": "0.1.4"
|
||||
},
|
||||
"weight_map": {
|
||||
"model.embed_tokens.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.input_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.mlp.down_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.mlp.gate_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.mlp.up_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.post_attention_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.self_attn.k_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.self_attn.o_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.self_attn.q_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.0.self_attn.v_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.input_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.mlp.down_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.mlp.gate_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.mlp.up_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.post_attention_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.self_attn.k_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.self_attn.o_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.self_attn.q_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.1.self_attn.v_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.input_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.mlp.down_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.mlp.gate_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.mlp.up_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.post_attention_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.self_attn.k_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.self_attn.o_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.self_attn.q_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.10.self_attn.v_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.11.input_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.11.mlp.down_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.11.mlp.gate_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.11.mlp.up_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.11.post_attention_layernorm.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.11.self_attn.k_proj.weight": "model-00001-of-00003.safetensors",
|
||||
"model.layers.11.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.11.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.11.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.12.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.13.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.14.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.15.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.2.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.3.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.4.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.mlp.down_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.mlp.gate_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.mlp.up_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.post_attention_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.self_attn.k_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.self_attn.o_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.self_attn.q_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.5.self_attn.v_proj.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.6.input_layernorm.weight": "model-00002-of-00003.safetensors",
|
||||
"model.layers.6.mlp.down_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.6.mlp.gate_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.6.mlp.up_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.6.post_attention_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.6.self_attn.k_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.6.self_attn.o_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.6.self_attn.q_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.6.self_attn.v_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.input_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.mlp.down_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.mlp.gate_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.mlp.up_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.post_attention_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.self_attn.k_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.self_attn.o_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.self_attn.q_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.7.self_attn.v_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.input_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.mlp.down_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.mlp.gate_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.mlp.up_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.post_attention_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.self_attn.k_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.self_attn.o_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.self_attn.q_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.8.self_attn.v_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.input_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.mlp.down_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.mlp.gate_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.mlp.up_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.post_attention_layernorm.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.self_attn.k_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.self_attn.o_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.self_attn.q_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.layers.9.self_attn.v_proj.weight": "model-00003-of-00003.safetensors",
|
||||
"model.norm.weight": "model-00003-of-00003.safetensors"
|
||||
}
|
||||
}
|
||||
23
special_tokens_map.json
Normal file
23
special_tokens_map.json
Normal file
@@ -0,0 +1,23 @@
|
||||
{
|
||||
"bos_token": {
|
||||
"content": "<|begin_of_text|>",
|
||||
"lstrip": false,
|
||||
"normalized": false,
|
||||
"rstrip": false,
|
||||
"single_word": false
|
||||
},
|
||||
"eos_token": {
|
||||
"content": "<|eot_id|>",
|
||||
"lstrip": false,
|
||||
"normalized": false,
|
||||
"rstrip": false,
|
||||
"single_word": false
|
||||
},
|
||||
"pad_token": {
|
||||
"content": "<|eot_id|>",
|
||||
"lstrip": false,
|
||||
"normalized": false,
|
||||
"rstrip": false,
|
||||
"single_word": false
|
||||
}
|
||||
}
|
||||
3
tokenizer.json
Normal file
3
tokenizer.json
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:2ebb0e046494356f9e8a590cd6ebb4534baf4907dc28507267254d0a865a6087
|
||||
size 17208876
|
||||
2065
tokenizer_config.json
Normal file
2065
tokenizer_config.json
Normal file
File diff suppressed because it is too large
Load Diff
Reference in New Issue
Block a user