初始化项目,由ModelHub XC社区提供模型

Model: KoboldAI/LLAMA2-13B-Holodeck-1
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-11 09:34:06 +08:00
commit 42391cd65e
40 changed files with 761 additions and 0 deletions

49
.gitattributes vendored Normal file
View File

@@ -0,0 +1,49 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bin.* filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zstandard filter=lfs diff=lfs merge=lfs -text
*.tfevents* filter=lfs diff=lfs merge=lfs -text
*.db* filter=lfs diff=lfs merge=lfs -text
*.ark* filter=lfs diff=lfs merge=lfs -text
**/*ckpt*data* filter=lfs diff=lfs merge=lfs -text
**/*ckpt*.meta filter=lfs diff=lfs merge=lfs -text
**/*ckpt*.index filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.gguf* filter=lfs diff=lfs merge=lfs -text
*.ggml filter=lfs diff=lfs merge=lfs -text
*.llamafile* filter=lfs diff=lfs merge=lfs -text
*.pt2 filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
tokenizer.json filter=lfs diff=lfs merge=lfs -text

125
LICENSE.md Normal file
View File

@@ -0,0 +1,125 @@
LLAMA 2 COMMUNITY LICENSE AGREEMENT
Llama 2 Version Release Date: July 18, 2023
"Agreement" means the terms and conditions for use, reproduction, distribution and
modification of the Llama Materials set forth herein.
"Documentation" means the specifications, manuals and documentation
accompanying Llama 2 distributed by Meta at ai.meta.com/resources/models-and-
libraries/llama-downloads/.
"Licensee" or "you" means you, or your employer or any other person or entity (if
you are entering into this Agreement on such person or entity's behalf), of the age
required under applicable laws, rules or regulations to provide legal consent and that
has legal authority to bind your employer or such other person or entity if you are
entering in this Agreement on their behalf.
"Llama 2" means the foundational large language models and software and
algorithms, including machine-learning model code, trained model weights,
inference-enabling code, training-enabling code, fine-tuning enabling code and other
elements of the foregoing distributed by Meta at ai.meta.com/resources/models-and-
libraries/llama-downloads/.
"Llama Materials" means, collectively, Meta's proprietary Llama 2 and
Documentation (and any portion thereof) made available under this Agreement.
"Meta" or "we" means Meta Platforms Ireland Limited (if you are located in or, if you
are an entity, your principal place of business is in the EEA or Switzerland) and Meta
Platforms, Inc. (if you are located outside of the EEA or Switzerland).
By clicking "I Accept" below or by using or distributing any portion or element of the
Llama Materials, you agree to be bound by this Agreement.
1. License Rights and Redistribution.
a. Grant of Rights. You are granted a non-exclusive, worldwide, non-
transferable and royalty-free limited license under Meta's intellectual property or
other rights owned by Meta embodied in the Llama Materials to use, reproduce,
distribute, copy, create derivative works of, and make modifications to the Llama
Materials.
b. Redistribution and Use.
i. If you distribute or make the Llama Materials, or any derivative works
thereof, available to a third party, you shall provide a copy of this Agreement to such
third party.
ii. If you receive Llama Materials, or any derivative works thereof, from
a Licensee as part of an integrated end user product, then Section 2 of this
Agreement will not apply to you.
iii. You must retain in all copies of the Llama Materials that you
distribute the following attribution notice within a "Notice" text file distributed as a
part of such copies: "Llama 2 is licensed under the LLAMA 2 Community License,
Copyright (c) Meta Platforms, Inc. All Rights Reserved."
iv. Your use of the Llama Materials must comply with applicable laws
and regulations (including trade compliance laws and regulations) and adhere to the
Acceptable Use Policy for the Llama Materials (available at
https://ai.meta.com/llama/use-policy), which is hereby incorporated by reference into
this Agreement.
v. You will not use the Llama Materials or any output or results of the
Llama Materials to improve any other large language model (excluding Llama 2 or
derivative works thereof).
2. Additional Commercial Terms. If, on the Llama 2 version release date, the
monthly active users of the products or services made available by or for Licensee,
or Licensee's affiliates, is greater than 700 million monthly active users in the
preceding calendar month, you must request a license from Meta, which Meta may
grant to you in its sole discretion, and you are not authorized to exercise any of the
rights under this Agreement unless or until Meta otherwise expressly grants you
such rights.
3. Disclaimer of Warranty. UNLESS REQUIRED BY APPLICABLE LAW, THE
LLAMA MATERIALS AND ANY OUTPUT AND RESULTS THEREFROM ARE
PROVIDED ON AN "AS IS" BASIS, WITHOUT WARRANTIES OF ANY KIND,
EITHER EXPRESS OR IMPLIED, INCLUDING, WITHOUT LIMITATION, ANY
WARRANTIES OF TITLE, NON-INFRINGEMENT, MERCHANTABILITY, OR
FITNESS FOR A PARTICULAR PURPOSE. YOU ARE SOLELY RESPONSIBLE
FOR DETERMINING THE APPROPRIATENESS OF USING OR REDISTRIBUTING
THE LLAMA MATERIALS AND ASSUME ANY RISKS ASSOCIATED WITH YOUR
USE OF THE LLAMA MATERIALS AND ANY OUTPUT AND RESULTS.
4. Limitation of Liability. IN NO EVENT WILL META OR ITS AFFILIATES BE
LIABLE UNDER ANY THEORY OF LIABILITY, WHETHER IN CONTRACT, TORT,
NEGLIGENCE, PRODUCTS LIABILITY, OR OTHERWISE, ARISING OUT OF THIS
AGREEMENT, FOR ANY LOST PROFITS OR ANY INDIRECT, SPECIAL,
CONSEQUENTIAL, INCIDENTAL, EXEMPLARY OR PUNITIVE DAMAGES, EVEN
IF META OR ITS AFFILIATES HAVE BEEN ADVISED OF THE POSSIBILITY OF
ANY OF THE FOREGOING.
5. Intellectual Property.
a. No trademark licenses are granted under this Agreement, and in
connection with the Llama Materials, neither Meta nor Licensee may use any name
or mark owned by or associated with the other or any of its affiliates, except as
required for reasonable and customary use in describing and redistributing the
Llama Materials.
b. Subject to Meta's ownership of Llama Materials and derivatives made by or
for Meta, with respect to any derivative works and modifications of the Llama
Materials that are made by you, as between you and Meta, you are and will be the
owner of such derivative works and modifications.
c. If you institute litigation or other proceedings against Meta or any entity
(including a cross-claim or counterclaim in a lawsuit) alleging that the Llama
Materials or Llama 2 outputs or results, or any portion of any of the foregoing,
constitutes infringement of intellectual property or other rights owned or licensable
by you, then any licenses granted to you under this Agreement shall terminate as of
the date such litigation or claim is filed or instituted. You will indemnify and hold
harmless Meta from and against any claim by any third party arising out of or related
to your use or distribution of the Llama Materials.
6. Term and Termination. The term of this Agreement will commence upon your
acceptance of this Agreement or access to the Llama Materials and will continue in
full force and effect until terminated in accordance with the terms and conditions
herein. Meta may terminate this Agreement if you are in breach of any term or
condition of this Agreement. Upon termination of this Agreement, you shall delete
and cease use of the Llama Materials. Sections 3, 4 and 7 shall survive the
termination of this Agreement.
7. Governing Law and Jurisdiction. This Agreement will be governed and
construed under the laws of the State of California without regard to choice of law
principles, and the UN Convention on Contracts for the International Sale of Goods
does not apply to this Agreement. The courts of California shall have exclusive
jurisdiction of any dispute arising out of this Agreement.

31
README.md Normal file
View File

@@ -0,0 +1,31 @@
---
license: other
language: en
commercial: no
inference: true
---
# LLAMA2 13B - Holodeck
## Model Description
LLAMA2 13B-Holodeck is a finetune created using Meta's llama 2 model.
## Training data
The training data contains around 3000 ebooks in various genres.
Most parts of the dataset have been prepended using the following text: `[Genre: <genre1>, <genre2>]`
### How to use
You can use this model directly with a pipeline for text generation. This example generates a different sequence each time it's run:
```py
>>> from transformers import pipeline
>>> generator = pipeline('text-generation', model='KoboldAI/LLAMA2-13B-Holodeck-1')
>>> generator("Welcome Captain Janeway, I apologize for the delay.", do_sample=True, min_length=50)
[{'generated_text': 'Welcome Captain Janeway, I apologize for the delay."\nIt's all right," Janeway said. "I'm certain that you're doing your best to keep me informed of what\'s going on."'}]
```
### Limitations and Biases
Based on known problems with NLP technology, potential relevant factors include bias (gender, profession, race and religion).
### License
Llama 2 is licensed under the LLAMA 2 Community License, Copyright (c) Meta Platforms, Inc. All Rights Reserved.
**Extra clause:**
You shall use the Materials and Products solely for research purposes or personal use and not for any commercial purpose. Nothing in the Community License shall be construed as granting you a license to use the Materials or Products for any other purpose.
### BibTeX entry and citation info
https://huggingface.co/meta-llama/Llama-2-13b-hf

26
config.json Normal file
View File

@@ -0,0 +1,26 @@
{
"_name_or_path": "mrseeker/llama2-13b-pike-v2-2",
"architectures": [
"LlamaForCausalLM"
],
"bos_token_id": 1,
"eos_token_id": 2,
"hidden_act": "silu",
"hidden_size": 5120,
"initializer_range": 0.02,
"intermediate_size": 13824,
"max_position_embeddings": 4096,
"model_type": "llama",
"num_attention_heads": 40,
"num_hidden_layers": 40,
"num_key_value_heads": 40,
"pad_token_id": 0,
"pretraining_tp": 1,
"rms_norm_eps": 1e-05,
"rope_scaling": null,
"tie_word_embeddings": false,
"torch_dtype": "float16",
"transformers_version": "4.32.0.dev0",
"use_cache": false,
"vocab_size": 32000
}

1
configuration.json Normal file
View File

@@ -0,0 +1 @@
{"framework": "pytorch", "task": "text-generation", "allow_remote": true}

9
generation_config.json Normal file
View File

@@ -0,0 +1,9 @@
{
"bos_token_id": 1,
"eos_token_id": 2,
"max_length": 4096,
"pad_token_id": 0,
"temperature": 0.9,
"top_p": 0.6,
"transformers_version": "4.32.0.dev0"
}

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3d3eb91816affeb64fc1f1a4e01a86eaf6a9f1df84d1f99507dc82312f84304a
size 1947773640

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:5080365135be5097e54c945ba2d9989fad541127ac6d92c8d9cb60d67cca73ba
size 1903229976

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:30218f9b6ab6af8fe6ba88378add37d45d5e282cfab17969e1bb8887648981a0
size 1903229976

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0f22c8a3f8b3dc8001182018f6b3087498e2e47bff8f9ecbfcba9a1cb717dc63
size 1903229992

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2392331eb3fc7212ec02734d819fbad568467ae7c21897edfb73c421928d4fc6
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f4de532f58a6829499f8681e957cfdc45b59331a616f1cf07a73305a22b2b23a
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c544675d8a94c1a685b0e717c2646134618e2249da93f529f246248110f5233a
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0116341747a3a1bde448abe4a533081775c11a842ca69a8fefa321af34d0b33a
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f8c9948db79a0a729337922806b558647c94e94e9a3a03a69505562c428c6908
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8915f00c9e86f552d5302469bc69b43568fae189efdc330d7f0f75f2692a1200
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b9a80aa4a01a58c3d95aba31f4c2927456bdedf5e0b3f501e154a438cd10ecac
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2465f4ddbc09c1a32c860ff9769e6aeee121497a85f72b63d42297047a91849c
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1c7f46e8a89cf8432229e7838a8c15b706fb9f36620987d19380a42b5b7bce0b
size 1903230008

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f5055f20b22d2b17f95dec4fa5d07f3be9648c0132a0575f0b0f7aa3338d11fd
size 1245236904

View File

@@ -0,0 +1,370 @@
{
"metadata": {
"total_size": 26031728640
},
"weight_map": {
"lm_head.weight": "model-00014-of-00014.safetensors",
"model.embed_tokens.weight": "model-00001-of-00014.safetensors",
"model.layers.0.input_layernorm.weight": "model-00001-of-00014.safetensors",
"model.layers.0.mlp.down_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.0.mlp.gate_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.0.mlp.up_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.0.post_attention_layernorm.weight": "model-00001-of-00014.safetensors",
"model.layers.0.self_attn.k_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.0.self_attn.o_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.0.self_attn.q_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.0.self_attn.v_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.1.input_layernorm.weight": "model-00001-of-00014.safetensors",
"model.layers.1.mlp.down_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.1.mlp.gate_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.1.mlp.up_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.1.post_attention_layernorm.weight": "model-00001-of-00014.safetensors",
"model.layers.1.self_attn.k_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.1.self_attn.o_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.1.self_attn.q_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.1.self_attn.v_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.10.input_layernorm.weight": "model-00004-of-00014.safetensors",
"model.layers.10.mlp.down_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.10.mlp.gate_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.10.mlp.up_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.10.post_attention_layernorm.weight": "model-00004-of-00014.safetensors",
"model.layers.10.self_attn.k_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.10.self_attn.o_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.10.self_attn.q_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.10.self_attn.v_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.11.input_layernorm.weight": "model-00005-of-00014.safetensors",
"model.layers.11.mlp.down_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.11.mlp.gate_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.11.mlp.up_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.11.post_attention_layernorm.weight": "model-00005-of-00014.safetensors",
"model.layers.11.self_attn.k_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.11.self_attn.o_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.11.self_attn.q_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.11.self_attn.v_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.12.input_layernorm.weight": "model-00005-of-00014.safetensors",
"model.layers.12.mlp.down_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.12.mlp.gate_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.12.mlp.up_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.12.post_attention_layernorm.weight": "model-00005-of-00014.safetensors",
"model.layers.12.self_attn.k_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.12.self_attn.o_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.12.self_attn.q_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.12.self_attn.v_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.13.input_layernorm.weight": "model-00005-of-00014.safetensors",
"model.layers.13.mlp.down_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.13.mlp.gate_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.13.mlp.up_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.13.post_attention_layernorm.weight": "model-00005-of-00014.safetensors",
"model.layers.13.self_attn.k_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.13.self_attn.o_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.13.self_attn.q_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.13.self_attn.v_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.14.input_layernorm.weight": "model-00006-of-00014.safetensors",
"model.layers.14.mlp.down_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.14.mlp.gate_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.14.mlp.up_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.14.post_attention_layernorm.weight": "model-00006-of-00014.safetensors",
"model.layers.14.self_attn.k_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.14.self_attn.o_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.14.self_attn.q_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.14.self_attn.v_proj.weight": "model-00005-of-00014.safetensors",
"model.layers.15.input_layernorm.weight": "model-00006-of-00014.safetensors",
"model.layers.15.mlp.down_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.15.mlp.gate_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.15.mlp.up_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.15.post_attention_layernorm.weight": "model-00006-of-00014.safetensors",
"model.layers.15.self_attn.k_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.15.self_attn.o_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.15.self_attn.q_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.15.self_attn.v_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.16.input_layernorm.weight": "model-00006-of-00014.safetensors",
"model.layers.16.mlp.down_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.16.mlp.gate_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.16.mlp.up_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.16.post_attention_layernorm.weight": "model-00006-of-00014.safetensors",
"model.layers.16.self_attn.k_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.16.self_attn.o_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.16.self_attn.q_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.16.self_attn.v_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.17.input_layernorm.weight": "model-00007-of-00014.safetensors",
"model.layers.17.mlp.down_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.17.mlp.gate_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.17.mlp.up_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.17.post_attention_layernorm.weight": "model-00007-of-00014.safetensors",
"model.layers.17.self_attn.k_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.17.self_attn.o_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.17.self_attn.q_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.17.self_attn.v_proj.weight": "model-00006-of-00014.safetensors",
"model.layers.18.input_layernorm.weight": "model-00007-of-00014.safetensors",
"model.layers.18.mlp.down_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.18.mlp.gate_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.18.mlp.up_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.18.post_attention_layernorm.weight": "model-00007-of-00014.safetensors",
"model.layers.18.self_attn.k_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.18.self_attn.o_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.18.self_attn.q_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.18.self_attn.v_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.19.input_layernorm.weight": "model-00007-of-00014.safetensors",
"model.layers.19.mlp.down_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.19.mlp.gate_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.19.mlp.up_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.19.post_attention_layernorm.weight": "model-00007-of-00014.safetensors",
"model.layers.19.self_attn.k_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.19.self_attn.o_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.19.self_attn.q_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.19.self_attn.v_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.2.input_layernorm.weight": "model-00002-of-00014.safetensors",
"model.layers.2.mlp.down_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.2.mlp.gate_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.2.mlp.up_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.2.post_attention_layernorm.weight": "model-00002-of-00014.safetensors",
"model.layers.2.self_attn.k_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.2.self_attn.o_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.2.self_attn.q_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.2.self_attn.v_proj.weight": "model-00001-of-00014.safetensors",
"model.layers.20.input_layernorm.weight": "model-00008-of-00014.safetensors",
"model.layers.20.mlp.down_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.20.mlp.gate_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.20.mlp.up_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.20.post_attention_layernorm.weight": "model-00008-of-00014.safetensors",
"model.layers.20.self_attn.k_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.20.self_attn.o_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.20.self_attn.q_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.20.self_attn.v_proj.weight": "model-00007-of-00014.safetensors",
"model.layers.21.input_layernorm.weight": "model-00008-of-00014.safetensors",
"model.layers.21.mlp.down_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.21.mlp.gate_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.21.mlp.up_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.21.post_attention_layernorm.weight": "model-00008-of-00014.safetensors",
"model.layers.21.self_attn.k_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.21.self_attn.o_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.21.self_attn.q_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.21.self_attn.v_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.22.input_layernorm.weight": "model-00008-of-00014.safetensors",
"model.layers.22.mlp.down_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.22.mlp.gate_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.22.mlp.up_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.22.post_attention_layernorm.weight": "model-00008-of-00014.safetensors",
"model.layers.22.self_attn.k_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.22.self_attn.o_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.22.self_attn.q_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.22.self_attn.v_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.23.input_layernorm.weight": "model-00009-of-00014.safetensors",
"model.layers.23.mlp.down_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.23.mlp.gate_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.23.mlp.up_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.23.post_attention_layernorm.weight": "model-00009-of-00014.safetensors",
"model.layers.23.self_attn.k_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.23.self_attn.o_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.23.self_attn.q_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.23.self_attn.v_proj.weight": "model-00008-of-00014.safetensors",
"model.layers.24.input_layernorm.weight": "model-00009-of-00014.safetensors",
"model.layers.24.mlp.down_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.24.mlp.gate_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.24.mlp.up_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.24.post_attention_layernorm.weight": "model-00009-of-00014.safetensors",
"model.layers.24.self_attn.k_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.24.self_attn.o_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.24.self_attn.q_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.24.self_attn.v_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.25.input_layernorm.weight": "model-00009-of-00014.safetensors",
"model.layers.25.mlp.down_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.25.mlp.gate_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.25.mlp.up_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.25.post_attention_layernorm.weight": "model-00009-of-00014.safetensors",
"model.layers.25.self_attn.k_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.25.self_attn.o_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.25.self_attn.q_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.25.self_attn.v_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.26.input_layernorm.weight": "model-00010-of-00014.safetensors",
"model.layers.26.mlp.down_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.26.mlp.gate_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.26.mlp.up_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.26.post_attention_layernorm.weight": "model-00010-of-00014.safetensors",
"model.layers.26.self_attn.k_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.26.self_attn.o_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.26.self_attn.q_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.26.self_attn.v_proj.weight": "model-00009-of-00014.safetensors",
"model.layers.27.input_layernorm.weight": "model-00010-of-00014.safetensors",
"model.layers.27.mlp.down_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.27.mlp.gate_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.27.mlp.up_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.27.post_attention_layernorm.weight": "model-00010-of-00014.safetensors",
"model.layers.27.self_attn.k_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.27.self_attn.o_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.27.self_attn.q_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.27.self_attn.v_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.28.input_layernorm.weight": "model-00010-of-00014.safetensors",
"model.layers.28.mlp.down_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.28.mlp.gate_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.28.mlp.up_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.28.post_attention_layernorm.weight": "model-00010-of-00014.safetensors",
"model.layers.28.self_attn.k_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.28.self_attn.o_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.28.self_attn.q_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.28.self_attn.v_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.29.input_layernorm.weight": "model-00011-of-00014.safetensors",
"model.layers.29.mlp.down_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.29.mlp.gate_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.29.mlp.up_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.29.post_attention_layernorm.weight": "model-00011-of-00014.safetensors",
"model.layers.29.self_attn.k_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.29.self_attn.o_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.29.self_attn.q_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.29.self_attn.v_proj.weight": "model-00010-of-00014.safetensors",
"model.layers.3.input_layernorm.weight": "model-00002-of-00014.safetensors",
"model.layers.3.mlp.down_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.3.mlp.gate_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.3.mlp.up_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.3.post_attention_layernorm.weight": "model-00002-of-00014.safetensors",
"model.layers.3.self_attn.k_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.3.self_attn.o_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.3.self_attn.q_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.3.self_attn.v_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.30.input_layernorm.weight": "model-00011-of-00014.safetensors",
"model.layers.30.mlp.down_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.30.mlp.gate_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.30.mlp.up_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.30.post_attention_layernorm.weight": "model-00011-of-00014.safetensors",
"model.layers.30.self_attn.k_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.30.self_attn.o_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.30.self_attn.q_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.30.self_attn.v_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.31.input_layernorm.weight": "model-00011-of-00014.safetensors",
"model.layers.31.mlp.down_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.31.mlp.gate_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.31.mlp.up_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.31.post_attention_layernorm.weight": "model-00011-of-00014.safetensors",
"model.layers.31.self_attn.k_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.31.self_attn.o_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.31.self_attn.q_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.31.self_attn.v_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.32.input_layernorm.weight": "model-00012-of-00014.safetensors",
"model.layers.32.mlp.down_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.32.mlp.gate_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.32.mlp.up_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.32.post_attention_layernorm.weight": "model-00012-of-00014.safetensors",
"model.layers.32.self_attn.k_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.32.self_attn.o_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.32.self_attn.q_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.32.self_attn.v_proj.weight": "model-00011-of-00014.safetensors",
"model.layers.33.input_layernorm.weight": "model-00012-of-00014.safetensors",
"model.layers.33.mlp.down_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.33.mlp.gate_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.33.mlp.up_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.33.post_attention_layernorm.weight": "model-00012-of-00014.safetensors",
"model.layers.33.self_attn.k_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.33.self_attn.o_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.33.self_attn.q_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.33.self_attn.v_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.34.input_layernorm.weight": "model-00012-of-00014.safetensors",
"model.layers.34.mlp.down_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.34.mlp.gate_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.34.mlp.up_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.34.post_attention_layernorm.weight": "model-00012-of-00014.safetensors",
"model.layers.34.self_attn.k_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.34.self_attn.o_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.34.self_attn.q_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.34.self_attn.v_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.35.input_layernorm.weight": "model-00013-of-00014.safetensors",
"model.layers.35.mlp.down_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.35.mlp.gate_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.35.mlp.up_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.35.post_attention_layernorm.weight": "model-00013-of-00014.safetensors",
"model.layers.35.self_attn.k_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.35.self_attn.o_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.35.self_attn.q_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.35.self_attn.v_proj.weight": "model-00012-of-00014.safetensors",
"model.layers.36.input_layernorm.weight": "model-00013-of-00014.safetensors",
"model.layers.36.mlp.down_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.36.mlp.gate_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.36.mlp.up_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.36.post_attention_layernorm.weight": "model-00013-of-00014.safetensors",
"model.layers.36.self_attn.k_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.36.self_attn.o_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.36.self_attn.q_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.36.self_attn.v_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.37.input_layernorm.weight": "model-00013-of-00014.safetensors",
"model.layers.37.mlp.down_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.37.mlp.gate_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.37.mlp.up_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.37.post_attention_layernorm.weight": "model-00013-of-00014.safetensors",
"model.layers.37.self_attn.k_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.37.self_attn.o_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.37.self_attn.q_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.37.self_attn.v_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.38.input_layernorm.weight": "model-00014-of-00014.safetensors",
"model.layers.38.mlp.down_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.38.mlp.gate_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.38.mlp.up_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.38.post_attention_layernorm.weight": "model-00014-of-00014.safetensors",
"model.layers.38.self_attn.k_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.38.self_attn.o_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.38.self_attn.q_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.38.self_attn.v_proj.weight": "model-00013-of-00014.safetensors",
"model.layers.39.input_layernorm.weight": "model-00014-of-00014.safetensors",
"model.layers.39.mlp.down_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.39.mlp.gate_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.39.mlp.up_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.39.post_attention_layernorm.weight": "model-00014-of-00014.safetensors",
"model.layers.39.self_attn.k_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.39.self_attn.o_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.39.self_attn.q_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.39.self_attn.v_proj.weight": "model-00014-of-00014.safetensors",
"model.layers.4.input_layernorm.weight": "model-00002-of-00014.safetensors",
"model.layers.4.mlp.down_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.4.mlp.gate_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.4.mlp.up_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.4.post_attention_layernorm.weight": "model-00002-of-00014.safetensors",
"model.layers.4.self_attn.k_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.4.self_attn.o_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.4.self_attn.q_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.4.self_attn.v_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.5.input_layernorm.weight": "model-00003-of-00014.safetensors",
"model.layers.5.mlp.down_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.5.mlp.gate_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.5.mlp.up_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.5.post_attention_layernorm.weight": "model-00003-of-00014.safetensors",
"model.layers.5.self_attn.k_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.5.self_attn.o_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.5.self_attn.q_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.5.self_attn.v_proj.weight": "model-00002-of-00014.safetensors",
"model.layers.6.input_layernorm.weight": "model-00003-of-00014.safetensors",
"model.layers.6.mlp.down_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.6.mlp.gate_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.6.mlp.up_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.6.post_attention_layernorm.weight": "model-00003-of-00014.safetensors",
"model.layers.6.self_attn.k_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.6.self_attn.o_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.6.self_attn.q_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.6.self_attn.v_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.7.input_layernorm.weight": "model-00003-of-00014.safetensors",
"model.layers.7.mlp.down_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.7.mlp.gate_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.7.mlp.up_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.7.post_attention_layernorm.weight": "model-00003-of-00014.safetensors",
"model.layers.7.self_attn.k_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.7.self_attn.o_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.7.self_attn.q_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.7.self_attn.v_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.8.input_layernorm.weight": "model-00004-of-00014.safetensors",
"model.layers.8.mlp.down_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.8.mlp.gate_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.8.mlp.up_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.8.post_attention_layernorm.weight": "model-00004-of-00014.safetensors",
"model.layers.8.self_attn.k_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.8.self_attn.o_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.8.self_attn.q_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.8.self_attn.v_proj.weight": "model-00003-of-00014.safetensors",
"model.layers.9.input_layernorm.weight": "model-00004-of-00014.safetensors",
"model.layers.9.mlp.down_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.9.mlp.gate_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.9.mlp.up_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.9.post_attention_layernorm.weight": "model-00004-of-00014.safetensors",
"model.layers.9.self_attn.k_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.9.self_attn.o_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.9.self_attn.q_proj.weight": "model-00004-of-00014.safetensors",
"model.layers.9.self_attn.v_proj.weight": "model-00004-of-00014.safetensors",
"model.norm.weight": "model-00014-of-00014.safetensors"
}
}

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bde146f45c778df6532713881cd9bbf5b76ede26b17eaba948cc8bfcb1f20a91
size 1947779007

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f038273e8c54913e92a7d8d0c4a8e976ea30c422eb1a442bf65aa4185949a16f
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:edbb93511d0b62bc14bc984462d50dcf267659adcb567c0a9a71320e09f3de36
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1df35468a0f604a84a3ea16e0422d01345cd2845e9b7e38d59eab607a61ba0af
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:92bee1cbf712311b5ee758644b683c4188a7f36dcbdcca3242970b7e50096813
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d7ecde442b249ff36852709f95bd89814973cb677629a3e63331715c79812362
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:86c544fa9ddc611c8eb8ba46bf9b5ca8d8bca051f52047fdacbb987988c16ab9
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d0c91a9016d0ebf239b12e0bd5c3eeaeec48ae08e825ea7b03e582db30351979
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:75f1bb4728833877060d12e61971c04b1841b6693f19d66c5a69672a7d0fb20a
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a3a3caf9553bbcabbf0044d0c857371bb840ba1622ff3c5f16feae5544174626
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:dbd924169fd93b6a71033a8ed5dac38f409aeac954e90906001e4fed90615f84
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:185a51f31a5baff58a106da322b9299d27826790e2740ade35ee828a557fee65
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bc016fc3dd5fc816c688c8924069b1c832d5c0e9a97e82ba5ffcbf8aa05d2f62
size 1903235957

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f0a8779b8b848cd1296ad49f038e23c080f0401d33d23c3d53c7fef602a6c096
size 1245240477

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d53ecdeb07c7983b6cad0b68e676703399582edc0cc6ecc6cbfaebcb47ec847a
size 29894

23
special_tokens_map.json Normal file
View File

@@ -0,0 +1,23 @@
{
"bos_token": {
"content": "<s>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
},
"eos_token": {
"content": "</s>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
},
"unk_token": {
"content": "<unk>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
}
}

BIN
tokenizer.json (Stored with Git LFS) Normal file

Binary file not shown.

BIN
tokenizer.model (Stored with Git LFS) Normal file

Binary file not shown.

34
tokenizer_config.json Normal file
View File

@@ -0,0 +1,34 @@
{
"add_bos_token": true,
"add_eos_token": false,
"bos_token": {
"__type": "AddedToken",
"content": "<s>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
},
"clean_up_tokenization_spaces": false,
"eos_token": {
"__type": "AddedToken",
"content": "</s>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
},
"legacy": false,
"model_max_length": 1000000000000000019884624838656,
"pad_token": null,
"sp_model_kwargs": {},
"tokenizer_class": "LlamaTokenizer",
"unk_token": {
"__type": "AddedToken",
"content": "<unk>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
}
}