初始化项目,由ModelHub XC社区提供模型

Model: bofenghuang/vigogne-2-7b-chat
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-04 13:54:32 +08:00
commit 1fca3b08f3
17 changed files with 646 additions and 0 deletions

35
.gitattributes vendored Normal file
View File

@@ -0,0 +1,35 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text

3
.gitignore vendored Normal file
View File

@@ -0,0 +1,3 @@
checkpoint-*/
tmp*

192
README.md Normal file
View File

@@ -0,0 +1,192 @@
---
license: llama2
language: fr
pipeline_tag: text-generation
inference: false
tags:
- LLM
- llama-2
- finetuned
---
<p align="center" width="100%">
<img src="https://huggingface.co/bofenghuang/vigogne-2-7b-chat/resolve/v2.0/logo_v2.jpg" alt="Vigogne" style="width: 30%; min-width: 300px; display: block; margin: auto;">
</p>
# Vigogne-2-7B-Chat-V2.0: A Llama-2-based French Chat LLM
Vigogne-2-7B-Chat-V2.0 is a French chat LLM, based on [LLaMA-2-7B](https://ai.meta.com/llama), optimized to generate helpful and coherent responses in conversations with users.
Check out our [release blog](https://github.com/bofenghuang/vigogne/blob/main/blogs/2023-08-17-vigogne-chat-v2_0.md) and [GitHub repository](https://github.com/bofenghuang/vigogne) for more information.
**Usage and License Notices**: Vigogne-2-7B-Chat-V2.0 follows Llama-2's [usage policy](https://ai.meta.com/llama/use-policy). A significant portion of the training data is distilled from GPT-3.5-Turbo and GPT-4, kindly use it cautiously to avoid any violations of OpenAI's [terms of use](https://openai.com/policies/terms-of-use).
## Changelog
All previous versions are accessible through branches.
- **V1.0**: Trained on 420K chat data.
- **V2.0**: Trained on 520K data. Check out our [release blog](https://github.com/bofenghuang/vigogne/blob/main/blogs/2023-08-17-vigogne-chat-v2_0.md) for more details.
## Prompt Template
We utilized prefix tokens `<user>:` and `<assistant>:` to distinguish between user and assistant utterances.
You can apply this formatting using the [chat template](https://huggingface.co/docs/transformers/main/chat_templating) through the `apply_chat_template()` method.
```python
from transformers import AutoTokenizer
tokenizer = AutoTokenizer.from_pretrained("bofenghuang/vigogne-2-7b-chat")
conversation = [
{"role": "user", "content": "Bonjour ! Comment ça va aujourd'hui ?"},
{"role": "assistant", "content": "Bonjour ! Je suis une IA, donc je n'ai pas de sentiments, mais je suis prêt à vous aider. Comment puis-je vous assister aujourd'hui ?"},
{"role": "user", "content": "Quelle est la hauteur de la Tour Eiffel ?"},
{"role": "assistant", "content": "La Tour Eiffel mesure environ 330 mètres de hauteur."},
{"role": "user", "content": "Comment monter en haut ?"},
]
print(tokenizer.apply_chat_template(conversation, tokenize=False, add_generation_prompt=True))
```
You will get
```
<s><|system|>: Vous êtes Vigogne, un assistant IA créé par Zaion Lab. Vous suivez extrêmement bien les instructions. Aidez autant que vous le pouvez.
<|user|>: Bonjour ! Comment ça va aujourd'hui ?
<|assistant|>: Bonjour ! Je suis une IA, donc je n'ai pas de sentiments, mais je suis prêt à vous aider. Comment puis-je vous assister aujourd'hui ?</s>
<|user|>: Quelle est la hauteur de la Tour Eiffel ?
<|assistant|>: La Tour Eiffel mesure environ 330 mètres de hauteur.</s>
<|user|>: Comment monter en haut ?
<|assistant|>:
```
## Usage
### Inference using the quantized versions
The quantized versions of this model are generously provided by [TheBloke](https://huggingface.co/TheBloke)!
- AWQ for GPU inference: [TheBloke/Vigogne-2-7B-Chat-AWQ](https://huggingface.co/TheBloke/Vigogne-2-7B-Chat-AWQ)
- GTPQ for GPU inference: [TheBloke/Vigogne-2-7B-Chat-GPTQ](https://huggingface.co/TheBloke/Vigogne-2-7B-Chat-GPTQ)
- GGUF for CPU+GPU inference: [TheBloke/Vigogne-2-7B-Chat-GGUF](https://huggingface.co/TheBloke/Vigogne-2-7B-Chat-GGUF)
These versions facilitate testing and development with various popular frameworks, including [AutoAWQ](https://github.com/casper-hansen/AutoAWQ), [vLLM](https://github.com/vllm-project/vllm), [AutoGPTQ](https://github.com/PanQiWei/AutoGPTQ), [GPTQ-for-LLaMa](https://github.com/qwopqwop200/GPTQ-for-LLaMa), [llama.cpp](https://github.com/ggerganov/llama.cpp), [text-generation-webui](https://github.com/oobabooga/text-generation-webui), and more.
### Inference using the unquantized model with 🤗 Transformers
```python
from typing import Dict, List, Optional
import torch
from transformers import AutoModelForCausalLM, AutoTokenizer, GenerationConfig, TextStreamer
model_name_or_path = "bofenghuang/vigogne-2-7b-chat"
revision = "v2.0"
tokenizer = AutoTokenizer.from_pretrained(model_name_or_path, revision=revision, padding_side="right", use_fast=False)
model = AutoModelForCausalLM.from_pretrained(model_name_or_path, revision=revision, torch_dtype=torch.float16, device_map="auto")
streamer = TextStreamer(tokenizer, timeout=10.0, skip_prompt=True, skip_special_tokens=True)
def chat(
query: str,
history: Optional[List[Dict]] = None,
temperature: float = 0.7,
top_p: float = 1.0,
top_k: float = 0,
repetition_penalty: float = 1.1,
max_new_tokens: int = 1024,
**kwargs,
):
if history is None:
history = []
history.append({"role": "user", "content": query})
input_ids = tokenizer.apply_chat_template(history, add_generation_prompt=True, return_tensors="pt").to(model.device)
input_length = input_ids.shape[1]
generated_outputs = model.generate(
input_ids=input_ids,
generation_config=GenerationConfig(
temperature=temperature,
do_sample=temperature > 0.0,
top_p=top_p,
top_k=top_k,
repetition_penalty=repetition_penalty,
max_new_tokens=max_new_tokens,
pad_token_id=tokenizer.eos_token_id,
**kwargs,
),
streamer=streamer,
return_dict_in_generate=True,
)
generated_tokens = generated_outputs.sequences[0, input_length:]
generated_text = tokenizer.decode(generated_tokens, skip_special_tokens=True)
history.append({"role": "assistant", "content": generated_text})
return generated_text, history
# 1st round
response, history = chat("Un escargot parcourt 100 mètres en 5 heures. Quelle est sa vitesse ?", history=None)
# 2nd round
response, history = chat("Quand il peut dépasser le lapin ?", history=history)
# 3rd round
response, history = chat("Écris une histoire imaginative qui met en scène une compétition de course entre un escargot et un lapin.", history=history)
```
You can also use the Google Colab Notebook provided below.
<a href="https://colab.research.google.com/github/bofenghuang/vigogne/blob/main/notebooks/infer_chat.ipynb" target="_blank"><img src="https://colab.research.google.com/assets/colab-badge.svg" alt="Open In Colab"/></a>
### Inference using the unquantized model with vLLM
Set up an OpenAI-compatible server with the following command:
```bash
# Install vLLM
# This may take 5-10 minutes.
# pip install vllm
# Start server for Vigogne-Chat models
python -m vllm.entrypoints.openai.api_server --model bofenghuang/vigogne-2-7b-chat
# List models
# curl http://localhost:8000/v1/models
```
Query the model using the openai python package.
```python
import openai
# Modify OpenAI's API key and API base to use vLLM's API server.
openai.api_key = "EMPTY"
openai.api_base = "http://localhost:8000/v1"
# First model
models = openai.Model.list()
model = models["data"][0]["id"]
# Chat completion API
chat_completion = openai.ChatCompletion.create(
model=model,
messages=[
{"role": "user", "content": "Parle-moi de toi-même."},
],
max_tokens=1024,
temperature=0.7,
)
print("Chat completion results:", chat_completion)
```
## Limitations
Vigogne is still under development, and there are many limitations that have to be addressed. Please note that it is possible that the model generates harmful or biased content, incorrect information or generally unhelpful answers.

25
config.json Normal file
View File

@@ -0,0 +1,25 @@
{
"_name_or_path": "outputs/chat/llama-2-7b-ft-chat-max2048-packing2048-ep3-bs128-lr1e4-gpt3_5data-merged",
"architectures": [
"LlamaForCausalLM"
],
"bos_token_id": 1,
"eos_token_id": 2,
"hidden_act": "silu",
"hidden_size": 4096,
"initializer_range": 0.02,
"intermediate_size": 11008,
"max_position_embeddings": 4096,
"model_type": "llama",
"num_attention_heads": 32,
"num_hidden_layers": 32,
"num_key_value_heads": 32,
"pretraining_tp": 1,
"rms_norm_eps": 1e-05,
"rope_scaling": null,
"tie_word_embeddings": false,
"torch_dtype": "float16",
"transformers_version": "4.32.0.dev0",
"use_cache": true,
"vocab_size": 32000
}

10
generation_config.json Normal file
View File

@@ -0,0 +1,10 @@
{
"bos_token_id": 1,
"do_sample": true,
"eos_token_id": 2,
"max_length": 4096,
"pad_token_id": 0,
"temperature": 0.6,
"top_p": 0.9,
"transformers_version": "4.32.0.dev0"
}

BIN
logo_v2.jpg Normal file

Binary file not shown.

After

Width:  |  Height:  |  Size: 190 KiB

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:598f272292e5fd95a1b8c55661f75e8cd723e7c03261e3a3c87d419b2d02d2db
size 1981887519

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0c601f9ad5961862c6d33b2ce20a87fa4faace97a0bf15f15c79c25ef8d83759
size 1990293863

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:bd5c437796e4ff0d52f09d6285284c05a39b11771c4ab20d323ce2ca0d2f14be
size 1990293863

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:63a461d1338d87ff5b300b99a94be2387eba61bdf0583b4b1dc7a5bc98193f4b
size 1990293863

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e586d815091d3fe896b493e5c6205216a58c5c0eb404bc6736e0093f04390cff
size 1933653763

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c96192318a0c5cf3fd5939b121cb84d3377bb38b6f5a1650853e544b101e84a5
size 1933670759

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b6bd313fae56049023afba7dcdc334fe3fb73fbfabe35bf9647d2e602e98fbf1
size 1656834785

View File

@@ -0,0 +1,298 @@
{
"metadata": {
"total_size": 13476831232
},
"weight_map": {
"lm_head.weight": "pytorch_model-00007-of-00007.bin",
"model.embed_tokens.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.input_layernorm.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.mlp.down_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.mlp.gate_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.mlp.up_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.post_attention_layernorm.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.self_attn.k_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.self_attn.o_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.self_attn.q_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.0.self_attn.v_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.input_layernorm.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.mlp.down_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.mlp.gate_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.mlp.up_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.post_attention_layernorm.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.self_attn.k_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.self_attn.o_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.self_attn.q_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.1.self_attn.v_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.10.input_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.10.mlp.down_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.10.mlp.gate_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.10.mlp.up_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.10.post_attention_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.10.self_attn.k_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.10.self_attn.o_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.10.self_attn.q_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.10.self_attn.v_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.input_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.mlp.down_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.mlp.gate_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.mlp.up_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.post_attention_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.self_attn.k_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.self_attn.o_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.self_attn.q_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.11.self_attn.v_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.input_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.mlp.down_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.mlp.gate_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.mlp.up_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.post_attention_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.self_attn.k_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.self_attn.o_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.self_attn.q_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.12.self_attn.v_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.input_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.mlp.down_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.mlp.gate_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.mlp.up_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.post_attention_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.self_attn.k_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.self_attn.o_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.self_attn.q_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.13.self_attn.v_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.14.input_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.14.mlp.down_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.14.mlp.gate_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.14.mlp.up_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.14.post_attention_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.14.self_attn.k_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.14.self_attn.o_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.14.self_attn.q_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.14.self_attn.v_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.input_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.mlp.down_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.mlp.gate_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.mlp.up_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.post_attention_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.self_attn.k_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.self_attn.o_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.self_attn.q_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.15.self_attn.v_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.input_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.mlp.down_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.mlp.gate_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.mlp.up_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.post_attention_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.self_attn.k_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.self_attn.o_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.self_attn.q_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.16.self_attn.v_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.input_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.mlp.down_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.mlp.gate_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.mlp.up_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.post_attention_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.self_attn.k_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.self_attn.o_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.self_attn.q_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.17.self_attn.v_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.input_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.mlp.down_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.mlp.gate_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.mlp.up_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.post_attention_layernorm.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.self_attn.k_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.self_attn.o_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.self_attn.q_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.18.self_attn.v_proj.weight": "pytorch_model-00004-of-00007.bin",
"model.layers.19.input_layernorm.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.19.mlp.down_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.19.mlp.gate_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.19.mlp.up_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.19.post_attention_layernorm.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.19.self_attn.k_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.19.self_attn.o_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.19.self_attn.q_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.19.self_attn.v_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.2.input_layernorm.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.2.mlp.down_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.2.mlp.gate_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.2.mlp.up_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.2.post_attention_layernorm.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.2.self_attn.k_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.2.self_attn.o_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.2.self_attn.q_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.2.self_attn.v_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.20.input_layernorm.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.20.mlp.down_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.20.mlp.gate_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.20.mlp.up_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.20.post_attention_layernorm.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.20.self_attn.k_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.20.self_attn.o_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.20.self_attn.q_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.20.self_attn.v_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.input_layernorm.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.mlp.down_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.mlp.gate_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.mlp.up_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.post_attention_layernorm.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.self_attn.k_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.self_attn.o_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.self_attn.q_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.21.self_attn.v_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.input_layernorm.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.mlp.down_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.mlp.gate_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.mlp.up_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.post_attention_layernorm.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.self_attn.k_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.self_attn.o_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.self_attn.q_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.22.self_attn.v_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.23.input_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.23.mlp.down_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.23.mlp.gate_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.23.mlp.up_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.23.post_attention_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.23.self_attn.k_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.23.self_attn.o_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.23.self_attn.q_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.23.self_attn.v_proj.weight": "pytorch_model-00005-of-00007.bin",
"model.layers.24.input_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.24.mlp.down_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.24.mlp.gate_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.24.mlp.up_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.24.post_attention_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.24.self_attn.k_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.24.self_attn.o_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.24.self_attn.q_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.24.self_attn.v_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.input_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.mlp.down_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.mlp.gate_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.mlp.up_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.post_attention_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.self_attn.k_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.self_attn.o_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.self_attn.q_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.25.self_attn.v_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.input_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.mlp.down_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.mlp.gate_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.mlp.up_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.post_attention_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.self_attn.k_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.self_attn.o_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.self_attn.q_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.26.self_attn.v_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.input_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.mlp.down_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.mlp.gate_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.mlp.up_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.post_attention_layernorm.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.self_attn.k_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.self_attn.o_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.self_attn.q_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.27.self_attn.v_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.28.input_layernorm.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.28.mlp.down_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.28.mlp.gate_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.28.mlp.up_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.28.post_attention_layernorm.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.28.self_attn.k_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.28.self_attn.o_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.28.self_attn.q_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.28.self_attn.v_proj.weight": "pytorch_model-00006-of-00007.bin",
"model.layers.29.input_layernorm.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.29.mlp.down_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.29.mlp.gate_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.29.mlp.up_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.29.post_attention_layernorm.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.29.self_attn.k_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.29.self_attn.o_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.29.self_attn.q_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.29.self_attn.v_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.3.input_layernorm.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.3.mlp.down_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.3.mlp.gate_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.3.mlp.up_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.3.post_attention_layernorm.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.3.self_attn.k_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.3.self_attn.o_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.3.self_attn.q_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.3.self_attn.v_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.30.input_layernorm.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.30.mlp.down_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.30.mlp.gate_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.30.mlp.up_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.30.post_attention_layernorm.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.30.self_attn.k_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.30.self_attn.o_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.30.self_attn.q_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.30.self_attn.v_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.input_layernorm.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.mlp.down_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.mlp.gate_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.mlp.up_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.post_attention_layernorm.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.self_attn.k_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.self_attn.o_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.self_attn.q_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.31.self_attn.v_proj.weight": "pytorch_model-00007-of-00007.bin",
"model.layers.4.input_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.4.mlp.down_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.4.mlp.gate_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.4.mlp.up_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.4.post_attention_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.4.self_attn.k_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.4.self_attn.o_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.4.self_attn.q_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.4.self_attn.v_proj.weight": "pytorch_model-00001-of-00007.bin",
"model.layers.5.input_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.5.mlp.down_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.5.mlp.gate_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.5.mlp.up_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.5.post_attention_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.5.self_attn.k_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.5.self_attn.o_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.5.self_attn.q_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.5.self_attn.v_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.input_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.mlp.down_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.mlp.gate_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.mlp.up_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.post_attention_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.self_attn.k_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.self_attn.o_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.self_attn.q_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.6.self_attn.v_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.input_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.mlp.down_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.mlp.gate_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.mlp.up_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.post_attention_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.self_attn.k_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.self_attn.o_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.self_attn.q_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.7.self_attn.v_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.input_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.mlp.down_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.mlp.gate_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.mlp.up_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.post_attention_layernorm.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.self_attn.k_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.self_attn.o_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.self_attn.q_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.8.self_attn.v_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.9.input_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.9.mlp.down_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.9.mlp.gate_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.9.mlp.up_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.9.post_attention_layernorm.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.9.self_attn.k_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.9.self_attn.o_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.layers.9.self_attn.q_proj.weight": "pytorch_model-00002-of-00007.bin",
"model.layers.9.self_attn.v_proj.weight": "pytorch_model-00003-of-00007.bin",
"model.norm.weight": "pytorch_model-00007-of-00007.bin"
}
}

23
special_tokens_map.json Normal file
View File

@@ -0,0 +1,23 @@
{
"bos_token": {
"content": "<s>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
},
"eos_token": {
"content": "</s>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
},
"unk_token": {
"content": "<unk>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
}
}

BIN
tokenizer.model (Stored with Git LFS) Normal file

Binary file not shown.

36
tokenizer_config.json Normal file
View File

@@ -0,0 +1,36 @@
{
"add_bos_token": true,
"add_eos_token": false,
"bos_token": {
"__type": "AddedToken",
"content": "<s>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
},
"clean_up_tokenization_spaces": false,
"eos_token": {
"__type": "AddedToken",
"content": "</s>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
},
"legacy": false,
"chat_template": "{{ bos_token }}{% if messages[0]['role'] == 'system' %}{% set loop_messages = messages[1:] %}{% set system_message = messages[0]['content'] %}{% elif true == true %}{% set loop_messages = messages %}{% set system_message = 'Vous êtes Vigogne, un assistant IA créé par Zaion Lab. Vous suivez extrêmement bien les instructions. Aidez autant que vous le pouvez.' %}{% else %}{% set loop_messages = messages %}{% set system_message = false %}{% endif %}{% if system_message != false %}{{ '<|system|>: ' + system_message + '\\n' }}{% endif %}{% for message in loop_messages %}{% if (message['role'] == 'user') != (loop.index0 % 2 == 0) %}{{ raise_exception('Conversation roles must alternate user/assistant/user/assistant/...') }}{% endif %}{% if message['role'] == 'user' %}{{ '<|user|>: ' + message['content'].strip() + '\\n' }}{% elif message['role'] == 'assistant' %}{{ '<|assistant|>: ' + message['content'].strip() + eos_token + '\\n' }}{% endif %}{% endfor %}{% if add_generation_prompt %}{{ '<|assistant|>:' }}{% endif %}",
"model_max_length": 1000000000000000019884624838656,
"pad_token": null,
"padding_side": "right",
"sp_model_kwargs": {},
"tokenizer_class": "LlamaTokenizer",
"unk_token": {
"__type": "AddedToken",
"content": "<unk>",
"lstrip": false,
"normalized": false,
"rstrip": false,
"single_word": false
}
}