初始化项目,由ModelHub XC社区提供模型

Model: RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-09-27 13:01:17 +08:00
commit 607d7bd912
21 changed files with 288 additions and 0 deletions

54
.gitattributes vendored Normal file
View File

@@ -0,0 +1,54 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q3_K.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.IQ4_NL.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q4_K.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q4_1.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q5_0.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q5_K.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q5_1.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
gemma2-9b-cpt-sea-lionv3-instruct.Q8_0.gguf filter=lfs diff=lfs merge=lfs -text

177
README.md Normal file
View File

@@ -0,0 +1,177 @@
Quantization made by Richard Erkhov.
[Github](https://github.com/RichardErkhov)
[Discord](https://discord.gg/pvy7H8DZMG)
[Request more models](https://github.com/RichardErkhov/quant_request)
gemma2-9b-cpt-sea-lionv3-instruct - GGUF
- Model creator: https://huggingface.co/aisingapore/
- Original model: https://huggingface.co/aisingapore/gemma2-9b-cpt-sea-lionv3-instruct/
| Name | Quant method | Size |
| ---- | ---- | ---- |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q2_K.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q2_K.gguf) | Q2_K | 3.54GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_S.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_S.gguf) | Q3_K_S | 4.04GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q3_K.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q3_K.gguf) | Q3_K | 4.43GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_M.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_M.gguf) | Q3_K_M | 4.43GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_L.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q3_K_L.gguf) | Q3_K_L | 4.78GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.IQ4_XS.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.IQ4_XS.gguf) | IQ4_XS | 4.86GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q4_0.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q4_0.gguf) | Q4_0 | 5.07GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.IQ4_NL.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.IQ4_NL.gguf) | IQ4_NL | 5.1GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q4_K_S.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q4_K_S.gguf) | Q4_K_S | 5.1GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q4_K.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q4_K.gguf) | Q4_K | 5.37GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q4_K_M.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q4_K_M.gguf) | Q4_K_M | 5.37GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q4_1.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q4_1.gguf) | Q4_1 | 5.55GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q5_0.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q5_0.gguf) | Q5_0 | 6.04GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q5_K_S.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q5_K_S.gguf) | Q5_K_S | 6.04GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q5_K.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q5_K.gguf) | Q5_K | 6.19GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q5_K_M.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q5_K_M.gguf) | Q5_K_M | 6.19GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q5_1.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q5_1.gguf) | Q5_1 | 6.52GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q6_K.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q6_K.gguf) | Q6_K | 7.07GB |
| [gemma2-9b-cpt-sea-lionv3-instruct.Q8_0.gguf](https://huggingface.co/RichardErkhov/aisingapore_-_gemma2-9b-cpt-sea-lionv3-instruct-gguf/blob/main/gemma2-9b-cpt-sea-lionv3-instruct.Q8_0.gguf) | Q8_0 | 9.15GB |
Original model description:
---
library_name: transformers
pipeline_tag: text-generation
base_model:
- aisingapore/gemma2-9b-cpt-sea-lionv3-base
language:
- en
- zh
- vi
- id
- th
- fil
- ta
- ms
- km
- lo
- my
- jv
- su
license: gemma
---
# Gemma2 9B CPT SEA-LIONv3 Instruct
SEA-LION is a collection of Large Language Models (LLMs) which have been pretrained and instruct-tuned for the Southeast Asia (SEA) region.
Gemma2 9B CPT SEA-LIONv3 Instruct is a multilingual model which has been fine-tuned with around **500,000 English instruction-completion pairs** alongside a larger pool of around **1,000,000 instruction-completion pairs** from other ASEAN languages, such as Indonesian, Thai and Vietnamese.
SEA-LION stands for _Southeast Asian Languages In One Network_.
- **Developed by:** Products Pillar, AI Singapore
- **Funded by:** Singapore NRF
- **Model type:** Decoder
- **Languages:** English, Chinese, Vietnamese, Indonesian, Thai, Filipino, Tamil, Malay, Khmer, Lao, Burmese, Javanese, Sundanese
- **License:** [Gemma Community License](https://ai.google.dev/gemma/terms)
## Model Details
### Model Description
We performed instruction tuning in English and also in ASEAN languages such as Indonesian, Thai and Vietnamese on our [continued pre-trained Gemma2 9B CPT SEA-LIONv3](https://huggingface.co/aisingapore/gemma2-9b-cpt-sea-lionv3-base), a decoder model using the Gemma2 architecture, to create Gemma2 9B CPT SEA-LIONv3 Instruct.
For tokenisation, the model employs the default tokenizer used in Gemma-2-9B. The model has a context length of 8192.
### Benchmark Performance
We evaluated Gemma2 9B CPT SEA-LIONv3 Instruct on both general language capabilities and instruction-following capabilities.
#### General Language Capabilities
For the evaluation of general language capabilities, we employed the [SEA HELM (also known as BHASA) evaluation benchmark](https://arxiv.org/abs/2309.06085v2) across a variety of tasks.
These tasks include Question Answering (QA), Sentiment Analysis (Sentiment), Toxicity Detection (Toxicity), Translation in both directions (Eng>Lang & Lang>Eng), Abstractive Summarization (Summ), Causal Reasoning (Causal) and Natural Language Inference (NLI).
Note: SEA HELM is implemented using prompts to elicit answers in a strict format. For all tasks, the model is expected to provide an answer tag from which the answer is automatically extracted. For tasks where options are provided, the answer should comprise one of the pre-defined options. The scores for each task is normalised to account for baseline performance due to random chance.
The evaluation was done **zero-shot** with native prompts on a sample of 100-1000 instances for each dataset.
#### Instruction-following Capabilities
Since Gemma2 9B CPT SEA-LIONv3 Instruct is an instruction-following model, we also evaluated it on instruction-following capabilities with two datasets, [IFEval](https://arxiv.org/abs/2311.07911) and [MT-Bench](https://arxiv.org/abs/2306.05685).
As these two datasets were originally in English, the linguists and native speakers in the team worked together to filter, localize and translate the datasets into the respective target languages to ensure that the examples remained reasonable, meaningful and natural.
**IFEval**
IFEval evaluates a model's ability to adhere to constraints provided in the prompt, for example beginning a response with a specific word/phrase or answering with a certain number of sections. Additionally, accuracy is normalized by the proportion of responses in the correct language (if the model performs the task correctly but responds in the wrong language, it is judged to have failed the task).
**MT-Bench**
MT-Bench evaluates a model's ability to engage in multi-turn (2 turns) conversations and respond in ways that align with human needs. We use `gpt-4-1106-preview` as the judge model and compare against `gpt-3.5-turbo-0125` as the baseline model. The metric used is the weighted win rate against the baseline model (i.e. average win rate across each category: Math, Reasoning, STEM, Humanities, Roleplay, Writing, Extraction). A tie is given a score of 0.5.
For more details on Gemma2 9B CPT SEA-LIONv3 Instruct benchmark performance, please refer to the SEA HELM leaderboard, https://leaderboard.sea-lion.ai/
### Usage
Gemma2 9B CPT SEA-LIONv3 Instruct can be run using the 🤗 Transformers library
```python
# Please use transformers==4.45.2
import transformers
import torch
model_id = "aisingapore/gemma2-9b-cpt-sea-lionv3-instruct"
pipeline = transformers.pipeline(
"text-generation",
model=model_id,
model_kwargs={"torch_dtype": torch.bfloat16},
device_map="auto",
)
messages = [
{"role": "user", "content": "Apa sentimen dari kalimat berikut ini?\nKalimat: Buku ini sangat membosankan.\nJawaban: "},
]
outputs = pipeline(
messages,
max_new_tokens=256,
)
print(outputs[0]["generated_text"][-1])
```
### Caveats
It is important for users to be aware that our model exhibits certain limitations that warrant consideration. Like many LLMs, the model can hallucinate and occasionally generates irrelevant content, introducing fictional elements that are not grounded in the provided context. Users should also exercise caution in interpreting and validating the model's responses due to the potential inconsistencies in its reasoning.
## Limitations
### Safety
Current SEA-LION models, including this commercially permissive release, have not been aligned for safety. Developers and users should perform their own safety fine-tuning and related security measures. In no event shall the authors be held liable for any claim, damages, or other liability arising from the use of the released weights and codes.
## Technical Specifications
### Fine-Tuning Details
Gemma2 9B CPT SEA-LIONv3 Instruct was built using a combination of a full parameter fine-tune, on-policy alignment, and model merges of the best performing checkpoints. The training process for fine-tuning was approximately 15 hours, with alignment taking 2 hours, both on 8x H100-80GB GPUs.
## Data
Gemma2 9B CPT SEA-LIONv3 Instruct was trained on a wide range of synthetic instructions, alongside publicly available instructions hand-curated by the team with the assistance of native speakers. In addition, special care was taken to ensure that the datasets used had commercially permissive licenses through verification with the original data source.
## Call for Contributions
We encourage researchers, developers, and language enthusiasts to actively contribute to the enhancement and expansion of SEA-LION. Contributions can involve identifying and reporting bugs, sharing pre-training, instruction, and preference data, improving documentation usability, proposing and implementing new model evaluation tasks and metrics, or training versions of the model in additional Southeast Asian languages. Join us in shaping the future of SEA-LION by sharing your expertise and insights to make these models more accessible, accurate, and versatile. Please check out our GitHub for further information on the call for contributions.
## The Team
Chan Adwin, Choa Esther, Cheng Nicholas, Huang Yuli, Lau Wayne, Lee Chwan Ren, Leong Wai Yi, Leong Wei Qi, Limkonchotiwat Peerat, Liu Bing Jie Darius, Montalan Jann Railey, Ng Boon Cheong Raymond, Ngui Jian Gang, Nguyen Thanh Ngan, Ong Brandon, Ong Tat-Wee David, Ong Zhi Hao, Rengarajan Hamsawardhini, Siow Bryan, Susanto Yosephine, Tai Ngee Chia, Tan Choon Meng, Teo Eng Sipp Leslie, Teo Wei Yi, Tjhi William, Teng Walter, Yeo Yeow Tong, Yong Xianbin
## Acknowledgements
[AI Singapore](​​https://aisingapore.org/) is a national programme supported by the National Research Foundation, Singapore and hosted by the National University of Singapore. Any opinions, findings and conclusions or recommendations expressed in this material are those of the author(s) and do not reflect the views of the National Research Foundation or the National University of Singapore.
## Contact
For more info, please contact us using this [SEA-LION Inquiry Form](https://forms.gle/sLCUVb95wmGf43hi6)
[Link to SEA-LION's GitHub repository](https://github.com/aisingapore/sealion)
## Disclaimer
This is the repository for the commercial instruction-tuned model.
The model has _not_ been aligned for safety.
Developers and users should perform their own safety fine-tuning and related security measures.
In no event shall the authors be held liable for any claims, damages, or other liabilities arising from the use of the released weights and codes.

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e4c27e79ff89b8c59e75fec4c8dc611e029939cd6e89672a8b38b66f441e4383
size 5475255712

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1b796eb1fc812dccad81eb3930b1bd0499441c70654cc9bd40f3fb037432a88b
size 5223171488

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a1d809bf344c2940263bf030fc80b8fe0152f20274c3555a9f8bed968c0de34a
size 3805398432

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b8c614eb1b833559b04cbc887a0282bf6a5679d63fdd86343a9406cbc7b55640
size 4761781664

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a1e1d4c71f1bff1cc1a02a09e18701673dcb8b80ad8dbcc3d4f2c10aa40c2d42
size 5132453280

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b8c614eb1b833559b04cbc887a0282bf6a5679d63fdd86343a9406cbc7b55640
size 4761781664

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f7d5b4efa060394c1bf4e02f18f9de748ff75e66c6ec28a110e80dbf853569bf
size 4337665440

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d536bd1ea5cd2d37e41a5585993666d4f4387b3066e5fcaafeb4db0b25634d72
size 5443143072

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3502ae46bc4ef5e9ebc912da1cd2ed04ce011b0bade4adee223ad74bfc1c48a8
size 5963367840

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:384988efff70d4158bdd89e7fb58ab72b550b245c5170b315f8c568586acfde3
size 5761058208

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:384988efff70d4158bdd89e7fb58ab72b550b245c5170b315f8c568586acfde3
size 5761058208

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:035ff2c19d6edf33eb4cb8223832e0d1b83f72eeb047f74eafe88477a51b7e1a
size 5478925728

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:06673cbd397835d3afea5e170ea67bcdd8469d8c76f632221fed5a10f9d72c09
size 6483592608

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:65838a70183646a910404decb31a7fa0b191c379e8ef0cb19945b44d601e5772
size 7003817376

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a57456468a8680b3a8c2180e994f60bf86dfc91529fa071af08e57eb44fe9a6c
size 6647367072

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a57456468a8680b3a8c2180e994f60bf86dfc91529fa071af08e57eb44fe9a6c
size 6647367072

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c0254024bf24a89f3600021772e2999b987fffb6c490528e4792201a51c54f99
size 6483592608

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e428591f398399bc229d7669162677f1d532fd9a7d3f22c145814616b45ae01d
size 7589070240

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9b01a4276d5c54da064f86c51a44cfc5e53593b49733581ac26a75507eec658b
size 9827149216