初始化项目,由ModelHub XC社区提供模型
Model: Lewdiculous/Lumimaid-v0.2-8B-GGUF-IQ-Imatrix Source: Original Platform
This commit is contained in:
47
.gitattributes
vendored
Normal file
47
.gitattributes
vendored
Normal file
@@ -0,0 +1,47 @@
|
||||
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||
*.model filter=lfs diff=lfs merge=lfs -text
|
||||
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||
imatrix.dat filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-BF16.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-F16.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-IQ3_M-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-IQ3_XXS-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-IQ4_XS-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-Q4_K_M-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-Q4_K_S-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-Q5_K_M-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-Q5_K_S-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-Q6_K-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
Lumimaid-v0.2-8B-Q8_0-imat.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
3
Lumimaid-v0.2-8B-BF16.gguf
Normal file
3
Lumimaid-v0.2-8B-BF16.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:d770bce46a00a3d5b1bd86e477ebebd37d562b72e5e02c46ed3dce1f702803f6
|
||||
size 16068891552
|
||||
3
Lumimaid-v0.2-8B-F16.gguf
Normal file
3
Lumimaid-v0.2-8B-F16.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:42d03de9a0b2e392c3edfbe71d21ab94bff94fa5683543afff4c65fea5c84256
|
||||
size 16068891552
|
||||
3
Lumimaid-v0.2-8B-IQ3_M-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-IQ3_M-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:e331e55ca24095d1ce21419521cdd3d9917409b98bcd47791c5cab3fa6e9c48b
|
||||
size 3784824000
|
||||
3
Lumimaid-v0.2-8B-IQ3_XXS-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-IQ3_XXS-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:761b00ed0352354b1784aa3d554e09ca5faeedb90177ff9c9047eac50362df03
|
||||
size 3274912960
|
||||
3
Lumimaid-v0.2-8B-IQ4_XS-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-IQ4_XS-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:261762e2d9e9edc2b5e5053cb48cdbf52748801f139a95c5f6ccc95092d89e6e
|
||||
size 4447663296
|
||||
3
Lumimaid-v0.2-8B-Q4_K_M-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-Q4_K_M-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:f7b9e1db31efcb91b6c8c338a7ccec8d7bdcbc78491cbd6a82d5068eb2f6ec40
|
||||
size 4920734912
|
||||
3
Lumimaid-v0.2-8B-Q4_K_S-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-Q4_K_S-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:71331ba953b9e6c428bbc7395a93ac449cea86170d10049cc653ad4092f255e2
|
||||
size 4692669632
|
||||
3
Lumimaid-v0.2-8B-Q5_K_M-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-Q5_K_M-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:b3cb02b76f149df78bc65eac4f59f16574cc799f35d358d557d2f1dca43c31d3
|
||||
size 5732988096
|
||||
3
Lumimaid-v0.2-8B-Q5_K_S-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-Q5_K_S-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:ef227f43d975f307a893c59c97ee86718cb755f53a26d5da8c6ee3afd858cb22
|
||||
size 5599294656
|
||||
3
Lumimaid-v0.2-8B-Q6_K-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-Q6_K-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:a9761044727e2a9a65065d2b04658a253b6e2d4d4a559a37f22c170c42b96dae
|
||||
size 6596007104
|
||||
3
Lumimaid-v0.2-8B-Q8_0-imat.gguf
Normal file
3
Lumimaid-v0.2-8B-Q8_0-imat.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:7532e1accff0594791994150ff11c7fd5695bb2da1c7bba44394bf20a875c572
|
||||
size 8540771520
|
||||
154
README.md
Normal file
154
README.md
Normal file
@@ -0,0 +1,154 @@
|
||||
---
|
||||
base_model: NeverSleep/Lumimaid-v0.2-8B
|
||||
quantized_by: Lewdiculous
|
||||
library_name: transformers
|
||||
license: cc-by-nc-4.0
|
||||
inference: false
|
||||
language:
|
||||
- en
|
||||
tags:
|
||||
- roleplay
|
||||
- llama3
|
||||
- sillytavern
|
||||
---
|
||||
|
||||
# #roleplay #sillytavern #llama3
|
||||
|
||||
My GGUF-IQ-Imatrix quants for [**NeverSleep/Lumimaid-v0.2-8B**](https://huggingface.co/NeverSleep/Lumimaid-v0.2-8B).
|
||||
|
||||
I recommend checking their page for feedback and support.
|
||||
|
||||
> [!IMPORTANT]
|
||||
> **Quantization process:** <br>
|
||||
> Imatrix data was generated from the FP16-GGUF and conversions directly from the BF16-GGUF. <br>
|
||||
> This is a bit more disk and compute intensive but hopefully avoids any losses during conversion. <br>
|
||||
> To run this model, please use the [**latest version of KoboldCpp**](https://github.com/LostRuins/koboldcpp/releases/latest). <br>
|
||||
> If you noticed any issues let me know in the discussions.
|
||||
|
||||
> [!NOTE]
|
||||
> **Presets:** <br>
|
||||
> * Llama-3. <br>
|
||||
>
|
||||
> Some compatible SillyTavern presets can be found [**here (Virt's Roleplay Presets - v1.9)**](https://huggingface.co/Virt-io/SillyTavern-Presets). <br>
|
||||
> Check [**discussions such as this one**](https://huggingface.co/Virt-io/SillyTavern-Presets/discussions/5#664d6fb87c563d4d95151baa) and [**this one**](https://www.reddit.com/r/SillyTavernAI/comments/1dff2tl/my_personal_llama3_stheno_presets/) for other presets and samplers recommendations. <br>
|
||||
> Lower temperatures are recommended by the authors, so make sure to experiment. <br>
|
||||
>
|
||||
> **General usage with KoboldCpp:** <br>
|
||||
> For **8GB VRAM** GPUs, I recommend the **Q4_K_M-imat** (4.89 BPW) quant for up to 12288 context sizes without the use of `--quantkv`. <br>
|
||||
> Using `--quantkv 1` (≈Q8) or even `--quantkv 2` (≈Q4) can get you to 32K context sizes with the caveat of not being compatible with Context Shifting, only relevant if you can manage to fill up that much context. <br>
|
||||
> [**Read more about it in the release here**](https://github.com/LostRuins/koboldcpp/releases/tag/v1.67).
|
||||
|
||||
|
||||
<details>
|
||||
<summary>⇲ Click here to expand/hide information – General chart with relative quant parformances.</summary>
|
||||
|
||||
> [!NOTE]
|
||||
> **Recommended read:** <br>
|
||||
>
|
||||
> [**"Which GGUF is right for me? (Opinionated)" by Artefact2**](https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9)
|
||||
>
|
||||
> *Click the image to view full size.*
|
||||
> 
|
||||
|
||||
</details>
|
||||
|
||||
> [!TIP]
|
||||
> **Personal-support:** <br>
|
||||
> I apologize for disrupting your experience. <br>
|
||||
> Eventually I may be able to use a dedicated server for this, but for now hopefully these quants are helpful. <br>
|
||||
> If you **want** and you are **able to**... <br>
|
||||
> You can [**spare some change over here (Ko-fi)**](https://ko-fi.com/Lewdiculous). <br>
|
||||
>
|
||||
> **Author-support:** <br>
|
||||
> You can support the authors [**at their pages**](https://ko-fi.com/undiai)/[**here**](https://ikaridevgit.github.io/).
|
||||
|
||||

|
||||
|
||||
<details>
|
||||
<summary>Original model card information.</summary>
|
||||
|
||||
## **Original card:**
|
||||
|
||||
## Lumimaid 0.2
|
||||
<img src="https://cdn-uploads.huggingface.co/production/uploads/63ab1241ad514ca8d1430003/TUcHg7LKNjfo0sni88Ps7.png" alt="Image" style="display: block; margin-left: auto; margin-right: auto; width: 65%;">
|
||||
<div style="text-align: center; font-size: 30px;">
|
||||
<a href="https://huggingface.co/NeverSleep/Lumimaid-v0.2-8B">[8b]</a> -
|
||||
<a href="https://huggingface.co/NeverSleep/Lumimaid-v0.2-12B">12b</a> -
|
||||
<a href="https://huggingface.co/NeverSleep/Lumimaid-v0.2-70B">70b</a> -
|
||||
<a href="https://huggingface.co/NeverSleep/Lumimaid-v0.2-123B">123b</a>
|
||||
</div>
|
||||
|
||||
### This model is based on: [Meta-Llama-3.1-8B-Instruct](https://huggingface.co/meta-llama/Meta-Llama-3.1-8B-Instruct)
|
||||
Wandb: https://wandb.ai/undis95/Lumi-Llama-3-1-8B?nw=nwuserundis95
|
||||
|
||||
Lumimaid 0.1 -> 0.2 is a HUGE step up dataset wise.
|
||||
|
||||
As some people have told us our models are sloppy, Ikari decided to say fuck it and literally nuke all chats out with most slop.
|
||||
|
||||
Our dataset stayed the same since day one, we added data over time, cleaned them, and repeat. After not releasing model for a while because we were never satisfied, we think it's time to come back!
|
||||
|
||||
|
||||
## Prompt template: Llama-3-Instruct
|
||||
|
||||
```
|
||||
<|begin_of_text|><|start_header_id|>system<|end_header_id|>
|
||||
|
||||
{system_prompt}<|eot_id|><|start_header_id|>user<|end_header_id|>
|
||||
|
||||
{input}<|eot_id|><|start_header_id|>assistant<|end_header_id|>
|
||||
|
||||
{output}<|eot_id|>
|
||||
```
|
||||
|
||||
## Credits:
|
||||
- Undi
|
||||
- IkariDev
|
||||
|
||||
## Training data we used to make our dataset:
|
||||
|
||||
- [Epiculous/Gnosis](https://huggingface.co/Epiculous/Gnosis)
|
||||
- [ChaoticNeutrals/Luminous_Opus](https://huggingface.co/datasets/ChaoticNeutrals/Luminous_Opus)
|
||||
- [ChaoticNeutrals/Synthetic-Dark-RP](https://huggingface.co/datasets/ChaoticNeutrals/Synthetic-Dark-RP)
|
||||
- [ChaoticNeutrals/Synthetic-RP](https://huggingface.co/datasets/ChaoticNeutrals/Synthetic-RP)
|
||||
- [Gryphe/Sonnet3.5-SlimOrcaDedupCleaned](https://huggingface.co/datasets/Gryphe/Sonnet3.5-SlimOrcaDedupCleaned)
|
||||
- [Gryphe/Opus-WritingPrompts](https://huggingface.co/datasets/Gryphe/Opus-WritingPrompts)
|
||||
- [meseca/writing-opus-6k](https://huggingface.co/datasets/meseca/writing-opus-6k)
|
||||
- [meseca/opus-instruct-9k](https://huggingface.co/datasets/meseca/opus-instruct-9k)
|
||||
- [PJMixers/grimulkan_theory-of-mind-ShareGPT](https://huggingface.co/datasets/PJMixers/grimulkan_theory-of-mind-ShareGPT)
|
||||
- [NobodyExistsOnTheInternet/ToxicQAFinal](https://huggingface.co/datasets/NobodyExistsOnTheInternet/ToxicQAFinal)
|
||||
- [Undi95/toxic-dpo-v0.1-sharegpt](https://huggingface.co/datasets/Undi95/toxic-dpo-v0.1-sharegpt)
|
||||
- [cgato/SlimOrcaDedupCleaned](https://huggingface.co/datasets/cgato/SlimOrcaDedupCleaned)
|
||||
- [kalomaze/Opus_Instruct_25k](https://huggingface.co/datasets/kalomaze/Opus_Instruct_25k)
|
||||
- [Doctor-Shotgun/no-robots-sharegpt](https://huggingface.co/datasets/Doctor-Shotgun/no-robots-sharegpt)
|
||||
- [Norquinal/claude_multiround_chat_30k](https://huggingface.co/datasets/Norquinal/claude_multiround_chat_30k)
|
||||
- [nothingiisreal/Claude-3-Opus-Instruct-15K](https://huggingface.co/datasets/nothingiisreal/Claude-3-Opus-Instruct-15K)
|
||||
- All the Aesirs dataset, cleaned, unslopped
|
||||
- All le luminae dataset, cleaned, unslopped
|
||||
- Small part of Airoboros reduced
|
||||
|
||||
We sadly didn't find the sources of the following, DM us if you recognize your set !
|
||||
|
||||
- Opus_Instruct-v2-6.5K-Filtered-v2-sharegpt
|
||||
- claude_sharegpt_trimmed
|
||||
- CapybaraPure_Decontaminated-ShareGPT_reduced
|
||||
|
||||
## Datasets credits:
|
||||
- Epiculous
|
||||
- ChaoticNeutrals
|
||||
- Gryphe
|
||||
- meseca
|
||||
- PJMixers
|
||||
- NobodyExistsOnTheInternet
|
||||
- cgato
|
||||
- kalomaze
|
||||
- Doctor-Shotgun
|
||||
- Norquinal
|
||||
- nothingiisreal
|
||||
|
||||
## Others
|
||||
|
||||
Undi: If you want to support us, you can [here](https://ko-fi.com/undiai).
|
||||
|
||||
IkariDev: Visit my [retro/neocities style website](https://ikaridevgit.github.io/) please kek
|
||||
|
||||
</details>
|
||||
2422
imatrix-with-rp-ex.txt
Normal file
2422
imatrix-with-rp-ex.txt
Normal file
File diff suppressed because it is too large
Load Diff
3
imatrix.dat
Normal file
3
imatrix.dat
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:c46f4a465cc0de6d2cda125f85992823d327cca13d7cb51002c432a35e4d526c
|
||||
size 4988171
|
||||
Reference in New Issue
Block a user