初始化项目,由ModelHub XC社区提供模型

Model: Lewdiculous/Lumimaid-v0.2-8B-GGUF-IQ-Imatrix
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-25 14:15:10 +08:00
commit c830bf984a
15 changed files with 2659 additions and 0 deletions

47
.gitattributes vendored Normal file
View File

@@ -0,0 +1,47 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
imatrix.dat filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-BF16.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-F16.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-IQ3_M-imat.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-IQ3_XXS-imat.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-IQ4_XS-imat.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-Q4_K_M-imat.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-Q4_K_S-imat.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-Q5_K_M-imat.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-Q5_K_S-imat.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-Q6_K-imat.gguf filter=lfs diff=lfs merge=lfs -text
Lumimaid-v0.2-8B-Q8_0-imat.gguf filter=lfs diff=lfs merge=lfs -text

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d770bce46a00a3d5b1bd86e477ebebd37d562b72e5e02c46ed3dce1f702803f6
size 16068891552

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:42d03de9a0b2e392c3edfbe71d21ab94bff94fa5683543afff4c65fea5c84256
size 16068891552

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e331e55ca24095d1ce21419521cdd3d9917409b98bcd47791c5cab3fa6e9c48b
size 3784824000

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:761b00ed0352354b1784aa3d554e09ca5faeedb90177ff9c9047eac50362df03
size 3274912960

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:261762e2d9e9edc2b5e5053cb48cdbf52748801f139a95c5f6ccc95092d89e6e
size 4447663296

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f7b9e1db31efcb91b6c8c338a7ccec8d7bdcbc78491cbd6a82d5068eb2f6ec40
size 4920734912

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:71331ba953b9e6c428bbc7395a93ac449cea86170d10049cc653ad4092f255e2
size 4692669632

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b3cb02b76f149df78bc65eac4f59f16574cc799f35d358d557d2f1dca43c31d3
size 5732988096

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ef227f43d975f307a893c59c97ee86718cb755f53a26d5da8c6ee3afd858cb22
size 5599294656

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a9761044727e2a9a65065d2b04658a253b6e2d4d4a559a37f22c170c42b96dae
size 6596007104

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:7532e1accff0594791994150ff11c7fd5695bb2da1c7bba44394bf20a875c572
size 8540771520

154
README.md Normal file
View File

@@ -0,0 +1,154 @@
---
base_model: NeverSleep/Lumimaid-v0.2-8B
quantized_by: Lewdiculous
library_name: transformers
license: cc-by-nc-4.0
inference: false
language:
- en
tags:
- roleplay
- llama3
- sillytavern
---
# #roleplay #sillytavern #llama3
My GGUF-IQ-Imatrix quants for [**NeverSleep/Lumimaid-v0.2-8B**](https://huggingface.co/NeverSleep/Lumimaid-v0.2-8B).
I recommend checking their page for feedback and support.
> [!IMPORTANT]
> **Quantization process:** <br>
> Imatrix data was generated from the FP16-GGUF and conversions directly from the BF16-GGUF. <br>
> This is a bit more disk and compute intensive but hopefully avoids any losses during conversion. <br>
> To run this model, please use the [**latest version of KoboldCpp**](https://github.com/LostRuins/koboldcpp/releases/latest). <br>
> If you noticed any issues let me know in the discussions.
> [!NOTE]
> **Presets:** <br>
> * Llama-3. <br>
>
> Some compatible SillyTavern presets can be found [**here (Virt's Roleplay Presets - v1.9)**](https://huggingface.co/Virt-io/SillyTavern-Presets). <br>
> Check [**discussions such as this one**](https://huggingface.co/Virt-io/SillyTavern-Presets/discussions/5#664d6fb87c563d4d95151baa) and [**this one**](https://www.reddit.com/r/SillyTavernAI/comments/1dff2tl/my_personal_llama3_stheno_presets/) for other presets and samplers recommendations. <br>
> Lower temperatures are recommended by the authors, so make sure to experiment. <br>
>
> **General usage with KoboldCpp:** <br>
> For **8GB VRAM** GPUs, I recommend the **Q4_K_M-imat** (4.89 BPW) quant for up to 12288 context sizes without the use of `--quantkv`. <br>
> Using `--quantkv 1` (≈Q8) or even `--quantkv 2` (≈Q4) can get you to 32K context sizes with the caveat of not being compatible with Context Shifting, only relevant if you can manage to fill up that much context. <br>
> [**Read more about it in the release here**](https://github.com/LostRuins/koboldcpp/releases/tag/v1.67).
<details>
<summary>⇲ Click here to expand/hide information General chart with relative quant parformances.</summary>
> [!NOTE]
> **Recommended read:** <br>
>
> [**"Which GGUF is right for me? (Opinionated)" by Artefact2**](https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9)
>
> *Click the image to view full size.*
> !["Which GGUF is right for me? (Opinionated)" by Artefact2 - Firs Graph](https://cdn-uploads.huggingface.co/production/uploads/65d4cf2693a0a3744a27536c/fScWdHIPix5IzNJ8yswCB.webp)
</details>
> [!TIP]
> **Personal-support:** <br>
> I apologize for disrupting your experience. <br>
> Eventually I may be able to use a dedicated server for this, but for now hopefully these quants are helpful. <br>
> If you **want** and you are **able to**... <br>
> You can [**spare some change over here (Ko-fi)**](https://ko-fi.com/Lewdiculous). <br>
>
> **Author-support:** <br>
> You can support the authors [**at their pages**](https://ko-fi.com/undiai)/[**here**](https://ikaridevgit.github.io/).
![image/png](https://cdn-uploads.huggingface.co/production/uploads/65d4cf2693a0a3744a27536c/qEH7KuSGfUGXSHeyWSwS-.png)
<details>
<summary>Original model card information.</summary>
## **Original card:**
## Lumimaid 0.2
<img src="https://cdn-uploads.huggingface.co/production/uploads/63ab1241ad514ca8d1430003/TUcHg7LKNjfo0sni88Ps7.png" alt="Image" style="display: block; margin-left: auto; margin-right: auto; width: 65%;">
<div style="text-align: center; font-size: 30px;">
<a href="https://huggingface.co/NeverSleep/Lumimaid-v0.2-8B">[8b]</a> -
<a href="https://huggingface.co/NeverSleep/Lumimaid-v0.2-12B">12b</a> -
<a href="https://huggingface.co/NeverSleep/Lumimaid-v0.2-70B">70b</a> -
<a href="https://huggingface.co/NeverSleep/Lumimaid-v0.2-123B">123b</a>
</div>
### This model is based on: [Meta-Llama-3.1-8B-Instruct](https://huggingface.co/meta-llama/Meta-Llama-3.1-8B-Instruct)
Wandb: https://wandb.ai/undis95/Lumi-Llama-3-1-8B?nw=nwuserundis95
Lumimaid 0.1 -> 0.2 is a HUGE step up dataset wise.
As some people have told us our models are sloppy, Ikari decided to say fuck it and literally nuke all chats out with most slop.
Our dataset stayed the same since day one, we added data over time, cleaned them, and repeat. After not releasing model for a while because we were never satisfied, we think it's time to come back!
## Prompt template: Llama-3-Instruct
```
<|begin_of_text|><|start_header_id|>system<|end_header_id|>
{system_prompt}<|eot_id|><|start_header_id|>user<|end_header_id|>
{input}<|eot_id|><|start_header_id|>assistant<|end_header_id|>
{output}<|eot_id|>
```
## Credits:
- Undi
- IkariDev
## Training data we used to make our dataset:
- [Epiculous/Gnosis](https://huggingface.co/Epiculous/Gnosis)
- [ChaoticNeutrals/Luminous_Opus](https://huggingface.co/datasets/ChaoticNeutrals/Luminous_Opus)
- [ChaoticNeutrals/Synthetic-Dark-RP](https://huggingface.co/datasets/ChaoticNeutrals/Synthetic-Dark-RP)
- [ChaoticNeutrals/Synthetic-RP](https://huggingface.co/datasets/ChaoticNeutrals/Synthetic-RP)
- [Gryphe/Sonnet3.5-SlimOrcaDedupCleaned](https://huggingface.co/datasets/Gryphe/Sonnet3.5-SlimOrcaDedupCleaned)
- [Gryphe/Opus-WritingPrompts](https://huggingface.co/datasets/Gryphe/Opus-WritingPrompts)
- [meseca/writing-opus-6k](https://huggingface.co/datasets/meseca/writing-opus-6k)
- [meseca/opus-instruct-9k](https://huggingface.co/datasets/meseca/opus-instruct-9k)
- [PJMixers/grimulkan_theory-of-mind-ShareGPT](https://huggingface.co/datasets/PJMixers/grimulkan_theory-of-mind-ShareGPT)
- [NobodyExistsOnTheInternet/ToxicQAFinal](https://huggingface.co/datasets/NobodyExistsOnTheInternet/ToxicQAFinal)
- [Undi95/toxic-dpo-v0.1-sharegpt](https://huggingface.co/datasets/Undi95/toxic-dpo-v0.1-sharegpt)
- [cgato/SlimOrcaDedupCleaned](https://huggingface.co/datasets/cgato/SlimOrcaDedupCleaned)
- [kalomaze/Opus_Instruct_25k](https://huggingface.co/datasets/kalomaze/Opus_Instruct_25k)
- [Doctor-Shotgun/no-robots-sharegpt](https://huggingface.co/datasets/Doctor-Shotgun/no-robots-sharegpt)
- [Norquinal/claude_multiround_chat_30k](https://huggingface.co/datasets/Norquinal/claude_multiround_chat_30k)
- [nothingiisreal/Claude-3-Opus-Instruct-15K](https://huggingface.co/datasets/nothingiisreal/Claude-3-Opus-Instruct-15K)
- All the Aesirs dataset, cleaned, unslopped
- All le luminae dataset, cleaned, unslopped
- Small part of Airoboros reduced
We sadly didn't find the sources of the following, DM us if you recognize your set !
- Opus_Instruct-v2-6.5K-Filtered-v2-sharegpt
- claude_sharegpt_trimmed
- CapybaraPure_Decontaminated-ShareGPT_reduced
## Datasets credits:
- Epiculous
- ChaoticNeutrals
- Gryphe
- meseca
- PJMixers
- NobodyExistsOnTheInternet
- cgato
- kalomaze
- Doctor-Shotgun
- Norquinal
- nothingiisreal
## Others
Undi: If you want to support us, you can [here](https://ko-fi.com/undiai).
IkariDev: Visit my [retro/neocities style website](https://ikaridevgit.github.io/) please kek
</details>

2422
imatrix-with-rp-ex.txt Normal file

File diff suppressed because it is too large Load Diff

3
imatrix.dat Normal file
View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c46f4a465cc0de6d2cda125f85992823d327cca13d7cb51002c432a35e4d526c
size 4988171