初始化项目,由ModelHub XC社区提供模型
Model: qwp4w3hyb/Llama-3-11.5B-Instruct-V2-iMat-GGUF Source: Original Platform
This commit is contained in:
55
.gitattributes
vendored
Normal file
55
.gitattributes
vendored
Normal file
@@ -0,0 +1,55 @@
|
||||
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||
*.model filter=lfs diff=lfs merge=lfs -text
|
||||
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||
imat-bf16-gmerged.dat filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-bf16.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ1_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ2_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ2_XXS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ3_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ3_XXS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ4_NL.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
llama-3-11.5b-instruct-v2-imat-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||
50
README.md
Normal file
50
README.md
Normal file
@@ -0,0 +1,50 @@
|
||||
---
|
||||
license: other
|
||||
license_name: llama-3
|
||||
license_link: https://llama.meta.com/llama3/license/
|
||||
pipeline_tag: text-generation
|
||||
base_model: Replete-AI/Llama-3-11.5B-Instruct-V2
|
||||
tags:
|
||||
- facebook
|
||||
- meta
|
||||
- pytorch
|
||||
- llama
|
||||
- llama-3
|
||||
- instruct
|
||||
- finetune
|
||||
- frankenmerge
|
||||
- merge
|
||||
- gguf
|
||||
- imatrix
|
||||
- importance matrix
|
||||
model-index:
|
||||
- name: Yi-1.5-34B-Chat-16K-iMat-GGUF
|
||||
results: []
|
||||
---
|
||||
|
||||
# Quant Infos
|
||||
|
||||
- quants done with an importance matrix for improved quantization loss
|
||||
- ggufs & imatrix generated from bf16 for "optimal" accuracy loss
|
||||
- Wide coverage of different gguf quant types from Q\_8\_0 down to IQ1\_S
|
||||
- Quantized with [llama.cpp](https://github.com/ggerganov/llama.cpp) commit [fabf30b4c4fca32e116009527180c252919ca922](https://github.com/ggerganov/llama.cpp/commit/fabf30b4c4fca32e116009527180c252919ca922) (master as of 2024-05-20)
|
||||
- Imatrix generated with [this](https://github.com/ggerganov/llama.cpp/discussions/5263#discussioncomment-8395384) multi-purpose dataset.
|
||||
```
|
||||
./imatrix -c 512 -m $model_name-f16.gguf -f $llama_cpp_path/groups_merged.txt -o $out_path/imat-f16-gmerged.dat
|
||||
```
|
||||
|
||||
# Original Model Card:
|
||||
|
||||
## Llama-3-11.5B-v2
|
||||
|
||||
Thank you to Meta for the weights for Meta-Llama-3-8B
|
||||
|
||||

|
||||
|
||||
This is an upscaling of the Meta-Llama-3-8B Ai using techniques created for chargoddard/mistral-11b-slimorca. This Ai model has been upscaled from 8b parameters to 11.5b parameters without any continuous pretraining or fine-tuning.
|
||||
|
||||
Unlike version 1 this model has no issues at fp16 or any quantizations.
|
||||
|
||||
The model that was used to create this one is linked below:
|
||||
|
||||
https://huggingface.co/meta-llama/Meta-Llama-3-8B
|
||||
3
imat-bf16-gmerged.dat
Normal file
3
imat-bf16-gmerged.dat
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:f8b33a6b085d1955a21273fc0cc40fd7827c6b8f471e4837c7c438e1bfff2188
|
||||
size 7482281
|
||||
3
llama-3-11.5b-instruct-v2-bf16.gguf
Normal file
3
llama-3-11.5b-instruct-v2-bf16.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:a948423fe2b42e15c1e91586cb458a2b01c1e61aad7059c5fd7fd71475ea7f9c
|
||||
size 23048745664
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ1_S.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ1_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:93a4eb11d507b50c3753cec6ebae1e62ac864862d55f313f9fc606920f1e5f6c
|
||||
size 2758751232
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ2_M.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ2_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:eb5fe05407f2740167d6d7d7893ac9c7bf85d4ca9f38e777a303cf09a1750ab7
|
||||
size 4125053952
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ2_S.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ2_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:51ab7409dc5522d69b2dad23d439e9826fb86d7034db1c0e8f6857f29987a4d9
|
||||
size 3840365568
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ2_XS.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ2_XS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:cb809506bebcca5b294d8bf1cfb4fd92bf4ae131ffe0fe20ba4f4bdad01a9532
|
||||
size 3637982208
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ2_XXS.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ2_XXS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:371b1dd4c9545d7f2c0fb8154a102d7618fea7f3c1076b5298e1fa6686e71b83
|
||||
size 3328128000
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ3_M.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ3_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:f7634b9b25be2d1540f1386af5a245a7493ae00e5f527811a1bfe2e336018abc
|
||||
size 5344982016
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ3_S.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ3_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:a2066640d61cd194c09ff50712e4d0f2295ad2f06cd20f771dd83dfae9d3f1b3
|
||||
size 5191234560
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ3_XS.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ3_XS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:252279b8d0fb4930c55a759476a0d0a7a8e40618e82153f2c4c4fdae1a0459ac
|
||||
size 4945867776
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ3_XXS.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ3_XXS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:97ddb238bd3d974b597634eae6777cafded47749ed5371bb3b63f496ed399255
|
||||
size 4615001088
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ4_NL.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ4_NL.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:c844890368bfeb4062144a89abdb74c27502539bb14177210a67adb0fd2b1243
|
||||
size 6649844736
|
||||
3
llama-3-11.5b-instruct-v2-imat-IQ4_XS.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-IQ4_XS.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:483be3b26cca8ce429b9831e3d5ed01b56d0d4112d078e36231fcda4d5e8b59b
|
||||
size 6312563712
|
||||
3
llama-3-11.5b-instruct-v2-imat-Q4_0.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-Q4_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:de13ef7dc5369235baba3f222e439bc6ecad80a292f2eb30b503312eb765cb05
|
||||
size 6646699008
|
||||
3
llama-3-11.5b-instruct-v2-imat-Q4_K_M.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-Q4_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:5711150c8f440163f09bad9e000fdade2066a0b0b5d80334dba977b619ba6374
|
||||
size 7013962752
|
||||
3
llama-3-11.5b-instruct-v2-imat-Q4_K_S.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-Q4_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:25617b45263871f55c6abfe12b21be3a995f243af36f68641d4f986e0be62516
|
||||
size 6670816256
|
||||
3
llama-3-11.5b-instruct-v2-imat-Q5_K_M.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-Q5_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:730a97287c5d10c18ef45b94ec0eabde6949b1f843e7264be84beff8c84dd618
|
||||
size 8199508992
|
||||
3
llama-3-11.5b-instruct-v2-imat-Q5_K_S.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-Q5_K_S.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:52cfc6f9990dcbb92428b8e86639f01decc802867231f1875d4cfbc2589e33bc
|
||||
size 7998968832
|
||||
3
llama-3-11.5b-instruct-v2-imat-Q6_K.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-Q6_K.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:a9d1ca7bbdeee5512d30014c5342a449454446e5774c1de437792fe5bb7f3105
|
||||
size 9459151872
|
||||
3
llama-3-11.5b-instruct-v2-imat-Q8_0.gguf
Normal file
3
llama-3-11.5b-instruct-v2-imat-Q8_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
||||
version https://git-lfs.github.com/spec/v1
|
||||
oid sha256:031bd32f3d875c15e9458c70826f9790adf01f37ccac733d0206459979c931d3
|
||||
size 12249068544
|
||||
Reference in New Issue
Block a user