初始化项目,由ModelHub XC社区提供模型

Model: mradermacher/LightGPT-7B-Llama2-i1-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-05-14 01:00:35 +08:00
commit 35549979ac
24 changed files with 353 additions and 0 deletions

57
.gitattributes vendored Normal file
View File

@@ -0,0 +1,57 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
imatrix.dat filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ3_M.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ3_XXS.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q3_K_M.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ2_M.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ1_M.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ4_XS.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ2_XXS.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q3_K_L.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ2_XS.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ2_S.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ1_S.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q4_0.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ3_XS.gguf filter=lfs diff=lfs merge=lfs -text
LightGPT-7B-Llama2.i1-IQ3_S.gguf filter=lfs diff=lfs merge=lfs -text

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:08336f31dcc75cfa0082dd2023835d66b07016006aebc7119ca07584b42c6d3e
size 1650971584

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:6bd2b935fd3d8b942363238ca8ccc4816bb91f3cf7ef4d07f95b199c7a2d8afa
size 1528583104

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ec6da0ae406518c7050cda0e8f2ee0b722276718c8d77092faa52525601e7cb1
size 2359751616

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1e01b1c4981c5b6b7b756650bfc407b474ed7feb480e2803620ba60b387d2d12
size 2196566976

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:dd0c3b2a9e8f025dd9e24688244c0797be02d231c9679831461c6181d732d6be
size 2034914240

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2301f42b336d00e21cc5a5d3c3cf3f27d7d4997329c63251b0c9f2590ccf26b8
size 1854952384

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:646440c75df862ef40790804e3b1c9ab424f94876e1ba6bd990ccbad776266b5
size 3114865600

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8b689c9991da82c4c7b7793ff5acf474ac71baff4610e4875f0b00a8cb261260
size 2948305856

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ebafbdce1d3b9f8894ae6e774ee91c32ae51eb692b47bf3437ec7b628ddddc11
size 2796524480

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e51a75dcececa400934266b80db19ae2c0810a8f368bc8f040e90c705edaeab6
size 2585392064

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f6a3ab72ab1a575b8f750eb7cccc1d53daf4055300cb8bb07567d0785fe68279
size 3619337152

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:e51d6ffacd09ae30959a21b4b1ea2dff966c28619b286e6a79071a07235d9886
size 2532864960

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2f00c4479ca47b4b30d502e3e3e087fbcc4adb4945e1dfde2d80d2b761fb7803
size 3597112256

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:42ef5148e36e0f18aa800bec8f5663a486c254952aec0103f0c5a9588dfd3b2c
size 3298005952

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:f5c71868685368162cfce42628d5c0aa9ff1e937aecaaa3dc71f08cb51f04fdd
size 2948305856

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:9028cd252bed77a94893ac4114d47ffeddc62fad0761b031bb130ba731842d48
size 3837080512

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b2709f16561bb101ccbf086a4520c6630d75dbd2de31b7e49f79ccd244a4cfb3
size 4081005504

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b5be2340b12b5cca8dda0fb0c4c0092add17b8d03e27cf5c6df2749007d9e5ba
size 3856741312

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:a2850272d17674e13fd1317c213bec060818980c1e5770101f4ff8531f779d0d
size 4783158208

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2317e260c21b0ff6dbc8fced59b68acd5a15547de0c6525e373c57a53951950e
size 4651692992

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:748105e7d1223a6d06be7dbfab91bc7a36fb27d0b207502122cb5299fc8bd011
size 5529195456

230
README.md Normal file
View File

@@ -0,0 +1,230 @@
---
base_model: lightgpt/LightGPT-7B-Llama2
extra_gated_button_content: Submit
extra_gated_fields:
Affiliation: text
? By clicking Submit below I accept the terms of the license and acknowledge that
the information I provide will be collected stored processed and shared in accordance
with the Meta Privacy Policy
: checkbox
Country: country
Date of birth: date_picker
First Name: text
Last Name: text
geo: ip_location
extra_gated_heading: Access LLMLight-LightGPT on Hugging Face
extra_gated_prompt: "### LLAMA 2 COMMUNITY LICENSE AGREEMENT\n\"Agreement\" means
the terms and conditions for use, reproduction, distribution and modification of
the Llama Materials set forth herein. \n\"Documentation\" means the specifications,
manuals and documentation accompanying Llama 2 distributed by Meta at https://ai.meta.com/resources/models-and-libraries/llama-downloads/.
\ \n\"Licensee\" or \"you\" means you, or your employer or any other person or entity
(if you are entering into this Agreement on such person or entity's behalf), of
the age required under applicable laws, rules or regulations to provide legal consent
and that has legal authority to bind your employer or such other person or entity
if you are entering in this Agreement on their behalf. \n\"Llama 2\" means the
foundational large language models and software and algorithms, including machine-learning
model code, trained model weights, inference-enabling code, training-enabling code,
fine-tuning enabling code and other elements of the foregoing distributed by Meta
at ai.meta.com/resources/models-and-libraries/llama-downloads/.\n\"Llama Materials\"
means, collectively, Meta's proprietary Llama 2 and documentation (and any portion
thereof) made available under this Agreement.\n\"Meta\" or \"we\" means Meta Platforms
Ireland Limited (if you are located in or, if you are an entity, your principal
place of business is in the EEA or Switzerland) and Meta Platforms, Inc. (if you
are located outside of the EEA or Switzerland). \n\nBy clicking \"I Accept\" below
or by using or distributing any portion or element of the Llama Materials, you agree
to be bound by this Agreement.\n1. License Rights and Redistribution. \na. Grant
of Rights. You are granted a non-exclusive, worldwide, non- transferable and royalty-free
limited license under Meta's intellectual property or other rights owned by Meta
embodied in the Llama Materials to use, reproduce, distribute, copy, create derivative
works of, and make modifications to the Llama Materials. \nb. Redistribution and
Use.\ni. If you distribute or make the Llama Materials, or any derivative works
\ thereof, available to a third party, you shall provide a copy of this Agreement
to such third party. \nii. If you receive Llama Materials, or any derivative works
thereof, from a Licensee as part of an integrated end user product, then Section
2 of this Agreement will not apply to you. \niii. You must retain in all copies
of the Llama Materials that you distribute the following attribution notice within
a \"Notice\" text file distributed as a part of such copies: \"Llama 2 is licensed
under the LLAMA 2 Community License, Copyright (c) Meta Platforms, Inc. All Rights
Reserved.\"\niv. Your use of the Llama Materials must comply with applicable laws
\ and regulations (including trade compliance laws and regulations) and adhere to
the Acceptable Use Policy for the Llama Materials (available at https://ai.meta.com/llama/use-policy),
which is hereby incorporated by reference into this Agreement.\nv. You will not
use the Llama Materials or any output or results of the Llama Materials to improve
any other large language model (excluding Llama 2 or derivative works thereof).
\ \n\n2. Additional Commercial Terms. If, on the Llama 2 version release date, the
\ monthly active users of the products or services made available by or for Licensee,
\ or Licensee's affiliates, is greater than 700 million monthly active users in
the preceding calendar month, you must request a license from Meta, which Meta
may grant to you in its sole discretion, and you are not authorized to exercise
any of the rights under this Agreement unless or until Meta otherwise expressly
grants you such rights.\n3. Disclaimer of Warranty. UNLESS REQUIRED BY APPLICABLE
LAW, THE LLAMA MATERIALS AND ANY OUTPUT AND RESULTS THEREFROM ARE PROVIDED ON
AN \"AS IS\" BASIS, WITHOUT WARRANTIES OF ANY KIND, EITHER EXPRESS OR IMPLIED,
INCLUDING, WITHOUT LIMITATION, ANY WARRANTIES OF TITLE, NON-INFRINGEMENT, MERCHANTABILITY,
OR FITNESS FOR A PARTICULAR PURPOSE. YOU ARE SOLELY RESPONSIBLE FOR DETERMINING
THE APPROPRIATENESS OF USING OR REDISTRIBUTING THE LLAMA MATERIALS AND ASSUME ANY
RISKS ASSOCIATED WITH YOUR USE OF THE LLAMA MATERIALS AND ANY OUTPUT AND RESULTS.\n4.
Limitation of Liability. IN NO EVENT WILL META OR ITS AFFILIATES BE LIABLE UNDER
ANY THEORY OF LIABILITY, WHETHER IN CONTRACT, TORT, NEGLIGENCE, PRODUCTS LIABILITY,
OR OTHERWISE, ARISING OUT OF THIS AGREEMENT, FOR ANY LOST PROFITS OR ANY INDIRECT,
SPECIAL, CONSEQUENTIAL, INCIDENTAL, EXEMPLARY OR PUNITIVE DAMAGES, EVEN IF META
OR ITS AFFILIATES HAVE BEEN ADVISED OF THE POSSIBILITY OF ANY OF THE FOREGOING.\n\n5.
Intellectual Property.\na. No trademark licenses are granted under this Agreement,
and in connection with the Llama Materials, neither Meta nor Licensee may use any
name or mark owned by or associated with the other or any of its affiliates, except
as required for reasonable and customary use in describing and redistributing the
\ Llama Materials.\nb. Subject to Meta's ownership of Llama Materials and derivatives
made by or for Meta, with respect to any derivative works and modifications of
the Llama Materials that are made by you, as between you and Meta, you are and
will be the owner of such derivative works and modifications.\nc. If you institute
litigation or other proceedings against Meta or any entity (including a cross-claim
or counterclaim in a lawsuit) alleging that the Llama Materials or Llama 2 outputs
or results, or any portion of any of the foregoing, constitutes infringement of
intellectual property or other rights owned or licensable by you, then any licenses
granted to you under this Agreement shall terminate as of the date such litigation
or claim is filed or instituted. You will indemnify and hold harmless Meta from
and against any claim by any third party arising out of or related to your use
or distribution of the Llama Materials.\n6. Term and Termination. The term of this
Agreement will commence upon your acceptance of this Agreement or access to the
Llama Materials and will continue in full force and effect until terminated in
accordance with the terms and conditions herein. Meta may terminate this Agreement
if you are in breach of any term or condition of this Agreement. Upon termination
of this Agreement, you shall delete and cease use of the Llama Materials. Sections
3, 4 and 7 shall survive the termination of this Agreement. \n7. Governing Law
and Jurisdiction. This Agreement will be governed and construed under the laws
of the State of California without regard to choice of law principles, and the
UN Convention on Contracts for the International Sale of Goods does not apply to
this Agreement. The courts of California shall have exclusive jurisdiction of any
dispute arising out of this Agreement. \n### Llama 2 Acceptable Use Policy\nMeta
is committed to promoting safe and fair use of its tools and features, including
Llama 2. If you access or use Llama 2, you agree to this Acceptable Use Policy (“Policy”).
The most recent copy of this policy can be found at [ai.meta.com/llama/use-policy](http://ai.meta.com/llama/use-policy).\n####
Prohibited Uses\nWe want everyone to use Llama 2 safely and responsibly. You agree
you will not use, or allow others to use, Llama 2 to:\n1. Violate the law or others’
rights, including to:\n 1. Engage in, promote, generate, contribute to, encourage,
plan, incite, or further illegal or unlawful activity or content, such as: \n 1.
Violence or terrorism \n 2. Exploitation or harm to children, including
the solicitation, creation, acquisition, or dissemination of child exploitative
content or failure to report Child Sexual Abuse Material\n 3. Human trafficking,
exploitation, and sexual violence\n 4. The illegal distribution of information
or materials to minors, including obscene materials, or failure to employ legally
required age-gating in connection with such information or materials.\n 5.
Sexual solicitation\n 6. Any other criminal activity\n 2. Engage in,
promote, incite, or facilitate the harassment, abuse, threatening, or bullying of
individuals or groups of individuals\n 3. Engage in, promote, incite, or facilitate
discrimination or other unlawful or harmful conduct in the provision of employment,
employment benefits, credit, housing, other economic benefits, or other essential
goods and services\n 4. Engage in the unauthorized or unlicensed practice of
any profession including, but not limited to, financial, legal, medical/health,
or related professional practices \n 5. Collect, process, disclose, generate,
or infer health, demographic, or other sensitive personal or private information
about individuals without rights and consents required by applicable laws\n 6.
Engage in or facilitate any action or generate any content that infringes, misappropriates,
or otherwise violates any third-party rights, including the outputs or results of
any products or services using the Llama 2 Materials\n 7. Create, generate,
or facilitate the creation of malicious code, malware, computer viruses or do anything
else that could disable, overburden, interfere with or impair the proper working,
integrity, operation or appearance of a website or computer system \n2. Engage in,
promote, incite, facilitate, or assist in the planning or development of activities
that present a risk of death or bodily harm to individuals, including use of Llama
2 related to the following:\n 1. Military, warfare, nuclear industries or applications,
espionage, use for materials or activities that are subject to the International
Traffic Arms Regulations (ITAR) maintained by the United States Department of State\n
\ 2. Guns and illegal weapons (including weapon development)\n 3. Illegal drugs
and regulated/controlled substances\n 4. Operation of critical infrastructure,
transportation technologies, or heavy machinery\n 5. Self-harm or harm to others,
including suicide, cutting, and eating disorders\n 6. Any content intended to
incite or promote violence, abuse, or any infliction of bodily harm to an individual\n3.
Intentionally deceive or mislead others, including use of Llama 2 related to the
following:\n 1. Generating, promoting, or furthering fraud or the creation or
promotion of disinformation\n 2. Generating, promoting, or furthering defamatory
content, including the creation of defamatory statements, images, or other content\n
\ 3. Generating, promoting, or further distributing spam\n 4. Impersonating
another individual without consent, authorization, or legal right\n 5. Representing
that the use of Llama 2 or outputs are human-generated\n 6. Generating or facilitating
false online engagement, including fake reviews and other means of fake online engagement
\n 4. Fail to appropriately disclose to end users any known dangers of your AI
system \nPlease report any violation of this Policy, software “bug,” or other problems
that could lead to a violation of this Policy through one of the following means:
\n * Reporting issues with the model: [github.com/facebookresearch/llama](http://github.com/facebookresearch/llama)\n
\ * Reporting risky content generated by the model: [developers.facebook.com/llama_output_feedback](http://developers.facebook.com/llama_output_feedback)\n
\ * Reporting bugs and security concerns: [facebook.com/whitehat/info](http://facebook.com/whitehat/info)
\n * Reporting violations of the Acceptable Use Policy or unlicensed uses of
Llama: [LlamaUseReport@meta.com](mailto:LlamaUseReport@meta.com)"
language:
- en
library_name: transformers
license: mit
quantized_by: mradermacher
tags:
- pytorch
- llama-2
- traffic signal control
- lightgpt
- llmlight
---
## About
<!-- ### quantize_version: 2 -->
<!-- ### output_tensor_quantised: 1 -->
<!-- ### convert_type: hf -->
<!-- ### vocab_type: -->
<!-- ### tags: nicoboss -->
weighted/imatrix quants of https://huggingface.co/lightgpt/LightGPT-7B-Llama2
<!-- provided-files -->
static quants are available at https://huggingface.co/mradermacher/LightGPT-7B-Llama2-GGUF
## Usage
If you are unsure how to use GGUF files, refer to one of [TheBloke's
READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for
more details, including on how to concatenate multi-part files.
## Provided Quants
(sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
| Link | Type | Size/GB | Notes |
|:-----|:-----|--------:|:------|
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ1_S.gguf) | i1-IQ1_S | 1.6 | for the desperate |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ1_M.gguf) | i1-IQ1_M | 1.8 | mostly desperate |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ2_XXS.gguf) | i1-IQ2_XXS | 2.0 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ2_XS.gguf) | i1-IQ2_XS | 2.1 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ2_S.gguf) | i1-IQ2_S | 2.3 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ2_M.gguf) | i1-IQ2_M | 2.5 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q2_K.gguf) | i1-Q2_K | 2.6 | IQ3_XXS probably better |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ3_XXS.gguf) | i1-IQ3_XXS | 2.7 | lower quality |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ3_XS.gguf) | i1-IQ3_XS | 2.9 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ3_S.gguf) | i1-IQ3_S | 3.0 | beats Q3_K* |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q3_K_S.gguf) | i1-Q3_K_S | 3.0 | IQ3_XS probably better |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ3_M.gguf) | i1-IQ3_M | 3.2 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q3_K_M.gguf) | i1-Q3_K_M | 3.4 | IQ3_S probably better |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q3_K_L.gguf) | i1-Q3_K_L | 3.7 | IQ3_M probably better |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-IQ4_XS.gguf) | i1-IQ4_XS | 3.7 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q4_0.gguf) | i1-Q4_0 | 3.9 | fast, low quality |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q4_K_S.gguf) | i1-Q4_K_S | 4.0 | optimal size/speed/quality |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q4_K_M.gguf) | i1-Q4_K_M | 4.2 | fast, recommended |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q5_K_S.gguf) | i1-Q5_K_S | 4.8 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q5_K_M.gguf) | i1-Q5_K_M | 4.9 | |
| [GGUF](https://huggingface.co/mradermacher/LightGPT-7B-Llama2-i1-GGUF/resolve/main/LightGPT-7B-Llama2.i1-Q6_K.gguf) | i1-Q6_K | 5.6 | practically like static Q6_K |
Here is a handy graph by ikawrakow comparing some lower-quality quant
types (lower is better):
![image.png](https://www.nethype.de/huggingface_embed/quantpplgraph.png)
And here are Artefact2's thoughts on the matter:
https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9
## FAQ / Model Request
See https://huggingface.co/mradermacher/model_requests for some answers to
questions you might have and/or if you want some other model quantized.
## Thanks
I thank my company, [nethype GmbH](https://www.nethype.de/), for letting
me use its servers and providing upgrades to my workstation to enable
this work in my free time. Additional thanks to [@nicoboss](https://huggingface.co/nicoboss) for giving me access to his private supercomputer, enabling me to provide many more imatrix quants, at much higher quality, than I would otherwise be able to.
<!-- end -->

3
imatrix.dat Normal file
View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d581b21bdb798f6dd7584c5b2f000c3fc2ee96f05eb1ff3f6864f2ad1f7b8428
size 4562173