初始化项目,由ModelHub XC社区提供模型

Model: Mungert/OpenCodeReasoning-Nemotron-14B-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-08-28 05:37:21 +08:00
commit f757bcd817
28 changed files with 359 additions and 0 deletions

75
.gitattributes vendored Normal file
View File

@@ -0,0 +1,75 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-f16.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-f16_q8_0.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-bf16_q8_0.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-f16_q6_k.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-bf16_q6_k.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-f16_q4_k.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-bf16_q4_k.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q2_k_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q3_k_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q5_k_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q4_0.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-bf16.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q5_0_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-iq2_s.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-iq2_xs.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q8_0.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q4_1_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q5_k_s.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q5_0.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q4_0_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-iq1_m.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q4_1.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B.imatrix filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q5_k_m.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-iq3_xxs.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q4_k_m.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q2_k_m.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q5_1_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q6_k_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-iq3_xs.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-iq1_s.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q4_k_s.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q4_k_l.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-iq2_m.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-iq2_xxs.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q3_k_m.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q5_1.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q3_k_s.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q2_k_s.gguf filter=lfs diff=lfs merge=lfs -text
OpenCodeReasoning-Nemotron-14B-q6_k_m.gguf filter=lfs diff=lfs merge=lfs -text

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:69cb118b4722c2193f8e681f239de0fcbae6abafd18f3fa70d76456569eab678
size 29547716864

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:31b508cb226dd72ce47fba99316afb08de2315e19bca01d09e388a38c2d1bc7e
size 17161412864

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:59adce99ab3250b9d505dac396c6a8662152b2575f72218f195498467795d1ea
size 17161412864

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d6edfbdca8504e774b6896dfe4bce219e2cc31ae9729fcd82c473d5ab644ba48
size 4840235552

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:130e14a93441765ee908aa4b26b8ec983cbfba4509bafc8fbb3f57ee86a2622e
size 4515586592

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:feb7594e721c330d1a318b8d32f3704f6f33f501ff899597c17c1967de7b81e0
size 5719319072

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:4e1ca918bc1e4e4df34bf06eb07a14b0ea1bf181750f3659f349cf857c80d4cf
size 5479457312

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c412b6cb907cc8ecd1ec171b43068bef2d63715d228e06c03926a4e50ea07676
size 5317992992

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:984170760b6a4951f36f36f529a785e77413d2e2309d584b8181fc9f064bb099
size 4940259872

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:4446b825586c246927f955f28e2da755ca8075be21de0d07ed954b3cd2018ff9
size 6473720352

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:332b53a26aa8269b19d3ae5e14fa8d861ec2a71a9e0d66d07aa2e38ae04dd01b
size 6173647392

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:c68d9d4da116ad2921467531d52123c782b78ba3944e852a699ca8ae76a9c722
size 6098813472

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:1b2854b225acd545dd9672261d688f977484e03c78c76ed4275127ba0a14701b
size 5534261792

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:0ac36565668eda2046119e8c2b36068ad5a4463b05b63efd6defe6c44bcc0bf4
size 7583531552

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d837a91cb201b248d9ba9826621bcd3169c063c86f59b16925daaa12b009ac72
size 6860751392

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:80cfd3f8f3d391491d3a0dfaa7b8e0834640e36ed23d863bff419d24185e624b
size 8317002272

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:04b04c308a5376eabc8c29113229f892039069615924d0c9117d2c5fce19352a
size 9240076832

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:ddc5a10ba6535ed6b2fba11ecec488d1c29713c1a5a30769d9dcb6a6569b2aff
size 9120637472

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:8d7dcfb7fe018567872d0a93ac1171bdb4cff0af7df8e97cabc9d572735e8c00
size 8855093792

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:d0a65a13b1ab1e949ee236feb1032128e2c3467de63a638594aa9eefda2a505a
size 10163151392

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:557bc4937f6cf52a1bc95b49ab622711f5d4450c8f1b57fb74812ec880ba7611
size 11086225952

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:224af6892375fb2311a9fe3e14af3f3e502e56c81d4e19d4ee27b088b4b5ee3d
size 10647093792

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:b0564bf6cceed8e9a0c09db225cb02709632123702332dc1b6d0c82ea62e7ef7
size 10505740832

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:80d0ff34db24401c642b437ffc441d5534c949c04a0549eee65ac13014b1ba86
size 12124684832

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:73aa4ba5ed7a9d0a6102de7e0425f5d4128def5aec80a6bcb7285630450717e3
size 15701598464

View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:76497cf67e076dd9ec67db18857aa0a39326ba556af02065977497f5c794ee7b
size 8563610

206
README.md Normal file
View File

@@ -0,0 +1,206 @@
---
base_model:
- Qwen/Qwen2.5-14B-Instruct
datasets:
- nvidia/OpenCodeReasoning
language:
- en
library_name: transformers
license: apache-2.0
tags:
- nvidia
- code
pipeline_tag: text-generation
---
# OpenCodeReasoning-Nemotron-14B Overview
## Description: <br>
OpenCodeReasoning-Nemotron-14B is a large language model (LLM) which is a derivative of Qwen2.5-14B-Instruct (AKA the reference model). It is a reasoning model that is post-trained for reasoning for code generation. The model supports a context length of 32K tokens. <br>
This model is ready for commercial/non-commercial use. <br>
![Evaluation Results](./results.png)
## Results from [OpenCodeReasoning](https://arxiv.org/abs/2504.01943)
Below results are the average of **64 evaluations** on each benchmark.
| Model | LiveCodeBench Avg. | CodeContest All |
|------------------------|--------------------|-----------------|
| DeepSeek-R1 | 65.6 | 26.2 |
| QwQ-32B | 61.3 | 20.2 |
| | | |
| **Distilled 7B+ Models** | | |
| | | |
| Bespoke-Stratos-7B | 14.7 | 2.0 |
| OpenThinker-7B | 25.5 | 5.0 |
| R1-Distill-Qwen-7B | 38.0 | 11.1 |
| OlympicCoder-7B | 40.9 | 10.6 |
| **OCR-Qwen-7B** | **48.5** | **16.3** |
| **OCR-Qwen-7B-Instruct** | **51.3** | **18.1** |
| | | |
| **Distilled 14B+ Models**| | |
| | | |
| R1-Distill-Qwen-14B | 51.3 | 17.6 |
| **OCR-Qwen-14B** | **57.7** | **22.6** |
| **OCR-Qwen-14B-Instruct**| **59.4** | **23.6** |
| | | |
| **Distilled 32B+ Models**| | |
| | | |
| Bespoke-Stratos-32B | 30.1 | 6.3 |
| OpenThinker-32B | 54.1 | 16.4 |
| R1-Distill-Qwen-32B | 58.1 | 18.3 |
| OlympicCoder-32B | 57.4 | 18.0 |
| **OCR-Qwen-32B** | **61.8** | **24.6** |
| **OCR-Qwen-32B-Instruct**| **61.7** | **24.4** |
## Reproducing our results
* [Models](https://huggingface.co/collections/nvidia/opencodereasoning-2-68168f37cd7c6beb1e3f92e7)
* [Dataset](https://huggingface.co/datasets/nvidia/OpenCodeReasoning)
* [Paper](https://arxiv.org/abs/2504.01943)
## How to use the models?
To run inference on coding problems:
````python
import transformers
import torch
model_id = "nvidia/OpenCodeReasoning-Nemotron-14B"
pipeline = transformers.pipeline(
"text-generation",
model=model_id,
model_kwargs={"torch_dtype": torch.bfloat16},
device_map="auto",
)
prompt = """You are a helpful and harmless assistant. You should think step-by-step before responding to the instruction below.
Please use python programming language only.
You must use ```python for just the final solution code block with the following format:
```python
# Your code here
```
{user}
"""
messages = [
{
"role": "user",
"content": prompt.format(user="Write a program to calculate the sum of the first $N$ fibonacci numbers")},
]
outputs = pipeline(
messages,
max_new_tokens=32768,
)
print(outputs[0]["generated_text"][-1]['content'])
````
## Citation
If you find the data useful, please cite:
```
@article{ahmad2025opencodereasoning,
title={OpenCodeReasoning: Advancing Data Distillation for Competitive Coding},
author={Wasi Uddin Ahmad, Sean Narenthiran, Somshubra Majumdar, Aleksander Ficek, Siddhartha Jain, Jocelyn Huang, Vahid Noroozi, Boris Ginsburg},
year={2025},
eprint={2504.01943},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2504.01943},
}
```
## Additional Information
## Model Architecture: <br>
Architecture Type: Dense decoder-only Transformer model
Network Architecture: Qwen-14B-Instruct
<br>
**This model was developed based on Qwen2.5-14B-Instruct and has 14B model parameters. <br>**
**OpenCodeReasoning-Nemotron-14B was developed based on Qwen2.5-14B-Instruct and has 14B model parameters. <br>**
## Input: <br>
**Input Type(s):** Text <br>
**Input Format(s):** String <br>
**Input Parameters:** One-Dimensional (1D) <br>
**Other Properties Related to Input:** Context length up to 32,768 tokens <br>
## Output: <br>
**Output Type(s):** Text <br>
**Output Format:** String <br>
**Output Parameters:** One-Dimensional (1D) <br>
**Other Properties Related to Output:** Context length up to 32,768 tokens <br>
Our AI models are designed and/or optimized to run on NVIDIA GPU-accelerated systems. By leveraging NVIDIAs hardware (e.g. GPU cores) and software frameworks (e.g., CUDA libraries), the model achieves faster training and inference times compared to CPU-only solutions. <br>
## Software Integration : <br>
* Runtime Engine: NeMo 2.3.0 <br>
* Recommended Hardware Microarchitecture Compatibility: <br>
NVIDIA Ampere <br>
NVIDIA Hopper <br>
* Preferred/Supported Operating System(s): Linux <br>
## Model Version(s):
1.0 (4/25/2025) <br>
OpenCodeReasoning-Nemotron-7B<br>
OpenCodeReasoning-Nemotron-14B<br>
OpenCodeReasoning-Nemotron-32B<br>
OpenCodeReasoning-Nemotron-32B-IOI<br>
# Training and Evaluation Datasets: <br>
## Training Dataset:
The training corpus for OpenCodeReasoning-Nemotron-14B is [OpenCodeReasoning](https://huggingface.co/datasets/nvidia/OpenCodeReasoning) dataset, which is composed of competitive programming questions and DeepSeek-R1 generated responses.
Data Collection Method: Hybrid: Automated, Human, Synthetic <br>
Labeling Method: Hybrid: Automated, Human, Synthetic <br>
Properties: 736k samples from OpenCodeReasoning (https://huggingface.co/datasets/nvidia/OpenCodeReasoning)
## Evaluation Dataset:
We used the datasets listed in the next section to evaluate OpenCodeReasoning-Nemotron-14B. <br>
**Data Collection Method: Hybrid: Automated, Human, Synthetic <br>**
**Labeling Method: Hybrid: Automated, Human, Synthetic <br>**
### License/Terms of Use: <br>
GOVERNING TERMS: Use of this model is governed by [Apache 2.0](https://huggingface.co/nvidia/OpenCode-Nemotron-2-14B/blob/main/LICENSE).
### Deployment Geography:
Global<br>
### Use Case: <br>
This model is intended for developers and researchers building LLMs. <br>
### Release Date: <br>
Huggingface [04/25/2025] via https://huggingface.co/nvidia/OpenCodeReasoning-Nemotron-7B/ <br>
## Reference(s):
[2504.01943] OpenCodeReasoning: Advancing Data Distillation for Competitive Coding
<br>
## Inference:
**Engine:** vLLM <br>
**Test Hardware** NVIDIA H100-80GB <br>
## Ethical Considerations:
NVIDIA believes Trustworthy AI is a shared responsibility and we have established policies and practices to enable development for a wide array of AI applications. When downloaded or used in accordance with our terms of service, developers should work with their internal model team to ensure this model meets requirements for the relevant industry and use case and addresses unforeseen product misuse.
Please report security vulnerabilities or NVIDIA AI Concerns here.