初始化项目,由ModelHub XC社区提供模型

Model: ForSureTesterSim/QwenR1-7B-Breadcrumbs-TIES
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-24 04:56:11 +08:00
commit 487fd8b094
12 changed files with 303838 additions and 0 deletions

45
README.md Normal file
View File

@@ -0,0 +1,45 @@
---
base_model:
- POLARIS-Project/Polaris-7B-Preview
- deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
- nvidia/AceReason-Nemotron-7B
library_name: transformers
tags:
- mergekit
- merge
---
# merged-breadcrumbs
This is a merge of pre-trained language models created using [mergekit](https://github.com/cg123/mergekit).
## Merge Details
### Merge Method
This model was merged using the [Model Breadcrumbs with TIES](https://arxiv.org/abs/2312.06795) merge method using [deepseek-ai/DeepSeek-R1-Distill-Qwen-7B](https://huggingface.co/deepseek-ai/DeepSeek-R1-Distill-Qwen-7B) as a base.
### Models Merged
The following models were included in the merge:
* [POLARIS-Project/Polaris-7B-Preview](https://huggingface.co/POLARIS-Project/Polaris-7B-Preview)
* ./padded-maestro
* [nvidia/AceReason-Nemotron-7B](https://huggingface.co/nvidia/AceReason-Nemotron-7B)
### Configuration
The following YAML configuration was used to produce this model:
```yaml
merge_method: breadcrumbs_ties
base_model: deepseek-ai/DeepSeek-R1-Distill-Qwen-7B
parameters:
density: 0.15 # Retain top 15% (Trims bottom 85% SFT noise)
gamma: 0.99 # Trim the top 1% (Destroys extreme RL outliers)
weight: 1.0
models:
- model: nvidia/AceReason-Nemotron-7B
- model: ./padded-maestro
- model: POLARIS-Project/Polaris-7B-Preview
dtype: bfloat16
```