187 lines
7.1 KiB
Markdown
187 lines
7.1 KiB
Markdown
---
|
|
license: apache-2.0
|
|
library_name: transformers
|
|
pipeline_tag: text-generation
|
|
language:
|
|
- ja
|
|
- en
|
|
base_model:
|
|
- nri-ai/Qwen3-14B-Ja-Fin-CPT
|
|
base_model_relation: finetune
|
|
tags:
|
|
- finance
|
|
- japanese
|
|
- reasoning
|
|
- thinking
|
|
- sft
|
|
datasets:
|
|
- nri-ai/nri-fin-reasoning
|
|
---
|
|
|
|
# Qwen3-14B-Ja-Fin-Thinking
|
|
|
|
<div align="center" style="line-height: 1;">
|
|
<a href="https://huggingface.co/nri-ai" target="_blank" style="margin: 2px;">
|
|
<img alt="Hugging Face" src="https://img.shields.io/badge/%F0%9F%A4%97%20Hugging%20Face-NRI--AI-005bac?color=005bac&logoColor=white" style="display: inline-block; vertical-align: middle;"/>
|
|
</a>
|
|
<a href="https://huggingface.co/nri-ai/Qwen3-14B-Ja-Fin-Thinking/blob/main/docs/README.ja.md" style="margin: 2px;">
|
|
<img alt="Japanese" src="https://img.shields.io/badge/%F0%9F%87%AF%F0%9F%87%B5%20%E6%97%A5%E6%9C%AC%E8%AA%9E-README-005bac?color=005bac&logoColor=white" style="display: inline-block; vertical-align: middle;"/>
|
|
</a>
|
|
</div>
|
|
<div align="center" style="line-height: 1;">
|
|
<a href="https://www.anlp.jp/proceedings/annual_meeting/2026/pdf_dir/C7-2.pdf" target="_blank" style="margin: 2px;">
|
|
<img alt="NLP2026" src="https://img.shields.io/badge/%F0%9F%93%9D%20NLP2026-Paper-005bac?color=005bac&logoColor=white" style="display: inline-block; vertical-align: middle;"/>
|
|
</a>
|
|
<a href="https://arxiv.org/abs/2603.01353" target="_blank" style="margin: 2px;">
|
|
<img alt="arXiv" src="https://img.shields.io/badge/%F0%9F%93%9D%20arXiv-Paper-b31b1b?color=b31b1b&logoColor=white" style="display: inline-block; vertical-align: middle;"/>
|
|
</a>
|
|
<a href="https://huggingface.co/datasets/nri-ai/nri-fin-reasoning" target="_blank" style="margin: 2px;">
|
|
<img alt="Dataset" src="https://img.shields.io/badge/%F0%9F%97%82%EF%B8%8F%20Dataset-nri--fin--reasoning-005bac?color=005bac&logoColor=white" style="display: inline-block; vertical-align: middle;"/>
|
|
</a>
|
|
</div>
|
|
<div align="center" style="line-height: 1;">
|
|
<a href="https://www.apache.org/licenses/LICENSE-2.0" style="margin: 2px;">
|
|
<img alt="License" src="https://img.shields.io/badge/License-Apache_2.0-f5de53?color=f5de53" style="display: inline-block; vertical-align: middle;"/>
|
|
</a>
|
|
</div>
|
|
|
|
A Japanese financial domain reasoning model, built through supervised fine-tuning of [Qwen3-14B-Ja-Fin-CPT](https://huggingface.co/nri-ai/Qwen3-14B-Ja-Fin-CPT).
|
|
|
|
## Model Overview
|
|
|
|
Trained to provide high-quality responses with explicit reasoning traces for Japanese financial domain tasks.
|
|
|
|
- **Base Model**: [Qwen3-14B-Ja-Fin-CPT](https://huggingface.co/nri-ai/Qwen3-14B-Ja-Fin-CPT)
|
|
- **Training Stage**: Supervised Fine-Tuning (SFT)
|
|
- **Domain**: Japanese Finance
|
|
- **Language**: Japanese, English
|
|
|
|
## Benchmark Results
|
|
|
|
### japanese-lm-fin-harness
|
|
|
|
| Model | Avg. | chabsa | cma | cpa | fp2 | ss1 |
|
|
|-------|:----:|:------:|:---:|:---:|:---:|:---:|
|
|
| Qwen3-14B (official) | 71.04 | **91.96** | **93.26** | **49.37** | 53.37 | 67.22 |
|
|
| **Qwen3-14B-Ja-Fin-Thinking (Ours)** | **71.78** | 91.62 | 91.45 | 48.59 | **60.00** | 67.27 |
|
|
|
|
### pfmt-bench-fin-ja
|
|
|
|
| Model | Avg. | turn1 | turn2 |
|
|
|-------|:----:|:-----:|:-----:|
|
|
| Qwen3-14B (official) | 8.104 | 8.211 | 7.997 |
|
|
| **Qwen3-14B-Ja-Fin-Thinking (Ours)** | **8.455** | **8.514** | **8.395** |
|
|
|
|
## Training
|
|
|
|
### Supervised Fine-Tuning
|
|
|
|
Fine-tuned on our synthetic instruction dataset with reasoning traces:
|
|
|
|
- **Dataset**: [nri-fin-reasoning](https://huggingface.co/datasets/nri-ai/nri-fin-reasoning) + supplementary data
|
|
- **Total samples**: ~1.44M
|
|
- **Total tokens**: ~9.5B
|
|
- **Epochs**: 2
|
|
|
|
**Training Infrastructure:**
|
|
- Hardware: AWS p5en.48xlarge (NVIDIA H200 Tensor Core GPU x 8)
|
|
- Training time: ~240 hours
|
|
|
|
## Usage
|
|
|
|
```python
|
|
from transformers import AutoModelForCausalLM, AutoTokenizer
|
|
|
|
model_name = "nri-ai/Qwen3-14B-Ja-Fin-Thinking"
|
|
|
|
tokenizer = AutoTokenizer.from_pretrained(model_name)
|
|
model = AutoModelForCausalLM.from_pretrained(
|
|
model_name,
|
|
torch_dtype="auto",
|
|
device_map="auto"
|
|
)
|
|
|
|
messages = [
|
|
{"role": "user", "content": "分散投資のメリットとデメリットを説明してください。"}
|
|
]
|
|
|
|
text = tokenizer.apply_chat_template(
|
|
messages,
|
|
tokenize=False,
|
|
add_generation_prompt=True
|
|
)
|
|
|
|
inputs = tokenizer([text], return_tensors="pt").to(model.device)
|
|
outputs = model.generate(**inputs, max_new_tokens=8192)
|
|
|
|
response = tokenizer.decode(outputs[0][inputs.input_ids.shape[-1]:], skip_special_tokens=True)
|
|
print(response)
|
|
```
|
|
|
|
## Intended Use
|
|
|
|
### Primary Use Cases
|
|
|
|
- Financial question answering in Japanese
|
|
- Financial document analysis and summarization
|
|
- Financial reasoning and calculation tasks
|
|
- Multi-turn financial advisory conversations
|
|
|
|
### Out-of-Scope Uses
|
|
|
|
- Production deployment without additional safety evaluation
|
|
- Professional financial advice (this is a research model)
|
|
- Non-financial domain applications
|
|
|
|
## Limitations
|
|
|
|
- **Domain specificity**: Optimized for Japanese financial domain; performance on other domains may vary
|
|
- **Synthetic training data**: May contain hallucinations despite quality filtering
|
|
- **Language coverage**: Primarily Japanese and English
|
|
|
|
## Ethical Considerations
|
|
|
|
- Financial information generated by this model should not be used as professional financial advice without review by qualified experts
|
|
- Users should verify important financial information against authoritative sources and professional guidance before making decisions
|
|
- The model may reflect biases present in training data
|
|
|
|
## License
|
|
|
|
This model is released under the Apache 2.0 license.
|
|
|
|
## Privacy Notice
|
|
|
|
For details on how personal information is handled, please see the [Privacy Notice](https://huggingface.co/nri-ai/Qwen3-14B-Ja-Fin-Thinking/blob/main/docs/PRIVACY_NOTICE.md) ([日本語](https://huggingface.co/nri-ai/Qwen3-14B-Ja-Fin-Thinking/blob/main/docs/PRIVACY_NOTICE.ja.md)).
|
|
|
|
## Citation
|
|
|
|
```bibtex
|
|
@inproceedings{okochiDomainSpecificLLM2026,
|
|
author = {大河内 悠磨 and Sim, Fabio Milentiansen and 岡田 智靖},
|
|
title = {ドメイン特化LLMの推論能力向上を目的とした合成指示データセットの構築と金融ドメインにおける評価},
|
|
booktitle = {言語処理学会第32回年次大会 (NLP2026) },
|
|
year = {2026},
|
|
month = mar,
|
|
address = {Utsunomiya, Tochigi, Japan},
|
|
publisher = {言語処理学会},
|
|
note = {Paper ID: C7-2},
|
|
url = {https://www.anlp.jp/proceedings/annual_meeting/2026/pdf_dir/C7-2.pdf}
|
|
}
|
|
```
|
|
|
|
```bibtex
|
|
@misc{okochi2026constructingsyntheticinstructiondatasets,
|
|
title = {Constructing Synthetic Instruction Datasets for Improving Reasoning in Domain-Specific LLMs: A Case Study in the Japanese Financial Domain},
|
|
author = {Yuma Okochi and Fabio Milentiansen Sim and Tomoyasu Okada},
|
|
year = {2026},
|
|
eprint = {2603.01353},
|
|
archivePrefix = {arXiv},
|
|
primaryClass = {cs.LG},
|
|
url = {https://arxiv.org/abs/2603.01353}
|
|
}
|
|
```
|
|
|
|
## Acknowledgments
|
|
|
|
This model was developed with the support of the "GENIAC (Generative AI Accelerator Challenge)" project, implemented by the Ministry of Economy, Trade and Industry (METI) and the New Energy and Industrial Technology Development Organization (NEDO), with the aim of strengthening Japan's development capabilities in generative AI.
|