初始化项目,由ModelHub XC社区提供模型

Model: saidutta69/Ghosty-7B
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-07-14 20:39:10 +08:00
commit e21e9a4770
11 changed files with 307 additions and 0 deletions

39
.gitattributes vendored Normal file
View File

@@ -0,0 +1,39 @@
*.7z filter=lfs diff=lfs merge=lfs -text
*.arrow filter=lfs diff=lfs merge=lfs -text
*.bin filter=lfs diff=lfs merge=lfs -text
*.bz2 filter=lfs diff=lfs merge=lfs -text
*.ckpt filter=lfs diff=lfs merge=lfs -text
*.ftz filter=lfs diff=lfs merge=lfs -text
*.gz filter=lfs diff=lfs merge=lfs -text
*.h5 filter=lfs diff=lfs merge=lfs -text
*.joblib filter=lfs diff=lfs merge=lfs -text
*.lfs.* filter=lfs diff=lfs merge=lfs -text
*.mlmodel filter=lfs diff=lfs merge=lfs -text
*.model filter=lfs diff=lfs merge=lfs -text
*.msgpack filter=lfs diff=lfs merge=lfs -text
*.npy filter=lfs diff=lfs merge=lfs -text
*.npz filter=lfs diff=lfs merge=lfs -text
*.onnx filter=lfs diff=lfs merge=lfs -text
*.ot filter=lfs diff=lfs merge=lfs -text
*.parquet filter=lfs diff=lfs merge=lfs -text
*.pb filter=lfs diff=lfs merge=lfs -text
*.pickle filter=lfs diff=lfs merge=lfs -text
*.pkl filter=lfs diff=lfs merge=lfs -text
*.pt filter=lfs diff=lfs merge=lfs -text
*.pth filter=lfs diff=lfs merge=lfs -text
*.rar filter=lfs diff=lfs merge=lfs -text
*.safetensors filter=lfs diff=lfs merge=lfs -text
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
*.tar.* filter=lfs diff=lfs merge=lfs -text
*.tar filter=lfs diff=lfs merge=lfs -text
*.tflite filter=lfs diff=lfs merge=lfs -text
*.tgz filter=lfs diff=lfs merge=lfs -text
*.wasm filter=lfs diff=lfs merge=lfs -text
*.xz filter=lfs diff=lfs merge=lfs -text
*.zip filter=lfs diff=lfs merge=lfs -text
*.zst filter=lfs diff=lfs merge=lfs -text
*tfevents* filter=lfs diff=lfs merge=lfs -text
tokenizer.json filter=lfs diff=lfs merge=lfs -text
ghosty-7b-q8_0.gguf filter=lfs diff=lfs merge=lfs -text
ghosty-7b-q5_k_m.gguf filter=lfs diff=lfs merge=lfs -text
ghosty-7b-q4_k_m.gguf filter=lfs diff=lfs merge=lfs -text

95
README.md Normal file
View File

@@ -0,0 +1,95 @@
---
license: apache-2.0
language:
- en
- hi
tags:
- qwen2.5
- paranormal
- ghost-hunting
- hinglish
- investigator
- 7b
- finetuned
pipeline_tag: text-generation
widget:
- text: "Maine ek raat ko ghar me ajeeb awaaz suni. Kya ho sakta hai?"
- text: "Kya bhoot sach me hote hain?"
- text: "Mere ghar me raat ko plates apne aap hilne lagti hain. Kya karun?"
---
# Ghosty-7B 👻
A **Paranormal Investigator LLM** — fine-tuned from Qwen2.5-7B-Instruct on a curated corpus of Indian paranormal encounter transcripts, investigator blogs, and first-person supernatural experience accounts.
Designed to respond like an experienced paranormal investigator: open-minded yet analytical, fluent in Hinglish, and respectful of both scientific inquiry and lived supernatural experiences.
## Model Capabilities
- 🎙️ **Hinglish-first**: Responds naturally in Hindi-English code-mixing, matching how Indian paranormal communities communicate
- 🔍 **Investigator persona**: Balances skepticism with belief — investigates claims seriously without dismissing or fear-mongering
- 🏠 **Indian context**: Understands local ghost lore (bhoot-pret, tantra-mantra, aatma), regional beliefs, and culturally relevant advice
- 📖 **Storytelling**: Can narrate Indian folklore and paranormal encounter stories in an engaging first-person investigator style
## Example Interaction
```
User: Maine ek raat ko ghar me ajeeb awaaz suni. Kya ho sakta hai?
Ghosty: Aise ajeeb awaaze ki koi common reasons hote hain:
1. Electrical interference (like faulty appliances)
2. Humming from AC/ventilator
3. Koi animal (katra, bat)
4. Wind noise through windows
...
User: Kya bhoot sach me hote hain?
Ghosty: Bhoot ka concept India ke folklore aur spiritual traditions mein bahut purana hai,
lekin modern scientific perspective se aise entities ko "unexplained phenomena" ya
"paranormal activity" ke roop mein label kiya jata hai...
```
## Training Data
Trained on a diverse collection of:
- **Paranormal investigator blogs & transcripts** (Creative Commons licensed, used with permission)
- **Indian ghost encounter narratives** from verified public sources
- **First-person supernatural experience accounts** from Indian online communities
- **Hinglish conversational data** spanning urban and rural paranormal reporting styles
Dataset: ~3,000 SFT question-answer pairs in Qwen chat format.
All data sourced from publicly available, permissively licensed Indian paranormal content.
## Training Details
| Parameter | Value |
|-----------|-------|
| Base Model | Qwen2.5-7B-Instruct (abliterated variant) |
| Fine-tuning | Full parameter FT (not LoRA) |
| Hardware | NVIDIA H100 80GB |
| Precision | BF16 |
| Attention | SDPA (PyTorch built-in) |
| Epochs | 5 |
| Batch Size | 8 (effective 16 with grad accum) |
| Learning Rate | 2e-5 (cosine, warmup 185 steps) |
| Max Length | 2048 tokens |
| Gradient Checkpointing | Yes |
## Intended Use
Ghosty-7B is designed for:
- **Paranormal investigation RP/chatbots**
- **Indian folklore & ghost story generation**
- **First-responder paranormal advice systems**
- **Educational demonstrations of Indian supernatural beliefs**
## Limitations
- May occasionally hallucinate specific case details (standard for LLMs)
- Best at Hinglish — performance in pure English or pure Hindi may vary
- Not a replacement for professional mental health or law enforcement advice
- Trained on publicly shared experiences — individual accounts may not be verified
## License
Apache 2.0

54
chat_template.jinja Normal file
View File

@@ -0,0 +1,54 @@
{%- if tools %}
{{- '<|im_start|>system\n' }}
{%- if messages[0]['role'] == 'system' %}
{{- messages[0]['content'] }}
{%- else %}
{{- 'You are Qwen, created by Alibaba Cloud. You are a helpful assistant.' }}
{%- endif %}
{{- "\n\n# Tools\n\nYou may call one or more functions to assist with the user query.\n\nYou are provided with function signatures within <tools></tools> XML tags:\n<tools>" }}
{%- for tool in tools %}
{{- "\n" }}
{{- tool | tojson }}
{%- endfor %}
{{- "\n</tools>\n\nFor each function call, return a json object with function name and arguments within <tool_call></tool_call> XML tags:\n<tool_call>\n{\"name\": <function-name>, \"arguments\": <args-json-object>}\n</tool_call><|im_end|>\n" }}
{%- else %}
{%- if messages[0]['role'] == 'system' %}
{{- '<|im_start|>system\n' + messages[0]['content'] + '<|im_end|>\n' }}
{%- else %}
{{- '<|im_start|>system\nYou are Qwen, created by Alibaba Cloud. You are a helpful assistant.<|im_end|>\n' }}
{%- endif %}
{%- endif %}
{%- for message in messages %}
{%- if (message.role == "user") or (message.role == "system" and not loop.first) or (message.role == "assistant" and not message.tool_calls) %}
{{- '<|im_start|>' + message.role + '\n' + message.content + '<|im_end|>' + '\n' }}
{%- elif message.role == "assistant" %}
{{- '<|im_start|>' + message.role }}
{%- if message.content %}
{{- '\n' + message.content }}
{%- endif %}
{%- for tool_call in message.tool_calls %}
{%- if tool_call.function is defined %}
{%- set tool_call = tool_call.function %}
{%- endif %}
{{- '\n<tool_call>\n{"name": "' }}
{{- tool_call.name }}
{{- '", "arguments": ' }}
{{- tool_call.arguments | tojson }}
{{- '}\n</tool_call>' }}
{%- endfor %}
{{- '<|im_end|>\n' }}
{%- elif message.role == "tool" %}
{%- if (loop.index0 == 0) or (messages[loop.index0 - 1].role != "tool") %}
{{- '<|im_start|>user' }}
{%- endif %}
{{- '\n<tool_response>\n' }}
{{- message.content }}
{{- '\n</tool_response>' }}
{%- if loop.last or (messages[loop.index0 + 1].role != "tool") %}
{{- '<|im_end|>\n' }}
{%- endif %}
{%- endif %}
{%- endfor %}
{%- if add_generation_prompt %}
{{- '<|im_start|>assistant\n' }}
{%- endif %}

61
config.json Normal file
View File

@@ -0,0 +1,61 @@
{
"architectures": [
"Qwen2ForCausalLM"
],
"attention_dropout": 0.0,
"bos_token_id": null,
"dtype": "bfloat16",
"eos_token_id": 151645,
"hidden_act": "silu",
"hidden_size": 3584,
"initializer_range": 0.02,
"intermediate_size": 18944,
"layer_types": [
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention",
"full_attention"
],
"max_position_embeddings": 32768,
"max_window_layers": 28,
"model_type": "qwen2",
"num_attention_heads": 28,
"num_hidden_layers": 28,
"num_key_value_heads": 4,
"pad_token_id": 151643,
"rms_norm_eps": 1e-06,
"rope_parameters": {
"rope_theta": 1000000.0,
"rope_type": "default"
},
"sliding_window": null,
"tie_word_embeddings": false,
"transformers_version": "5.13.0",
"use_cache": false,
"use_sliding_window": false,
"vocab_size": 152064
}

13
generation_config.json Normal file
View File

@@ -0,0 +1,13 @@
{
"do_sample": true,
"eos_token_id": [
151645,
151643
],
"pad_token_id": 151643,
"repetition_penalty": 1.05,
"temperature": 0.7,
"top_k": 20,
"top_p": 0.8,
"transformers_version": "5.13.0"
}

3
ghosty-7b-q4_k_m.gguf Normal file
View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:02217ab8a8c2cdbfe7fdab41f75ce6eac7062639aade4d123d329fc9df8e964c
size 4683073696

3
ghosty-7b-q5_k_m.gguf Normal file
View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:93a150971848b98c7fdbdbf171983b43032b09f14a072e787fde7085953b1b28
size 5444831392

3
ghosty-7b-q8_0.gguf Normal file
View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:acf960e879cec5e9c0a249471af8aecce01b0c44b26c9bac5fa9c1122890da2e
size 8098525344

3
model.safetensors Normal file
View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:2491c91ab7817fee3ca772645b1be25ae24a86c30b0a775997085fd884e10df8
size 15231272152

3
tokenizer.json Normal file
View File

@@ -0,0 +1,3 @@
version https://git-lfs.github.com/spec/v1
oid sha256:3fd169731d2cbde95e10bf356d66d5997fd885dd8dbb6fb4684da3f23b2585d8
size 11421892

30
tokenizer_config.json Normal file
View File

@@ -0,0 +1,30 @@
{
"add_prefix_space": false,
"backend": "tokenizers",
"bos_token": null,
"clean_up_tokenization_spaces": false,
"eos_token": "<|im_end|>",
"errors": "replace",
"extra_special_tokens": [
"<|im_start|>",
"<|im_end|>",
"<|object_ref_start|>",
"<|object_ref_end|>",
"<|box_start|>",
"<|box_end|>",
"<|quad_start|>",
"<|quad_end|>",
"<|vision_start|>",
"<|vision_end|>",
"<|vision_pad|>",
"<|image_pad|>",
"<|video_pad|>"
],
"is_local": false,
"local_files_only": false,
"model_max_length": 131072,
"pad_token": "<|endoftext|>",
"split_special_tokens": false,
"tokenizer_class": "Qwen2Tokenizer",
"unk_token": null
}