初始化项目,由ModelHub XC社区提供模型
Model: manvadariya1/Zynthos-1.2B-Instruct-GGUF Source: Original Platform
This commit is contained in:
38
.gitattributes
vendored
Normal file
38
.gitattributes
vendored
Normal file
@@ -0,0 +1,38 @@
|
|||||||
|
*.7z filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.arrow filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.bin filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.bz2 filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ckpt filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ftz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.gz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.h5 filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.joblib filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.lfs.* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.mlmodel filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.model filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.msgpack filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.npy filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.npz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.onnx filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.ot filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.parquet filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pb filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pickle filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pkl filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pt filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.pth filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.rar filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.safetensors filter=lfs diff=lfs merge=lfs -text
|
||||||
|
saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tar.* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tar filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tflite filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.tgz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.wasm filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.xz filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.zip filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*.zst filter=lfs diff=lfs merge=lfs -text
|
||||||
|
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Zynthos-1.2B-Instruct-Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Zynthos-1.2B-Instruct-Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
|
Zynthos-1.2B-Instruct-F16.gguf filter=lfs diff=lfs merge=lfs -text
|
||||||
80
README.md
Normal file
80
README.md
Normal file
@@ -0,0 +1,80 @@
|
|||||||
|
---
|
||||||
|
license: other
|
||||||
|
license_name: lfm-open-v1.0
|
||||||
|
license_link: https://huggingface.co/LiquidAI/LFM2.5-1.2B-Instruct/blob/main/LICENSE
|
||||||
|
base_model: LiquidAI/LFM2.5-1.2B-Instruct
|
||||||
|
model_creator: LiquidAI
|
||||||
|
model_name: Zynthos 1.2B Instruct
|
||||||
|
pipeline_tag: text-generation
|
||||||
|
quantized_by: manvadariya1
|
||||||
|
language:
|
||||||
|
- en
|
||||||
|
- zh
|
||||||
|
- fr
|
||||||
|
tags:
|
||||||
|
- text-generation
|
||||||
|
- gguf
|
||||||
|
- edge-ai
|
||||||
|
- on-device
|
||||||
|
- intent-router
|
||||||
|
- structured-outputs
|
||||||
|
- json-mode
|
||||||
|
- agent
|
||||||
|
---
|
||||||
|
|
||||||
|
# 🌌 Zynthos-1.2B-Instruct: The Edge AI Revolution
|
||||||
|
|
||||||
|
Zynthos-1.2B-Instruct represents a monumental paradigm shift in local, on-device intelligence. Moving entirely beyond the scaling limits and massive computational overhead of traditional Transformer models, Zynthos is a high-fidelity deployment lineage built upon Liquid AI’s revolutionary non-transformer sequential architecture (**LFM2.5-1.2B-Instruct**).
|
||||||
|
|
||||||
|
By redefining token processing logic from the ground up, Zynthos delivers unprecedented throughput, sub-millisecond execution loops, and infinitely scalable context efficiency—all within a microscopic hardware footprint.
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## ⚡ The Architectural Shift: Why Zynthos Changes Everything
|
||||||
|
|
||||||
|
Conventional small language models choke on memory bottlenecks and computational drain during long agent loops. **Zynthos-1.2B-Instruct** shatters these constraints, establishing a brand new class of localized ambient intelligence:
|
||||||
|
|
||||||
|
* **Sub-50ms Intelligent Routing:** Deployed instantly as a local "fast-lane" intent classifier to orchestrate multi-agent tasks before routing heavier workloads to deep reasoning engines.
|
||||||
|
* **Deterministic Structured Extraction:** Completely strips away conversational fluff to enforce flawless, schema-compliant JSON outputs and lightning-fast tool calls directly at the edge.
|
||||||
|
* **Flawless Infinite Scaling:** Leverages underlying non-transformer recurrent dynamics to process complex data arrays with virtually static memory allocations, saving critical hardware battery life.
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## 📊 Quantization & Performance Matrix
|
||||||
|
|
||||||
|
> ### ⭐ Execution Recommendation
|
||||||
|
> For professional deployments, local workflow automation, and multi-agent system pipelines, **`Zynthos-1.2B-Instruct-F16.gguf` is the highly recommended variant**. It preserves 100% of the raw, uncompressed model tensors, guaranteeing maximum semantic reasoning, perfect tool-calling accuracy, and zero quantization loss.
|
||||||
|
|
||||||
|
| File Artifact | Precision Bit-Weight | File Size | Memory Footprint | Deployment Classification |
|
||||||
|
| :--- | :--- | :--- | :--- | :--- |
|
||||||
|
| **`Zynthos-1.2B-Instruct-F16.gguf`** | **Full FP16 Master** | **~2.4 GB** | **8 GB RAM** | **🏆 Recommended Tier: Maximum Precision & Uncompromised Routing** |
|
||||||
|
| `Zynthos-1.2B-Instruct-Q8_0.gguf` | 8-bit Standard | ~1.2 GB | 4 GB RAM | Balanced Tier: Premium RAG parsing & local document scanning |
|
||||||
|
| `Zynthos-1.2B-Instruct-Q4_K_M.gguf`| 4-bit Medium | ~750 MB | 2 GB RAM | Ultra-Fast Tier: Extreme edge execution & restricted mobile hardware |
|
||||||
|
|
||||||
|
---
|
||||||
|
|
||||||
|
## 🛠️ High-Speed Integration Blueprint
|
||||||
|
|
||||||
|
### 1. Instant Desktop Setup (LM Studio / Ollama)
|
||||||
|
1. Navigate to the **Files and versions** tab and download the recommended **`Zynthos-1.2B-Instruct-F16.gguf`** file.
|
||||||
|
2. Drop the file directory path straight into your local workspace.
|
||||||
|
3. Select the model within your UI, maximize your **GPU Offload** toggles, and experience localized generation speeds that feel instantaneous.
|
||||||
|
|
||||||
|
### 2. Enterprise Workflow Orchestration (`llama-cpp-python`)
|
||||||
|
Build local agent loops, background intent filters, or rapid JSON parsers with this streamlined script:
|
||||||
|
|
||||||
|
```python
|
||||||
|
from llama_cpp import Llama
|
||||||
|
|
||||||
|
# Instantiate the recommended uncompressed master file for flawless execution
|
||||||
|
llm = Llama(
|
||||||
|
model_path="./Zynthos-1.2B-Instruct-F16.gguf",
|
||||||
|
n_ctx=4096,
|
||||||
|
n_gpu_layers=-1 # Completely offload all layer calculations to your hardware GPU
|
||||||
|
)
|
||||||
|
|
||||||
|
# Optimized syntax structure for Instruct execution
|
||||||
|
prompt = "<|im_start|>user\nAnalyze this payload and return only the target intent key: [JSON], [SQL], or [TEXT]. Payload: 'SELECT * FROM infrastructure_metrics WHERE cpu > 90;'<|im_end|>\n<|im_start|>assistant\n"
|
||||||
|
|
||||||
|
output = llm(prompt, max_tokens=16, stop=["<|im_end|>"])
|
||||||
|
print(f"⚡ Routed Intent: {output['choices'][0]['text'].strip()}")
|
||||||
3
Zynthos-1.2B-Instruct-F16.gguf
Normal file
3
Zynthos-1.2B-Instruct-F16.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:791932034400f9ecd3318d619d3ee1870ba8cd984e29db54d2e08e4fc6adeb77
|
||||||
|
size 2343326144
|
||||||
3
Zynthos-1.2B-Instruct-Q4_K_M.gguf
Normal file
3
Zynthos-1.2B-Instruct-Q4_K_M.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:9be6c0d1b7b85198d49a12cb43aa99d04935e684ac2ff324778e8e0ba5388407
|
||||||
|
size 730894784
|
||||||
3
Zynthos-1.2B-Instruct-Q8_0.gguf
Normal file
3
Zynthos-1.2B-Instruct-Q8_0.gguf
Normal file
@@ -0,0 +1,3 @@
|
|||||||
|
version https://git-lfs.github.com/spec/v1
|
||||||
|
oid sha256:c0d27fa540b5257b9a7c0d830bb17246c0822ecc068c4ec8ac7f909bd184759e
|
||||||
|
size 1246253504
|
||||||
Reference in New Issue
Block a user