Files
gemma4-cnc-ll-GGUF/README.md
ModelHub XC 8869760abf 初始化项目,由ModelHub XC社区提供模型
Model: Pauldyu57/gemma4-cnc-ll-GGUF
Source: Original Platform
2026-08-29 09:44:17 +08:00

1.6 KiB

license, base_model, library_name, tags, pipeline_tag
license base_model library_name tags pipeline_tag
gemma google/gemma-3-4b-it gguf
gguf
gemma
gemma-3
cnc
manufacturing
work-order
ll-machinery
fine-tuned
ollama
text-generation

gemma4-cnc-ll (GGUF) — Two-stage Fine-tune

Two-stage QLoRA fine-tune of google/gemma-3-4b-it:

Stage 1 — Work-order knowledge CNC work-order database schema, BOM categories, pricing models, Neo4j/SQL queries for 永詮機械 (Yong Chuan Machinery).

Stage 2 — L&L 永詮 specialization (anti-dilution training) 14 machine models (LLA/LLB/LLS/LL/LFM/LFS/LS-C/LS-S/LS-U/LD/LC/TA/MA/LA/A), 6 application industries, company background. Trained with 5 anti-dilution mechanisms including critical-facts validation.

Quick start (Ollama)

ollama run hf.co/Pauldyu57/gemma4-cnc-ll-GGUF:Q4_K_M

Other tags: Q5_K_M, Q8_0, F16.

If this repo is private, link your Ollama SSH key first:

cat ~/.ollama/id_ed25519.pub
# → paste into https://huggingface.co/settings/keys

Files

Quant Approx size Notes
F16 ~7.5 GB Reference
Q8_0 ~4.3 GB Near-lossless
Q5_K_M ~3.0 GB Good quality / size balance
Q4_K_M ~2.5 GB Default for CPU / small GPUs

Base model

google/gemma-3-4b-it — usage subject to the Gemma Terms of Use.