--- language: - en license: llama3 tags: - llama - coding - reasoning - logic - gguf base_model: meta-llama/Meta-Llama-3-8B-Instruct pipeline_tag: text-generation ---
# ๐Ÿ’  VEDA-8B-v1-COGNITIVE ### *Fine-tuned to think. Not just predict.* ![](https://img.shields.io/badge/Base-Llama--3--8B-blue?style=for-the-badge) ![](https://img.shields.io/badge/Method-LoRA-green?style=for-the-badge) ![](https://img.shields.io/badge/Format-GGUF-orange?style=for-the-badge) ![](https://img.shields.io/badge/Context-8192_tokens-purple?style=for-the-badge) **By Vibhansh Bamnia ยท 18 ยท Independent AI Developer ยท Rajasthan, India** [๐Ÿš€ Try Demo](https://huggingface.co/spaces/vibhansh/Veda-Chat) ยท [๐Ÿ“ฅ Download Model](#-quantization)
--- ## โšก Why Veda-8B? > Most 8B models pattern-match. Veda-8B **reasons.** LoRA fine-tuned on open-source reasoning datasets โ€” built specifically to solve where standard Llama-3 fails. | Capability | Standard Llama-3-8B | Veda-8B-v1 | |---|:---:|:---:| | Relational Chain Reasoning | โš ๏ธ Often fails | โœ… Reliable | | Multi-hop Logic | โš ๏ธ Inconsistent | โœ… Strong | | Code Hallucination | โŒ Common | โœ… Reduced | | Chain-of-Thought | โš ๏ธ Partial | โœ… Complete | --- ## ๐Ÿง  See It In Action
๐Ÿ”— Relational Reasoning **Prompt:** `A is B's father. C is A's sister. D is C's son. What is D's relation to B?` **Veda-8B:** ``` A is B's father โ†’ C is B's aunt โ†’ D is C's son โˆด D is B's cousin. ```
๐Ÿ”ฃ Logic Syllogism **Prompt:** `All Bloops are Razzies. All Razzies are Lazzies. Are all Bloops Lazzies?` **Veda-8B:** ``` Bloops โІ Razzies โІ Lazzies โˆด Yes โ€” by transitivity, all Bloops are definitely Lazzies. ```
๐Ÿ’ป Code (No Hallucination) **Prompt:** `Second largest unique number without sort()` **Veda-8B:** ```python def second_largest(nums): unique = set(nums) if len(unique) < 2: return None first = second = float('-inf') for n in unique: if n > first: second, first = first, n elif n > second: second = n return second ```
--- ## ๐Ÿ’Ž Quantization | Version | Size | RAM | Use | |:---:|:---:|:---:|:---| | ๐Ÿฅ‡ Q4_K_M | 4.92 GB | 8 GB | Local dev, edge | | ๐Ÿฅˆ Q8_0 | 8.54 GB | 16 GB | Production | | ๐Ÿ”ฌ F16 | 16 GB+ | 32 GB | Research | --- ## ๐Ÿ’ป Quick Start ```python import llama_cpp llm = llama_cpp.Llama(model_path="./Veda-8B-v1-Q4_K_M.gguf", n_ctx=8192, n_threads=4) response = llm("A is taller than B. B is taller than C. Is A taller than C?", max_tokens=512, temperature=0.1) ``` --- ## โš ๏ธ Limitations - Inherits Llama-3-8B base limitations - Not for real-time / factual queries - Verify math-heavy outputs - Multilingual not tested ---
*No institution. No GPU cluster. No team.* *Just a focused fine-tune built to make small models reason correctly.* **Veda AI Labs** ยท [HuggingFace](https://huggingface.co/vibhansh)