---
language:
- en
license: llama3
tags:
- llama
- coding
- reasoning
- logic
- gguf
base_model: meta-llama/Meta-Llama-3-8B-Instruct
pipeline_tag: text-generation
---
# ๐ VEDA-8B-v1-COGNITIVE
### *Fine-tuned to think. Not just predict.*




**By Vibhansh Bamnia ยท 18 ยท Independent AI Developer ยท Rajasthan, India**
[๐ Try Demo](https://huggingface.co/spaces/vibhansh/Veda-Chat) ยท [๐ฅ Download Model](#-quantization)
---
## โก Why Veda-8B?
> Most 8B models pattern-match. Veda-8B **reasons.**
LoRA fine-tuned on open-source reasoning datasets โ built specifically to solve where standard Llama-3 fails.
| Capability | Standard Llama-3-8B | Veda-8B-v1 |
|---|:---:|:---:|
| Relational Chain Reasoning | โ ๏ธ Often fails | โ
Reliable |
| Multi-hop Logic | โ ๏ธ Inconsistent | โ
Strong |
| Code Hallucination | โ Common | โ
Reduced |
| Chain-of-Thought | โ ๏ธ Partial | โ
Complete |
---
## ๐ง See It In Action
๐ Relational Reasoning
**Prompt:** `A is B's father. C is A's sister. D is C's son. What is D's relation to B?`
**Veda-8B:**
```
A is B's father โ C is B's aunt โ D is C's son
โด D is B's cousin.
```
๐ฃ Logic Syllogism
**Prompt:** `All Bloops are Razzies. All Razzies are Lazzies. Are all Bloops Lazzies?`
**Veda-8B:**
```
Bloops โ Razzies โ Lazzies
โด Yes โ by transitivity, all Bloops are definitely Lazzies.
```
๐ป Code (No Hallucination)
**Prompt:** `Second largest unique number without sort()`
**Veda-8B:**
```python
def second_largest(nums):
unique = set(nums)
if len(unique) < 2:
return None
first = second = float('-inf')
for n in unique:
if n > first:
second, first = first, n
elif n > second:
second = n
return second
```
---
## ๐ Quantization
| Version | Size | RAM | Use |
|:---:|:---:|:---:|:---|
| ๐ฅ Q4_K_M | 4.92 GB | 8 GB | Local dev, edge |
| ๐ฅ Q8_0 | 8.54 GB | 16 GB | Production |
| ๐ฌ F16 | 16 GB+ | 32 GB | Research |
---
## ๐ป Quick Start
```python
import llama_cpp
llm = llama_cpp.Llama(model_path="./Veda-8B-v1-Q4_K_M.gguf", n_ctx=8192, n_threads=4)
response = llm("A is taller than B. B is taller than C. Is A taller than C?",
max_tokens=512, temperature=0.1)
```
---
## โ ๏ธ Limitations
- Inherits Llama-3-8B base limitations
- Not for real-time / factual queries
- Verify math-heavy outputs
- Multilingual not tested
---
*No institution. No GPU cluster. No team.*
*Just a focused fine-tune built to make small models reason correctly.*
**Veda AI Labs** ยท [HuggingFace](https://huggingface.co/vibhansh)