初始化项目,由ModelHub XC社区提供模型

Model: eaddario/Hammer2.1-7b-GGUF
Source: Original Platform
This commit is contained in:
ModelHub XC
2026-08-25 11:51:19 +08:00
commit 8b81781d2d
110 changed files with 15439 additions and 0 deletions

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 29 key-value pairs and 339 tensors from ./Hammer2.1-7b-F16.gguf (version GGUF V3 (latest))
Final result: 58.9333 +/- 1.7976
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 6229.12 ms
llama_perf_context_print: prompt eval time = 102811.05 ms / 35972 tokens ( 2.86 ms per token, 349.88 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 103951.36 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 29 key-value pairs and 339 tensors from ./Hammer2.1-7b-F16.gguf (version GGUF V3 (latest))
750 72.13333333% [68.8181%, 75.2230%]
llama_perf_context_print: load time = 488.78 ms
llama_perf_context_print: prompt eval time = 371768.51 ms / 126038 tokens ( 2.95 ms per token, 339.02 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 377711.49 ms / 126039 tokens
ggml_metal_free: deallocating

1025
scores/Hammer2.1-7b-F16.md Normal file

File diff suppressed because it is too large Load Diff

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 29 key-value pairs and 339 tensors from ./Hammer2.1-7b-F16.gguf (version GGUF V3 (latest))
Final result: 39.0667 +/- 1.7827
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 527.95 ms
llama_perf_context_print: prompt eval time = 194228.90 ms / 67719 tokens ( 2.87 ms per token, 348.66 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 196042.71 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 29 key-value pairs and 339 tensors from ./Hammer2.1-7b-F16.gguf (version GGUF V3 (latest))
Final result: 30.4000 +/- 1.6807
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 508.17 ms
llama_perf_context_print: prompt eval time = 146171.42 ms / 49696 tokens ( 2.94 ms per token, 339.98 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 148488.63 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 29 key-value pairs and 339 tensors from ./Hammer2.1-7b-F16.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.5333 +/- 1.6658
llama_perf_context_print: load time = 536.90 ms
llama_perf_context_print: prompt eval time = 61935.07 ms / 21448 tokens ( 2.89 ms per token, 346.30 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 62681.37 ms / 21449 tokens
ggml_metal_free: deallocating

1029
scores/Hammer2.1-7b-IQ3_M.md Normal file

File diff suppressed because it is too large Load Diff

1029
scores/Hammer2.1-7b-IQ3_S.md Normal file

File diff suppressed because it is too large Load Diff

File diff suppressed because it is too large Load Diff

File diff suppressed because it is too large Load Diff

File diff suppressed because it is too large Load Diff

File diff suppressed because it is too large Load Diff

File diff suppressed because it is too large Load Diff

File diff suppressed because it is too large Load Diff

File diff suppressed because it is too large Load Diff

File diff suppressed because it is too large Load Diff

1029
scores/Hammer2.1-7b-Q6_K.md Normal file

File diff suppressed because it is too large Load Diff

1029
scores/Hammer2.1-7b-Q8_0.md Normal file

File diff suppressed because it is too large Load Diff

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_M.gguf (version GGUF V3 (latest))
Final result: 56.5333 +/- 1.8113
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 1682.53 ms
llama_perf_context_print: prompt eval time = 106353.34 ms / 35972 tokens ( 2.96 ms per token, 338.23 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 107373.11 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_M.gguf (version GGUF V3 (latest))
750 71.46666667% [68.1319%, 74.5827%]
llama_perf_context_print: load time = 253.78 ms
llama_perf_context_print: prompt eval time = 375081.86 ms / 126038 tokens ( 2.98 ms per token, 336.03 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 380601.65 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_M.gguf (version GGUF V3 (latest))
Final result: 38.0000 +/- 1.7736
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 252.71 ms
llama_perf_context_print: prompt eval time = 195840.26 ms / 67719 tokens ( 2.89 ms per token, 345.79 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 197498.82 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.827809 ± 0.071236
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 98.97%
Mean ln(PPL(Q)/PPL(base)) : 0.048068 ± 0.001042
Mean PPL(Q)/PPL(base) : 1.049242 ± 0.001093
Mean PPL(Q)-PPL(base) : 0.461232 ± 0.011000
====== KL divergence statistics ======
Mean KLD: 0.059970 ± 0.000226
Maximum KLD: 5.830572
99.9% KLD: 0.899464
99.0% KLD: 0.388867
99.0% KLD: 0.388867
Median KLD: 0.039933
10.0% KLD: 0.000901
5.0% KLD: 0.000182
1.0% KLD: 0.000013
Minimum KLD: -0.000098
====== Token probability statistics ======
Mean Δp: -0.217 ± 0.017 %
Maximum Δp: 90.326%
99.9% Δp: 34.151%
99.0% Δp: 18.912%
95.0% Δp: 9.734%
90.0% Δp: 5.703%
75.0% Δp: 1.068%
Median Δp: -0.002%
25.0% Δp: -1.292%
10.0% Δp: -6.110%
5.0% Δp: -10.513%
1.0% Δp: -22.188%
0.1% Δp: -43.583%
Minimum Δp: -81.978%
RMS Δp : 6.574 ± 0.032 %
Same top p: 87.097 ± 0.087 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_M.gguf (version GGUF V3 (latest))
Final result: 29.8667 +/- 1.6723
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 252.53 ms
llama_perf_context_print: prompt eval time = 151286.74 ms / 49696 tokens ( 3.04 ms per token, 328.49 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 153386.94 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_M.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.6667 +/- 1.6636
llama_perf_context_print: load time = 254.13 ms
llama_perf_context_print: prompt eval time = 63679.99 ms / 21448 tokens ( 2.97 ms per token, 336.81 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 64320.92 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_S.gguf (version GGUF V3 (latest))
Final result: 55.8667 +/- 1.8143
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 1546.80 ms
llama_perf_context_print: prompt eval time = 106501.37 ms / 35972 tokens ( 2.96 ms per token, 337.76 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 107504.12 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_S.gguf (version GGUF V3 (latest))
750 70.26666667% [66.8989%, 73.4279%]
llama_perf_context_print: load time = 251.76 ms
llama_perf_context_print: prompt eval time = 375564.87 ms / 126038 tokens ( 2.98 ms per token, 335.60 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 381034.82 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_S.gguf (version GGUF V3 (latest))
Final result: 38.6667 +/- 1.7794
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 252.58 ms
llama_perf_context_print: prompt eval time = 196145.96 ms / 67719 tokens ( 2.90 ms per token, 345.25 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 197800.20 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.942560 ± 0.071243
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 98.64%
Mean ln(PPL(Q)/PPL(base)) : 0.059677 ± 0.001178
Mean PPL(Q)/PPL(base) : 1.061493 ± 0.001251
Mean PPL(Q)-PPL(base) : 0.575982 ± 0.012339
====== KL divergence statistics ======
Mean KLD: 0.079935 ± 0.000289
Maximum KLD: 5.765717
99.9% KLD: 1.150883
99.0% KLD: 0.493128
99.0% KLD: 0.493128
Median KLD: 0.053589
10.0% KLD: 0.001480
5.0% KLD: 0.000323
1.0% KLD: 0.000027
Minimum KLD: -0.000038
====== Token probability statistics ======
Mean Δp: -0.885 ± 0.019 %
Maximum Δp: 91.523%
99.9% Δp: 34.667%
99.0% Δp: 19.093%
95.0% Δp: 9.310%
90.0% Δp: 5.250%
75.0% Δp: 0.738%
Median Δp: -0.033%
25.0% Δp: -2.046%
10.0% Δp: -7.921%
5.0% Δp: -13.116%
1.0% Δp: -26.866%
0.1% Δp: -52.295%
Minimum Δp: -90.266%
RMS Δp : 7.445 ± 0.038 %
Same top p: 85.475 ± 0.091 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_S.gguf (version GGUF V3 (latest))
Final result: 30.9333 +/- 1.6889
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 248.10 ms
llama_perf_context_print: prompt eval time = 151449.03 ms / 49696 tokens ( 3.05 ms per token, 328.14 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 153545.12 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ3_S.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.4000 +/- 1.6680
llama_perf_context_print: load time = 250.05 ms
llama_perf_context_print: prompt eval time = 63750.53 ms / 21448 tokens ( 2.97 ms per token, 336.44 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 64386.61 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ4_NL.gguf (version GGUF V3 (latest))
Final result: 57.3333 +/- 1.8072
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 1954.91 ms
llama_perf_context_print: prompt eval time = 110236.28 ms / 35972 tokens ( 3.06 ms per token, 326.32 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 111260.63 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ4_NL.gguf (version GGUF V3 (latest))
750 71.20000000% [67.8576%, 74.3263%]
llama_perf_context_print: load time = 270.76 ms
llama_perf_context_print: prompt eval time = 388008.19 ms / 126038 tokens ( 3.08 ms per token, 324.83 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 393490.42 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ4_NL.gguf (version GGUF V3 (latest))
Final result: 38.5333 +/- 1.7783
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 262.83 ms
llama_perf_context_print: prompt eval time = 202888.48 ms / 67719 tokens ( 3.00 ms per token, 333.77 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 204561.85 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.585928 ± 0.068060
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 99.51%
Mean ln(PPL(Q)/PPL(base)) : 0.023148 ± 0.000703
Mean PPL(Q)/PPL(base) : 1.023418 ± 0.000719
Mean PPL(Q)-PPL(base) : 0.219350 ± 0.006862
====== KL divergence statistics ======
Mean KLD: 0.026627 ± 0.000108
Maximum KLD: 2.930657
99.9% KLD: 0.451824
99.0% KLD: 0.171826
99.0% KLD: 0.171826
Median KLD: 0.017210
10.0% KLD: 0.000451
5.0% KLD: 0.000095
1.0% KLD: 0.000007
Minimum KLD: -0.000060
====== Token probability statistics ======
Mean Δp: -0.523 ± 0.011 %
Maximum Δp: 54.223%
99.9% Δp: 22.680%
99.0% Δp: 11.163%
95.0% Δp: 5.267%
90.0% Δp: 2.899%
75.0% Δp: 0.419%
Median Δp: -0.014%
25.0% Δp: -1.188%
10.0% Δp: -4.623%
5.0% Δp: -7.618%
1.0% Δp: -15.106%
0.1% Δp: -30.059%
Minimum Δp: -89.467%
RMS Δp : 4.346 ± 0.026 %
Same top p: 91.424 ± 0.073 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ4_NL.gguf (version GGUF V3 (latest))
Final result: 30.1333 +/- 1.6766
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 260.43 ms
llama_perf_context_print: prompt eval time = 156558.33 ms / 49696 tokens ( 3.15 ms per token, 317.43 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 158668.65 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-IQ4_NL.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 69.8667 +/- 1.6766
llama_perf_context_print: load time = 266.77 ms
llama_perf_context_print: prompt eval time = 65947.52 ms / 21448 tokens ( 3.07 ms per token, 325.23 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 66587.55 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_L.gguf (version GGUF V3 (latest))
Final result: 57.2000 +/- 1.8079
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 1728.15 ms
llama_perf_context_print: prompt eval time = 114247.06 ms / 35972 tokens ( 3.18 ms per token, 314.86 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 115262.68 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_L.gguf (version GGUF V3 (latest))
750 69.86666667% [66.4884%, 73.0424%]
llama_perf_context_print: load time = 256.26 ms
llama_perf_context_print: prompt eval time = 403213.28 ms / 126038 tokens ( 3.20 ms per token, 312.58 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 408748.80 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_L.gguf (version GGUF V3 (latest))
Final result: 36.6667 +/- 1.7608
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 252.71 ms
llama_perf_context_print: prompt eval time = 210303.01 ms / 67719 tokens ( 3.11 ms per token, 322.01 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 211970.11 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 10.010842 ± 0.071669
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 98.55%
Mean ln(PPL(Q)/PPL(base)) : 0.066521 ± 0.001216
Mean PPL(Q)/PPL(base) : 1.068783 ± 0.001299
Mean PPL(Q)-PPL(base) : 0.644264 ± 0.012880
====== KL divergence statistics ======
Mean KLD: 0.080230 ± 0.000317
Maximum KLD: 5.784035
99.9% KLD: 1.294543
99.0% KLD: 0.528458
99.0% KLD: 0.528458
Median KLD: 0.051107
10.0% KLD: 0.001336
5.0% KLD: 0.000286
1.0% KLD: 0.000024
Minimum KLD: -0.000067
====== Token probability statistics ======
Mean Δp: -1.245 ± 0.019 %
Maximum Δp: 81.317%
99.9% Δp: 34.733%
99.0% Δp: 17.968%
95.0% Δp: 8.306%
90.0% Δp: 4.460%
75.0% Δp: 0.555%
Median Δp: -0.050%
25.0% Δp: -2.346%
10.0% Δp: -8.488%
5.0% Δp: -13.739%
1.0% Δp: -28.248%
0.1% Δp: -54.584%
Minimum Δp: -91.716%
RMS Δp : 7.556 ± 0.040 %
Same top p: 85.515 ± 0.091 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_L.gguf (version GGUF V3 (latest))
Final result: 29.6000 +/- 1.6680
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 253.28 ms
llama_perf_context_print: prompt eval time = 162637.78 ms / 49696 tokens ( 3.27 ms per token, 305.56 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 164733.51 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_L.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.9333 +/- 1.6591
llama_perf_context_print: load time = 254.30 ms
llama_perf_context_print: prompt eval time = 68423.01 ms / 21448 tokens ( 3.19 ms per token, 313.46 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 69062.22 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_M.gguf (version GGUF V3 (latest))
Final result: 55.7333 +/- 1.8149
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 1623.86 ms
llama_perf_context_print: prompt eval time = 111857.08 ms / 35972 tokens ( 3.11 ms per token, 321.59 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 112882.88 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_M.gguf (version GGUF V3 (latest))
750 69.86666667% [66.4884%, 73.0424%]
llama_perf_context_print: load time = 263.27 ms
llama_perf_context_print: prompt eval time = 395258.72 ms / 126038 tokens ( 3.14 ms per token, 318.87 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 400785.88 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_M.gguf (version GGUF V3 (latest))
Final result: 37.2000 +/- 1.7661
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 248.74 ms
llama_perf_context_print: prompt eval time = 206021.07 ms / 67719 tokens ( 3.04 ms per token, 328.70 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 207679.96 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 10.109660 ± 0.072549
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 98.44%
Mean ln(PPL(Q)/PPL(base)) : 0.076344 ± 0.001263
Mean PPL(Q)/PPL(base) : 1.079333 ± 0.001363
Mean PPL(Q)-PPL(base) : 0.743083 ± 0.013719
====== KL divergence statistics ======
Mean KLD: 0.086444 ± 0.000342
Maximum KLD: 6.291467
99.9% KLD: 1.404206
99.0% KLD: 0.571756
99.0% KLD: 0.571756
Median KLD: 0.054763
10.0% KLD: 0.001465
5.0% KLD: 0.000315
1.0% KLD: 0.000028
Minimum KLD: -0.000235
====== Token probability statistics ======
Mean Δp: -1.387 ± 0.020 %
Maximum Δp: 82.526%
99.9% Δp: 35.574%
99.0% Δp: 18.394%
95.0% Δp: 8.338%
90.0% Δp: 4.407%
75.0% Δp: 0.508%
Median Δp: -0.066%
25.0% Δp: -2.517%
10.0% Δp: -8.904%
5.0% Δp: -14.322%
1.0% Δp: -29.340%
0.1% Δp: -57.123%
Minimum Δp: -93.887%
RMS Δp : 7.845 ± 0.041 %
Same top p: 85.114 ± 0.092 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_M.gguf (version GGUF V3 (latest))
Final result: 29.8667 +/- 1.6723
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 254.73 ms
llama_perf_context_print: prompt eval time = 159391.21 ms / 49696 tokens ( 3.21 ms per token, 311.79 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 161494.93 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_M.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.4000 +/- 1.6680
llama_perf_context_print: load time = 248.84 ms
llama_perf_context_print: prompt eval time = 67078.03 ms / 21448 tokens ( 3.13 ms per token, 319.75 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 67715.29 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_S.gguf (version GGUF V3 (latest))
Final result: 56.5333 +/- 1.8113
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 1475.41 ms
llama_perf_context_print: prompt eval time = 113557.50 ms / 35972 tokens ( 3.16 ms per token, 316.77 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 114575.72 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_S.gguf (version GGUF V3 (latest))
750 68.80000000% [65.3955%, 72.0129%]
llama_perf_context_print: load time = 247.18 ms
llama_perf_context_print: prompt eval time = 400796.81 ms / 126038 tokens ( 3.18 ms per token, 314.47 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 406267.35 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_S.gguf (version GGUF V3 (latest))
Final result: 37.3333 +/- 1.7674
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 252.83 ms
llama_perf_context_print: prompt eval time = 209065.30 ms / 67719 tokens ( 3.09 ms per token, 323.91 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 210718.75 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 10.224307 ± 0.073520
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 98.16%
Mean ln(PPL(Q)/PPL(base)) : 0.087620 ± 0.001374
Mean PPL(Q)/PPL(base) : 1.091573 ± 0.001499
Mean PPL(Q)-PPL(base) : 0.857730 ± 0.015181
====== KL divergence statistics ======
Mean KLD: 0.101951 ± 0.000397
Maximum KLD: 6.073646
99.9% KLD: 1.662331
99.0% KLD: 0.673299
99.0% KLD: 0.673299
Median KLD: 0.064775
10.0% KLD: 0.001756
5.0% KLD: 0.000385
1.0% KLD: 0.000033
Minimum KLD: -0.000036
====== Token probability statistics ======
Mean Δp: -1.602 ± 0.022 %
Maximum Δp: 88.808%
99.9% Δp: 37.347%
99.0% Δp: 19.623%
95.0% Δp: 8.796%
90.0% Δp: 4.589%
75.0% Δp: 0.491%
Median Δp: -0.086%
25.0% Δp: -2.849%
10.0% Δp: -9.793%
5.0% Δp: -15.740%
1.0% Δp: -31.943%
0.1% Δp: -61.801%
Minimum Δp: -98.023%
RMS Δp : 8.487 ± 0.044 %
Same top p: 84.095 ± 0.095 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_S.gguf (version GGUF V3 (latest))
Final result: 29.0667 +/- 1.6591
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 251.28 ms
llama_perf_context_print: prompt eval time = 161651.93 ms / 49696 tokens ( 3.25 ms per token, 307.43 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 163734.79 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q3_K_S.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 69.3333 +/- 1.6849
llama_perf_context_print: load time = 249.55 ms
llama_perf_context_print: prompt eval time = 68038.26 ms / 21448 tokens ( 3.17 ms per token, 315.23 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 68675.52 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_M.gguf (version GGUF V3 (latest))
Final result: 57.8667 +/- 1.8042
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 1953.65 ms
llama_perf_context_print: prompt eval time = 115394.35 ms / 35972 tokens ( 3.21 ms per token, 311.73 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 116424.56 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_M.gguf (version GGUF V3 (latest))
750 71.20000000% [67.8576%, 74.3263%]
llama_perf_context_print: load time = 260.90 ms
llama_perf_context_print: prompt eval time = 405989.14 ms / 126038 tokens ( 3.22 ms per token, 310.45 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 411436.74 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_M.gguf (version GGUF V3 (latest))
Final result: 38.0000 +/- 1.7736
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 260.22 ms
llama_perf_context_print: prompt eval time = 212336.08 ms / 67719 tokens ( 3.14 ms per token, 318.92 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 214000.87 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.513981 ± 0.067589
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 99.56%
Mean ln(PPL(Q)/PPL(base)) : 0.015615 ± 0.000668
Mean PPL(Q)/PPL(base) : 1.015737 ± 0.000679
Mean PPL(Q)-PPL(base) : 0.147404 ± 0.006420
====== KL divergence statistics ======
Mean KLD: 0.023792 ± 0.000107
Maximum KLD: 5.229656
99.9% KLD: 0.422463
99.0% KLD: 0.159748
99.0% KLD: 0.159748
Median KLD: 0.014752
10.0% KLD: 0.000388
5.0% KLD: 0.000078
1.0% KLD: 0.000005
Minimum KLD: -0.000044
====== Token probability statistics ======
Mean Δp: -0.337 ± 0.010 %
Maximum Δp: 67.097%
99.9% Δp: 22.363%
99.0% Δp: 11.173%
95.0% Δp: 5.170%
90.0% Δp: 2.864%
75.0% Δp: 0.466%
Median Δp: -0.008%
25.0% Δp: -0.987%
10.0% Δp: -3.948%
5.0% Δp: -6.542%
1.0% Δp: -13.717%
0.1% Δp: -28.959%
Minimum Δp: -91.120%
RMS Δp : 4.057 ± 0.026 %
Same top p: 92.039 ± 0.070 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_M.gguf (version GGUF V3 (latest))
Final result: 29.8667 +/- 1.6723
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 273.17 ms
llama_perf_context_print: prompt eval time = 163807.79 ms / 49696 tokens ( 3.30 ms per token, 303.38 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 165910.76 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_M.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 69.8667 +/- 1.6766
llama_perf_context_print: load time = 263.91 ms
llama_perf_context_print: prompt eval time = 69022.90 ms / 21448 tokens ( 3.22 ms per token, 310.74 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 69661.76 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_S.gguf (version GGUF V3 (latest))
Final result: 57.8667 +/- 1.8042
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 1866.16 ms
llama_perf_context_print: prompt eval time = 113639.43 ms / 35972 tokens ( 3.16 ms per token, 316.55 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 114663.24 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_S.gguf (version GGUF V3 (latest))
750 71.06666667% [67.7206%, 74.1981%]
llama_perf_context_print: load time = 262.17 ms
llama_perf_context_print: prompt eval time = 400033.89 ms / 126038 tokens ( 3.17 ms per token, 315.07 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 405454.55 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_S.gguf (version GGUF V3 (latest))
Final result: 37.8667 +/- 1.7724
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 258.75 ms
llama_perf_context_print: prompt eval time = 209132.23 ms / 67719 tokens ( 3.09 ms per token, 323.81 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 210787.92 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.529467 ± 0.067693
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 99.53%
Mean ln(PPL(Q)/PPL(base)) : 0.017241 ± 0.000692
Mean PPL(Q)/PPL(base) : 1.017391 ± 0.000704
Mean PPL(Q)-PPL(base) : 0.162890 ± 0.006661
====== KL divergence statistics ======
Mean KLD: 0.025429 ± 0.000116
Maximum KLD: 6.250518
99.9% KLD: 0.457138
99.0% KLD: 0.169697
99.0% KLD: 0.169697
Median KLD: 0.015727
10.0% KLD: 0.000429
5.0% KLD: 0.000085
1.0% KLD: 0.000006
Minimum KLD: -0.000044
====== Token probability statistics ======
Mean Δp: -0.378 ± 0.011 %
Maximum Δp: 64.248%
99.9% Δp: 23.260%
99.0% Δp: 11.368%
95.0% Δp: 5.279%
90.0% Δp: 2.904%
75.0% Δp: 0.455%
Median Δp: -0.010%
25.0% Δp: -1.037%
10.0% Δp: -4.112%
5.0% Δp: -6.788%
1.0% Δp: -14.402%
0.1% Δp: -30.405%
Minimum Δp: -91.281%
RMS Δp : 4.191 ± 0.026 %
Same top p: 91.792 ± 0.071 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_S.gguf (version GGUF V3 (latest))
Final result: 29.6000 +/- 1.6680
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 264.10 ms
llama_perf_context_print: prompt eval time = 161357.40 ms / 49696 tokens ( 3.25 ms per token, 307.99 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 163438.71 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q4_K_S.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 69.6000 +/- 1.6807
llama_perf_context_print: load time = 262.82 ms
llama_perf_context_print: prompt eval time = 67973.27 ms / 21448 tokens ( 3.17 ms per token, 315.54 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 68612.34 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_M.gguf (version GGUF V3 (latest))
Final result: 58.0000 +/- 1.8034
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 2350.20 ms
llama_perf_context_print: prompt eval time = 114769.46 ms / 35972 tokens ( 3.19 ms per token, 313.43 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 115764.61 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_M.gguf (version GGUF V3 (latest))
750 71.33333333% [67.9947%, 74.4545%]
llama_perf_context_print: load time = 287.60 ms
llama_perf_context_print: prompt eval time = 405854.89 ms / 126038 tokens ( 3.22 ms per token, 310.55 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 411351.99 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_M.gguf (version GGUF V3 (latest))
Final result: 38.6667 +/- 1.7794
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 281.78 ms
llama_perf_context_print: prompt eval time = 211602.07 ms / 67719 tokens ( 3.12 ms per token, 320.03 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 213278.72 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.411472 ± 0.066924
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 99.88%
Mean ln(PPL(Q)/PPL(base)) : 0.004782 ± 0.000343
Mean PPL(Q)/PPL(base) : 1.004793 ± 0.000345
Mean PPL(Q)-PPL(base) : 0.044894 ± 0.003261
====== KL divergence statistics ======
Mean KLD: 0.005881 ± 0.000035
Maximum KLD: 2.356096
99.9% KLD: 0.110341
99.0% KLD: 0.039253
99.0% KLD: 0.039253
Median KLD: 0.003644
10.0% KLD: 0.000086
5.0% KLD: 0.000015
1.0% KLD: -0.000000
Minimum KLD: -0.000153
====== Token probability statistics ======
Mean Δp: -0.012 ± 0.005 %
Maximum Δp: 40.226%
99.9% Δp: 11.947%
99.0% Δp: 6.009%
95.0% Δp: 2.947%
90.0% Δp: 1.726%
75.0% Δp: 0.356%
Median Δp: -0.000%
25.0% Δp: -0.356%
10.0% Δp: -1.732%
5.0% Δp: -2.967%
1.0% Δp: -6.302%
0.1% Δp: -13.317%
Minimum Δp: -62.412%
RMS Δp : 2.028 ± 0.015 %
Same top p: 95.904 ± 0.051 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_M.gguf (version GGUF V3 (latest))
Final result: 30.2667 +/- 1.6787
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 279.86 ms
llama_perf_context_print: prompt eval time = 163647.89 ms / 49696 tokens ( 3.29 ms per token, 303.68 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 165759.65 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_M.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.6667 +/- 1.6636
llama_perf_context_print: load time = 283.78 ms
llama_perf_context_print: prompt eval time = 68831.62 ms / 21448 tokens ( 3.21 ms per token, 311.60 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 69470.43 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_S.gguf (version GGUF V3 (latest))
Final result: 58.4000 +/- 1.8010
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 2246.68 ms
llama_perf_context_print: prompt eval time = 116329.15 ms / 35972 tokens ( 3.23 ms per token, 309.23 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 117348.24 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_S.gguf (version GGUF V3 (latest))
750 71.46666667% [68.1319%, 74.5827%]
llama_perf_context_print: load time = 275.63 ms
llama_perf_context_print: prompt eval time = 410573.07 ms / 126038 tokens ( 3.26 ms per token, 306.98 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 416121.78 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_S.gguf (version GGUF V3 (latest))
Final result: 38.8000 +/- 1.7805
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 277.93 ms
llama_perf_context_print: prompt eval time = 214179.80 ms / 67719 tokens ( 3.16 ms per token, 316.18 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 215843.77 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.410802 ± 0.066883
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 99.88%
Mean ln(PPL(Q)/PPL(base)) : 0.004710 ± 0.000355
Mean PPL(Q)/PPL(base) : 1.004722 ± 0.000357
Mean PPL(Q)-PPL(base) : 0.044224 ± 0.003363
====== KL divergence statistics ======
Mean KLD: 0.006252 ± 0.000036
Maximum KLD: 2.202329
99.9% KLD: 0.113542
99.0% KLD: 0.041912
99.0% KLD: 0.041912
Median KLD: 0.003874
10.0% KLD: 0.000095
5.0% KLD: 0.000017
1.0% KLD: 0.000000
Minimum KLD: -0.000135
====== Token probability statistics ======
Mean Δp: -0.049 ± 0.005 %
Maximum Δp: 48.849%
99.9% Δp: 12.232%
99.0% Δp: 6.155%
95.0% Δp: 2.940%
90.0% Δp: 1.698%
75.0% Δp: 0.326%
Median Δp: -0.000%
25.0% Δp: -0.399%
10.0% Δp: -1.846%
5.0% Δp: -3.125%
1.0% Δp: -6.593%
0.1% Δp: -13.502%
Minimum Δp: -61.344%
RMS Δp : 2.085 ± 0.015 %
Same top p: 95.765 ± 0.052 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_S.gguf (version GGUF V3 (latest))
Final result: 30.5333 +/- 1.6828
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 285.20 ms
llama_perf_context_print: prompt eval time = 165611.98 ms / 49696 tokens ( 3.33 ms per token, 300.07 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 167719.47 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q5_K_S.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.6667 +/- 1.6636
llama_perf_context_print: load time = 282.74 ms
llama_perf_context_print: prompt eval time = 69663.19 ms / 21448 tokens ( 3.25 ms per token, 307.88 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 70297.96 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q6_K.gguf (version GGUF V3 (latest))
Final result: 58.8000 +/- 1.7984
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 2767.04 ms
llama_perf_context_print: prompt eval time = 112965.93 ms / 35972 tokens ( 3.14 ms per token, 318.43 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 113983.11 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q6_K.gguf (version GGUF V3 (latest))
750 72.00000000% [68.6807%, 75.0950%]
llama_perf_context_print: load time = 303.64 ms
llama_perf_context_print: prompt eval time = 399056.52 ms / 126038 tokens ( 3.17 ms per token, 315.84 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 404573.51 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q6_K.gguf (version GGUF V3 (latest))
Final result: 38.9333 +/- 1.7816
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 309.26 ms
llama_perf_context_print: prompt eval time = 208947.61 ms / 67719 tokens ( 3.09 ms per token, 324.10 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 210629.00 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.384561 ± 0.066615
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 99.96%
Mean ln(PPL(Q)/PPL(base)) : 0.001918 ± 0.000206
Mean PPL(Q)/PPL(base) : 1.001920 ± 0.000207
Mean PPL(Q)-PPL(base) : 0.017984 ± 0.001945
====== KL divergence statistics ======
Mean KLD: 0.001779 ± 0.000018
Maximum KLD: 1.977490
99.9% KLD: 0.040032
99.0% KLD: 0.012342
99.0% KLD: 0.012342
Median KLD: 0.001053
10.0% KLD: 0.000028
5.0% KLD: 0.000005
1.0% KLD: -0.000001
Minimum KLD: -0.000152
====== Token probability statistics ======
Mean Δp: -0.027 ± 0.003 %
Maximum Δp: 49.334%
99.9% Δp: 7.267%
99.0% Δp: 3.236%
95.0% Δp: 1.482%
90.0% Δp: 0.845%
75.0% Δp: 0.164%
Median Δp: -0.000%
25.0% Δp: -0.202%
10.0% Δp: -0.938%
5.0% Δp: -1.620%
1.0% Δp: -3.443%
0.1% Δp: -7.356%
Minimum Δp: -46.022%
RMS Δp : 1.101 ± 0.012 %
Same top p: 97.794 ± 0.038 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q6_K.gguf (version GGUF V3 (latest))
Final result: 29.7333 +/- 1.6702
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 318.58 ms
llama_perf_context_print: prompt eval time = 161421.27 ms / 49696 tokens ( 3.25 ms per token, 307.87 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 163513.36 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q6_K.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.4000 +/- 1.6680
llama_perf_context_print: load time = 313.52 ms
llama_perf_context_print: prompt eval time = 67409.09 ms / 21448 tokens ( 3.14 ms per token, 318.18 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 68029.15 ms / 21449 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q8_0.gguf (version GGUF V3 (latest))
Final result: 58.5333 +/- 1.8002
Random chance: 25.0083 +/- 1.5824
llama_perf_context_print: load time = 3247.05 ms
llama_perf_context_print: prompt eval time = 106606.41 ms / 35972 tokens ( 2.96 ms per token, 337.43 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 107624.02 ms / 35973 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,12 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q8_0.gguf (version GGUF V3 (latest))
750 72.00000000% [68.6807%, 75.0950%]
llama_perf_context_print: load time = 318.54 ms
llama_perf_context_print: prompt eval time = 375734.27 ms / 126038 tokens ( 2.98 ms per token, 335.44 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 381209.78 ms / 126039 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q8_0.gguf (version GGUF V3 (latest))
Final result: 38.9333 +/- 1.7816
Random chance: 25.0000 +/- 1.5822
llama_perf_context_print: load time = 316.60 ms
llama_perf_context_print: prompt eval time = 197007.43 ms / 67719 tokens ( 2.91 ms per token, 343.74 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 198698.88 ms / 67720 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,37 @@
====== Perplexity statistics ======
Mean PPL(Q) : 9.379380 ± 0.066571
Mean PPL(base) : 9.366577 ± 0.066397
Cor(ln(PPL(Q)), ln(PPL(base))): 99.98%
Mean ln(PPL(Q)/PPL(base)) : 0.001366 ± 0.000154
Mean PPL(Q)/PPL(base) : 1.001367 ± 0.000154
Mean PPL(Q)-PPL(base) : 0.012802 ± 0.001454
====== KL divergence statistics ======
Mean KLD: 0.000692 ± 0.000012
Maximum KLD: 1.498512
99.9% KLD: 0.024913
99.0% KLD: 0.005779
99.0% KLD: 0.005779
Median KLD: 0.000331
10.0% KLD: 0.000011
5.0% KLD: 0.000002
1.0% KLD: -0.000002
Minimum KLD: -0.000134
====== Token probability statistics ======
Mean Δp: -0.028 ± 0.002 %
Maximum Δp: 21.822%
99.9% Δp: 4.935%
99.0% Δp: 1.949%
95.0% Δp: 0.826%
90.0% Δp: 0.458%
75.0% Δp: 0.085%
Median Δp: -0.000%
25.0% Δp: -0.129%
10.0% Δp: -0.558%
5.0% Δp: -0.952%
1.0% Δp: -2.125%
0.1% Δp: -5.152%
Minimum Δp: -39.341%
RMS Δp : 0.706 ± 0.010 %
Same top p: 98.711 ± 0.029 %

View File

@@ -0,0 +1,13 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q8_0.gguf (version GGUF V3 (latest))
Final result: 30.1333 +/- 1.6766
Random chance: 19.8992 +/- 1.4588
llama_perf_context_print: load time = 324.70 ms
llama_perf_context_print: prompt eval time = 151947.90 ms / 49696 tokens ( 3.06 ms per token, 327.06 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 154045.90 ms / 49697 tokens
ggml_metal_free: deallocating

View File

@@ -0,0 +1,11 @@
build: 5170 (658987cf) with Apple clang version 17.0.0 (clang-1700.0.13.3) for arm64-apple-darwin24.4.0
llama_model_load_from_file_impl: using device Metal (Apple M3 Pro) - 27647 MiB free
llama_model_loader: loaded meta data with 33 key-value pairs and 339 tensors from ./Hammer2.1-7b-Q8_0.gguf (version GGUF V3 (latest))
Final Winogrande score(750 tasks): 70.2667 +/- 1.6702
llama_perf_context_print: load time = 321.25 ms
llama_perf_context_print: prompt eval time = 63938.77 ms / 21448 tokens ( 2.98 ms per token, 335.45 tokens per second)
llama_perf_context_print: eval time = 0.00 ms / 1 runs ( 0.00 ms per token, inf tokens per second)
llama_perf_context_print: total time = 64599.27 ms / 21449 tokens
ggml_metal_free: deallocating