Re-convert with fixed rotary q/k permutation (baseRT-internal#114); NIAH-validated at 2k/8k. Q4 now uses the default-q4 quality profile (embeddings/head at higher precision).
Browse files
Llama-3.1-8B-Instruct-Q8.base
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 8288215040
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:833e8c57398d0878c7e2684bcb570dc9c3d0b5a9eb6107fa337e631b50af6f87
|
| 3 |
size 8288215040
|