IQ3_K_R4 (#145)

* iq3_k_r4 WIP * iq3_k_r4: Zen4 * iq3_k_r4: AVX2 * iq3_k_r4: NEON * iq3_k_r4: faster matrix x vector multiplication on NEON --------- Co-authored-by: Iwan Kawrakow <iwan.kawrakow@gmail.com>
author: Kawrakow <iwankawrakow@gmail.com> 2024-12-17 07:51:11 +0100
committer: GitHub <noreply@github.com> 2024-12-17 07:51:11 +0100
commit: d69344f8ea72c6fe6ec16300b939586fa9633e2e (patch)
tree: b8c0efb7322169372543b020360bf0e27549fba5 /include
parent: 1714e46f137318152370beee16af92991042d7b4 (diff)
1 files changed, 1 insertions, 0 deletions
diff --git a/include/llama.h b/include/llama.h
index 988ffec7..026cf08e 100644
--- a/include/llama.h
+++ b/include/llama.h
@@ -193,6 +193,7 @@ extern "C" {
         LLAMA_FTYPE_MOSTLY_Q6_0_R4       = 335, // except 1d tensors
         LLAMA_FTYPE_MOSTLY_BF16_R16      = 232, // except 1d tensors
         LLAMA_FTYPE_MOSTLY_IQ2_BN_R4     = 337, // except 1d tensors
+        LLAMA_FTYPE_MOSTLY_IQ3_K_R4      = 339, // except 1d tensors
         LLAMA_FTYPE_MOSTLY_IQ4_K_R4      = 340, // except 1d tensors
         LLAMA_FTYPE_MOSTLY_Q8_K_R8       = 399, // except 1d tensors
author	Kawrakow <iwankawrakow@gmail.com>	2024-12-17 07:51:11 +0100
committer	GitHub <noreply@github.com>	2024-12-17 07:51:11 +0100
commit	d69344f8ea72c6fe6ec16300b939586fa9633e2e (patch)
tree	b8c0efb7322169372543b020360bf0e27549fba5 /include
parent	1714e46f137318152370beee16af92991042d7b4 (diff)