diff options
author | Galunid <karolek1231456@gmail.com> | 2024-05-31 10:24:41 +0200 |
---|---|---|
committer | GitHub <noreply@github.com> | 2024-05-31 18:24:41 +1000 |
commit | 2e32f874e675f7bc5307cb7b4470ddbe090bab8f (patch) | |
tree | 491c7b156deb67802f75ee5a99266923d755be42 | |
parent | 1af511fc22cba4959dd8bced5501df9e8af6ddf9 (diff) |
Somehow '**' got lost (#7663)
-rw-r--r-- | README.md | 2 |
1 files changed, 1 insertions, 1 deletions
@@ -22,7 +22,7 @@ Inference of Meta's [LLaMA](https://arxiv.org/abs/2302.13971) model (and others) ### Hot topics -- **`convert.py` has been deprecated and moved to `examples/convert-legacy-llama.py`, please use `convert-hf-to-gguf.py` https://github.com/ggerganov/llama.cpp/pull/7430 +- **`convert.py` has been deprecated and moved to `examples/convert-legacy-llama.py`, please use `convert-hf-to-gguf.py`** https://github.com/ggerganov/llama.cpp/pull/7430 - Initial Flash-Attention support: https://github.com/ggerganov/llama.cpp/pull/5021 - BPE pre-tokenization support has been added: https://github.com/ggerganov/llama.cpp/pull/6920 - MoE memory layout has been updated - reconvert models for `mmap` support and regenerate `imatrix` https://github.com/ggerganov/llama.cpp/pull/6387 |