diff options
author | Georgi Gerganov <ggerganov@gmail.com> | 2023-06-26 19:45:09 +0300 |
---|---|---|
committer | GitHub <noreply@github.com> | 2023-06-26 19:45:09 +0300 |
commit | 412c60e4739367144e51e59add5dc7749d084115 (patch) | |
tree | 0e866ad81712d6b8636adce7002fb119933378ae /README.md | |
parent | 6769e944c727c63612dcafbef52009d21ae00fff (diff) |
readme : add link to new k-quants for visibility
Diffstat (limited to 'README.md')
-rw-r--r-- | README.md | 1 |
1 files changed, 1 insertions, 0 deletions
@@ -11,6 +11,7 @@ Inference of [LLaMA](https://arxiv.org/abs/2302.13971) model in pure C/C++ **Hot topics:** +- k-quants now support super-block size of 64: https://github.com/ggerganov/llama.cpp/pull/2001 - New roadmap: https://github.com/users/ggerganov/projects/7 - Azure CI brainstorming: https://github.com/ggerganov/llama.cpp/discussions/1985 - p1 : LLM-based code completion engine at the edge : https://github.com/ggml-org/p1/discussions/1 |