mirror of https://github.com/ggerganov/llama.cpp.git synced 2024-12-26 11:24:35 +00:00

History

Georgi Gerganov 88b5769487 gguf : deduplicate (#2629 ) * gguf : better type names * dedup : CPU + Metal is working * ggml : fix warnings about unused results * llama.cpp : fix line feed and compiler warning * llama : fix strncpy warning + note token_to_str does not write null * llama : restore the original load/save session implementation Will migrate this to GGUF in the future * convert-llama-h5-to-gguf.py : support alt ctx param name * ggml : assert when using ggml_mul with non-F32 src1 * examples : dedup simple --------- Co-authored-by: klosax <131523366+klosax@users.noreply.github.com>		2023-08-16 19:25:29 +03:00
..
CMakeLists.txt	cmake : install targets (#2256 )	2023-07-19 10:01:11 +03:00
quantize.cpp	gguf : deduplicate (#2629 )	2023-08-16 19:25:29 +03:00
README.md	Overhaul the examples structure	2023-03-25 20:26:40 +02:00

quantize

TODO