llama.cpp

mirror of https://github.com/ggerganov/llama.cpp.git synced 2024-12-24 18:34:36 +00:00

History

Brian 1666f92dcd gguf-hash : update clib.json to point to original xxhash repo (#8491 ) * Update clib.json to point to Cyan4973 original xxhash Convinced Cyan4973 to add clib.json directly to his repo, so can now point the clib package directly to him now. Previously pointed to my fork with the clib.json package metadata https://github.com/Cyan4973/xxHash/pull/954 * gguf-hash: readme update to point to Cyan4973 xxHash repo [no ci]		2024-07-16 10:14:16 +03:00
..
baby-llama	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
batched	Inference support for T5 and FLAN-T5 model families (#5763 )	2024-07-04 15:46:11 +02:00
batched-bench	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
batched.swift	Detokenizer fixes (#8039 )	2024-07-05 19:01:35 +02:00
benchmark	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
convert-llama2c-to-ggml	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
cvector-generator	cvector: better prompt handling, add "mean vector" method (#8069 )	2024-06-25 13:59:54 +02:00
deprecation-warning	Deprecation warning to assist with migration to new binary names (#8283 )	2024-07-09 11:54:43 -04:00
embedding	Removes multiple newlines at the end of files that is breaking the editorconfig step of CI. (#8258 )	2024-07-02 12:18:10 -04:00
eval-callback	examples : sprintf -> snprintf (#8434 )	2024-07-12 10:46:14 +03:00
export-lora	export-lora : handle help argument (#8497 )	2024-07-16 10:04:45 +03:00
finetune	py : type-check all Python scripts with Pyright (#8341 )	2024-07-07 15:04:39 -04:00
gbnf-validator	llama : return nullptr from llama_grammar_init (#8093 )	2024-06-25 15:07:28 -04:00
gguf	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
gguf-hash	gguf-hash : update clib.json to point to original xxhash repo (#8491 )	2024-07-16 10:14:16 +03:00
gguf-split	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
gritlm	llama : allow pooled embeddings on any model (#7477 )	2024-06-21 08:38:22 +03:00
imatrix	llama : reorganize source code + improve CMake (#8006 )	2024-06-26 18:33:02 +03:00
infill	infill : assert prefix/suffix tokens + remove old space logic (#8351 )	2024-07-08 09:34:35 +03:00
jeopardy	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
llama-bench	llama-bench : fix RPC indication (#7936 )	2024-06-14 16:47:41 +03:00
llama.android	Delete examples/llama.android/llama/CMakeLists.txt (#8165 )	2024-06-27 16:39:29 +02:00
llama.swiftui	Detokenizer fixes (#8039 )	2024-07-05 19:01:35 +02:00
llava	py : type-check all Python scripts with Pyright (#8341 )	2024-07-07 15:04:39 -04:00
lookahead	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
lookup	Removes multiple newlines at the end of files that is breaking the editorconfig step of CI. (#8258 )	2024-07-02 12:18:10 -04:00
main	main : print error on empty input (#8456 )	2024-07-12 14:48:04 +03:00
main-cmake-pkg	Removes multiple newlines at the end of files that is breaking the editorconfig step of CI. (#8258 )	2024-07-02 12:18:10 -04:00
parallel	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
passkey	passkey : add short intro to README.md [no-ci] (#8317 )	2024-07-05 09:14:24 +03:00
perplexity	ppl : fix n_seq_max for perplexity (#8277 )	2024-07-03 20:33:31 +03:00
quantize	llama : valign + remove unused ftype (#8502 )	2024-07-16 10:00:30 +03:00
quantize-stats	ggml : minor naming changes (#8433 )	2024-07-12 10:46:02 +03:00
retrieval	llama : allow pooled embeddings on any model (#7477 )	2024-06-21 08:38:22 +03:00
rpc	llama : reorganize source code + improve CMake (#8006 )	2024-06-26 18:33:02 +03:00
save-load-state	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
server	fix ci (#8494 )	2024-07-15 19:23:10 +02:00
simple	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
speculative	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
sycl	Removes multiple newlines at the end of files that is breaking the editorconfig step of CI. (#8258 )	2024-07-02 12:18:10 -04:00
tokenize	tokenize : add --no-parse-special option (#8423 )	2024-07-11 10:41:48 +03:00
train-text-from-scratch	py : type-check all Python scripts with Pyright (#8341 )	2024-07-07 15:04:39 -04:00
base-translate.sh	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
chat-13B.bat	Create chat-13B.bat (#592 )	2023-03-29 20:21:09 +03:00
chat-13B.sh	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
chat-persistent.sh	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
chat-vicuna.sh	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
chat.sh	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
CMakeLists.txt	gguf-hash: model wide and per tensor hashing using xxhash and sha1 (#8048 )	2024-07-07 22:58:43 +10:00
convert_legacy_llama.py	py : type-check all Python scripts with Pyright (#8341 )	2024-07-07 15:04:39 -04:00
json_schema_pydantic_example.py	py : type-check all Python scripts with Pyright (#8341 )	2024-07-07 15:04:39 -04:00
json_schema_to_grammar.py	py : type-check all Python scripts with Pyright (#8341 )	2024-07-07 15:04:39 -04:00
llama.vim	llama.vim : added api key support (#5090 )	2024-01-23 08:51:27 +02:00
llm.vim	llm.vim : stop generation at multiple linebreaks, bind to <F2> (#2879 )	2023-08-30 09:50:55 +03:00
Miku.sh	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
pydantic_models_to_grammar_examples.py	pydantic : replace uses of __annotations__ with get_type_hints (#8474 )	2024-07-14 19:51:21 -04:00
pydantic_models_to_grammar.py	pydantic : replace uses of __annotations__ with get_type_hints (#8474 )	2024-07-14 19:51:21 -04:00
reason-act.sh	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
regex_to_grammar.py	py : switch to snake_case (#8305 )	2024-07-05 07:53:33 +03:00
server_embd.py	py : type-check all Python scripts with Pyright (#8341 )	2024-07-07 15:04:39 -04:00
server-llama2-13B.sh	`build`: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (#7809 )	2024-06-13 00:41:52 +01:00
ts-type-to-grammar.sh	JSON schema conversion: ⚡️ faster repetitions, min/maxLength for strings, cap number length (#6555 )	2024-04-12 19:43:38 +01:00