llama.cpp/index.html.gz at 644fd71b44c4cdbfc6482fbf0353d289c3bc29e6

mirror of https://github.com/ggerganov/llama.cpp.git synced 2024-12-27 20:04:35 +00:00

Georgi Gerganov 644fd71b44

sampling : refactor + optimize penalties sampler (#10803 )

* sampling : refactor + optimize penalties sampler

ggml-ci

* common : apply ignore_eos as logit bias

ggml-ci

* batched : remove penalties sampler

* params : allow penalty_last_n == -1 to be equal to context size

ggml-ci

* common : by default, move the penalties at the end of the sampling chain

ggml-ci

* common : ignore all EOG tokens

Co-authored-by: Diego Devesa <slarengh@gmail.com>

* common : move back the penalties at the front of the sampling chain

ggml-ci

* readme : restore hint about --ignore-eos flag [no ci]

* llama : minor

ggml-ci

* webui : update

---------

Co-authored-by: Diego Devesa <slarengh@gmail.com>

2024-12-16 12:31:14 +02:00

1.1 MiB

Raw History

View Raw

1.1 MiB Raw History

1.1 MiB

Raw History