git.djapps.eu Git - pkg/ggml/sources/llama.cpp/commit

author	Georgi Gerganov <redacted>
	Mon, 21 Oct 2024 06:46:40 +0000 (09:46 +0300)
committer	GitHub <redacted>
	Mon, 21 Oct 2024 06:46:40 +0000 (09:46 +0300)
commit	55e47786e373c90fc7803e718e3e1dd6d53c3db6
tree	defc4984ec706d598676d4c2103e712706e4467f	tree
parent	bc219750845a59166d79f0d4ee3da1993b369b8a	commit \| diff

llama : default sampling changes + greedy update (#9897)

* llama : deprecate softmax sampler + fix dist sampler

ggml-ci

* tests : replace macros with functions

ggml-ci

* sampling : change temperature sampler logic

For t <= 0.0f, keep the max logit intact and set the rest to -inf

* cont : no need for special "greedy" logic

top-k == 1 is the same

* tests : init prob correctly

* llama : handle temp <= 0.0 in the temp_ext sampler too

ggml-ci

* cont : avoid extra loop in temperature sampler for sub-zero temp

ggml-ci

common/sampling.cpp		diff \| blob \| history
examples/llama.swiftui/llama.cpp.swift/LibLlama.swift		diff \| blob \| history
examples/save-load-state/save-load-state.cpp		diff \| blob \| history
examples/speculative/speculative.cpp		diff \| blob \| history
include/llama.h		diff \| blob \| history
src/llama-sampling.cpp		diff \| blob \| history
tests/test-sampling.cpp		diff \| blob \| history