git.djapps.eu Git - pkg/ggml/sources/llama.cpp/commit

author	Shintarou Okada <redacted>
	Sun, 24 Dec 2023 13:35:49 +0000 (22:35 +0900)
committer	GitHub <redacted>
	Sun, 24 Dec 2023 13:35:49 +0000 (15:35 +0200)
commit	753be377b69bda2d65a7e089f2b7f0c53ef3495e
tree	b32ae0b6fb10db974322edeeb22021bc43d1e210	tree
parent	5bf3953d7e9831ea22b0bc017ce97409b801ccf1	commit \| diff

llama : add PLaMo model (#3557)

* add plamo mock

* add tensor loading

* plamo convert

* update norm

* able to compile

* fix norm_rms_eps hparam

* runnable

* use inp_pos

* seems ok

* update kqv code

* remove develop code

* update README

* shuffle attn_q.weight and attn_output.weight for broadcasting

* remove plamo_llm_build_kqv and use llm_build_kqv

* fix style

* update

* llama : remove obsolete KQ_scale

* plamo : fix tensor names for correct GPU offload

---------

Co-authored-by: Georgi Gerganov <redacted>

README.md		diff \| blob \| history
convert-hf-to-gguf.py		diff \| blob \| history
gguf-py/gguf/constants.py		diff \| blob \| history
gguf-py/gguf/tensor_mapping.py		diff \| blob \| history
llama.cpp		diff \| blob \| history