]> git.djapps.eu Git - pkg/ggml/sources/llama.cpp/commit
llama : add PLaMo model (#3557)
authorShintarou Okada <redacted>
Sun, 24 Dec 2023 13:35:49 +0000 (22:35 +0900)
committerGitHub <redacted>
Sun, 24 Dec 2023 13:35:49 +0000 (15:35 +0200)
commit753be377b69bda2d65a7e089f2b7f0c53ef3495e
treeb32ae0b6fb10db974322edeeb22021bc43d1e210
parent5bf3953d7e9831ea22b0bc017ce97409b801ccf1
llama : add PLaMo model (#3557)

* add plamo mock

* add tensor loading

* plamo convert

* update norm

* able to compile

* fix norm_rms_eps hparam

* runnable

* use inp_pos

* seems ok

* update kqv code

* remove develop code

* update README

* shuffle attn_q.weight and attn_output.weight for broadcasting

* remove plamo_llm_build_kqv and use llm_build_kqv

* fix style

* update

* llama : remove obsolete KQ_scale

* plamo : fix tensor names for correct GPU offload

---------

Co-authored-by: Georgi Gerganov <redacted>
README.md
convert-hf-to-gguf.py
gguf-py/gguf/constants.py
gguf-py/gguf/tensor_mapping.py
llama.cpp