]> git.djapps.eu Git - pkg/ggml/sources/llama.cpp/commit
llama : add llama_model_ftype_name() (#25134)
authorAdrien Gallouët <redacted>
Thu, 2 Jul 2026 15:26:47 +0000 (17:26 +0200)
committerGitHub <redacted>
Thu, 2 Jul 2026 15:26:47 +0000 (17:26 +0200)
commitfdb1db877c526ec90f668eca1b858da5dba85560
tree30a695788a267f66ba277b330926c3b13d6cabfa
parent4fc4ec5541b243957ae5099edb67372f8f3b550e
llama : add llama_model_ftype_name() (#25134)

* llama : add llama_model_ftype_name()

Expose the model file type (quantization) name, e.g. "Q8_0" or
"Q4_K - Medium", through a new public C API. The returned pointer is
valid for the lifetime of the model and nullptr when the model is
invalid or the file type is unknown.

Signed-off-by: Adrien Gallouët <redacted>
* Export enum

Signed-off-by: Adrien Gallouët <redacted>
* s/llama_model_ftype_name/llama_ftype_name/

Signed-off-by: Adrien Gallouët <redacted>
* Move "(guessed)" to the front in llama_ftype_name

Prepend the "(guessed)" label instead of appending it. This allows removing
the non-thread-safe static std::string, making the function allocation-free.

Signed-off-by: Adrien Gallouët <redacted>
* Add LLAMA_FTYPE_PREFIX

Signed-off-by: Adrien Gallouët <redacted>
* Dont check for model

Signed-off-by: Adrien Gallouët <redacted>
---------

Signed-off-by: Adrien Gallouët <redacted>
include/llama.h
src/llama-model-loader.cpp
src/llama-model.cpp
src/llama-model.h
tools/cli/cli.cpp
tools/server/server-context.cpp
tools/server/server-context.h