git.djapps.eu Git - pkg/ggml/sources/llama.cpp/commit

author	Piotr Wilkin (ilintar) <redacted>
	Fri, 28 Nov 2025 11:02:56 +0000 (12:02 +0100)
committer	GitHub <redacted>
	Fri, 28 Nov 2025 11:02:56 +0000 (12:02 +0100)
commit	ff55414c42522adbeaa1bd9c52c0e9db16942484
tree	b783bd9a92bd7b716f817458d915862c92fcdab0	tree
parent	73955f7d2a3ce1f36d7ecc14495e08957b51d113	commit \| diff

model : Qwen3 Next (#16095)

* Qwen3 Next - cleaned up version

* Whitespaces and stuff

* Correct minor errors

* Update src/llama-model.cpp

Co-authored-by: Sigbjørn Skjæret <redacted>
* Misc. fixes.

* Clean up code, add missing hybrid qualifier

* Did someone transpose the SOLVE_TRI result matrix? Perhaps...

* Whitespace

* Proper tensors for cb calls

* Use llama-graph.h vertical alignment

* BROKEN: chunking

* Set new tensors as inputs.

* Proper chunk logic

* It's the circle of life...

* More shenanigans for n_seq > 1

* Nail in the coffin?

* Fix Windows build

* Eh, one fails on Windows, the other fails on Mac... just use general capture.

* quant : cleanup

* model : cleanup

* qwen3 : cleanup

* cont : cleanup

* cont : cleanup

* ggml : revert change

* qwen3 : cleanup

* cont : cleanup

* Readd cmath

* qwen3 : fix typo

* Update convert_hf_to_gguf.py

Co-authored-by: Sigbjørn Skjæret <redacted>
* Usual suspects

* fix my bad suggestion

---------

Co-authored-by: Sigbjørn Skjæret <redacted>
Co-authored-by: Georgi Gerganov <redacted>

convert_hf_to_gguf.py		diff \| blob \| history
examples/model-conversion/scripts/causal/run-converted-model.sh		diff \| blob \| history
examples/model-conversion/scripts/causal/run-org-model.py		diff \| blob \| history
ggml/src/ggml-cpu/ops.cpp		diff \| blob \| history
gguf-py/gguf/constants.py		diff \| blob \| history
gguf-py/gguf/tensor_mapping.py		diff \| blob \| history
src/CMakeLists.txt		diff \| blob \| history
src/llama-arch.cpp		diff \| blob \| history
src/llama-arch.h		diff \| blob \| history
src/llama-context.cpp		diff \| blob \| history
src/llama-hparams.h		diff \| blob \| history
src/llama-model.cpp		diff \| blob \| history
src/llama-model.h		diff \| blob \| history
src/llama-quant.cpp		diff \| blob \| history
src/models/models.h		diff \| blob \| history
src/models/qwen3next.cpp	[new file with mode: 0644]	blob