]> git.djapps.eu Git - pkg/ggml/sources/whisper.cpp/shortlog
pkg/ggml/sources/whisper.cpp
2026-06-02 Noah Lyonsserver : merge split utf-8 token text in verbose json...
2026-06-02 Patrice Levesquecmake : do not assume /usr/lib library installation...
2026-06-01 Georgi Gerganovrelease : v1.8.6
2026-06-01 Daniel Beveniusci : fix path to whisper.h in examples.yml [no ci]...
2026-05-31 Georgi Gerganovci : fix self-hosted paths to mnt
2026-05-31 Georgi Gerganovpi : add config
2026-05-31 Georgi Gerganovci : remove obsolete self-hosted label
2026-05-31 Georgi Gerganovcommon : pass sample rate to `ffmpeg_decode_audio()`
2026-05-31 Georgi Gerganovcommon : re-implement `ffmpeg-transcode.cpp` + clarify...
2026-05-29 Georgi Gerganovsync : ggml
2026-05-29 Georgi Gerganovggml : bump version to 0.13.1 (ggml/1523)
2026-05-29 Georgi Gerganovtalk-llama : sync llama.cpp
2026-05-29 Georgi Gerganovsync : ggml
2026-05-29 Andreas Kieslingercuda : disables launch_fattn PDL enrollment due to...
2026-05-29 Matt Corallometa : Add missing `buffer` set in allreduce fallback...
2026-05-29 Max Krasnyanskyhexagon: basic/generic op fusion support and RMS_NORM...
2026-05-29 lhezopencl: move backend info printing into its own functio...
2026-05-29 fl0rianrggml: auto apply iGPU flag CUDA/HIP if integrated devic...
2026-05-29 redfoxmmvq Optim: add MMVQ_PARAMETERS_TURING(mmvq_parameter_t...
2026-05-29 Jaden_MachCUDA: route batch>=4 quantized matmul to MMQ on AMD...
2026-05-29 Max Krasnyanskyhexagon: minor refresh for HMX FA and MM (llama/23796)
2026-05-29 Jeff Bolzvulkan: fast path for walsh-hadamard transform (llama...
2026-05-29 Winston Mavulkan: fix wrong index variable in inner loop (llama...
2026-05-29 Winston Mavulkan: Fix memory logger unsafe iterator access (llama...
2026-05-29 fairydreamingcuda : fix KQ mask offset integer overflow in fattn...
2026-05-29 Martin Klacerggml: fixed Arm SVE usage bug in vec.h, vec.cpp (llama...
2026-05-29 ymckiHexagon: OP_GATED_DELTA_NET K>1 support (llama/23531)
2026-05-29 ymckiopencl: OP_GATED_DELTA_NET (llama/23312)
2026-05-29 Reese Levineggml-webgpu: remove legacy constants (llama/23672)
2026-05-29 Max Krasnyanskyhexagon: add support for Q4_1 in MUL_MAT and MUL_MAT_ID...
2026-05-29 Masashi Yoshimuraggml-webgpu: Fix how to dispatch WG to some ops (llama...
2026-05-29 Matt Corallovulkan: Switch MUL_MAT_VEC to 4 K per iteration for...
2026-05-29 Jeff Bolzvulkan: use GL_NV_cooperative_matrix_decode_vector...
2026-05-29 l8bloomvulkan: add REPEAT op support for f16 to f16. (llama...
2026-05-29 Oliver SimonsCUDA: restrict PDL to CTK >= 12.3 due to MSVC issues...
2026-05-29 Winston Mavulkan: avoid preferring transfer queue on AMD UMA...
2026-05-29 Vladislavggml-zendnn : fixed naming of matmul function (llama...
2026-05-29 Jeff Bolzvulkan: optimize conv2d and implement coopmat1 support...
2026-05-29 Max Krasnyanskyhexagon: add support for CONCAT op (llama/23648)
2026-05-29 Alexey KopytkoSYCL: implement ggml_sycl_pool_vmm (llama/22862)
2026-05-29 Masashi Yoshimuraggml-webgpu: Add MMVQ path for Q4/Q8/Q2_K/Q4_K and...
2026-05-29 Nikhil JainCheck batch_compute_passes before sending passes when...
2026-05-29 Johannes GäßlerCUDA: missing PDL sync for FWHT, better fallback (llama...
2026-05-29 forforever73metal : add apple device id (llama/23566)
2026-05-29 Aman GuptaCUDA: add fast walsh-hadamard transform (llama/23615)
2026-05-28 Daniel Beveniusci : add ignore for bindings/{ruby, go} in build.yml...
2026-05-28 Daniel Beveniusci : fix include paths for bindings-go job [no ci]...
2026-05-28 Daniel Beveniusci : add on push/pull_request paths ruby job (#3833)
2026-05-28 Daniel Beveniusci : renable arm64 docker builds (#3832)
2026-05-28 Daniel Beveniusci : set GGML_NATIVE=OFF for bindings-java (#3830)
2026-05-27 Daniel Beveniusci : only run docker jobs when pushed to master [no...
2026-05-27 Daniel Beveniusdocs : add AGENTS.md and CONTRIBUTING.md [no ci] (...
2026-05-26 texasichcli : merge tokens split across UTF-8 boundaries in...
2026-05-25 Georgi Gerganovrelease : v1.8.5
2026-05-25 Georgi Gerganovbenches : update
2026-05-25 Georgi Gerganovsync : ggml
2026-05-25 Georgi Gerganovggml : bump version to 0.13.0 (ggml/1510)
2026-05-25 Johannes GäßlerTP: fix ggml context size calculation (llama/22616)
2026-05-25 Gilad Sggml: `gguf_init_from_callback` and `gguf_init_from_buf...
2026-05-25 Kaihui-AMDreadme : add AMD ROCm/HIP GPU build instructions (...
2026-05-25 Georgi Gerganovtalk-llama : sync llama.cpp
2026-05-25 Georgi Gerganovsync : ggml
2026-05-25 Georgi Gerganovggml : bump version to 0.12.1 (ggml/1508)
2026-05-25 Jeff Bolzggml : Parallelize quant LUT init (llama/23595)
2026-05-25 Johannes GäßlerTP: fix entirely zero-sized slices per device (llama...
2026-05-25 shaofeiqiopencl: batch profiling to improve speed and prevent...
2026-05-25 Yiwei Shaohexagon: apply repl optimization in flash attn softmax...
2026-05-25 dskweggml : Check the right iface method before using the...
2026-05-25 Jeff Bolzvulkan: fix windows find_package of SPIRV-Headers ...
2026-05-25 Shawn Guopencl: generalize Adreno MoE kernels on M (llama/23449)
2026-05-25 Alexey KopytkoSYCL: improve MoE prefill throughput (llama/23142)
2026-05-25 Alexey Kopytkosycl : Level Zero detection in ggml_sycl_init (llama...
2026-05-25 karavayevSYCL : gated_delta_net K>1 (llama/23174)
2026-05-25 KatostrofikSYCL: add BF16 to DMMV kernel path (~4x tg speedup...
2026-05-25 Sachin Sharmaggml-zendnn : add Q8_0 quantization support (llama...
2026-05-25 Johannes GäßlerCUDA: fix PDL CC check for JIT compilation (llama/23471)
2026-05-25 Pascalvulkan: fuse snake activation (mul, sin, sqr, mul,...
2026-05-25 Chen Yuanfix(flash-attn): replace f32 with kv_type and q_type...
2026-05-25 Georgi Gerganovmetal : optimize concat kernel and fix set kernel threa...
2026-05-25 Matt Coralloggml : Check the right iface method before using the...
2026-05-25 Todor Boinovskihexagon: ssm-conv fix for large prompts (llama/23307)
2026-05-25 lhezopencl: refactor backend initilization (llama/23318)
2026-05-25 Danielevulkan: optimize operations in the IM2COL shader (llama...
2026-05-25 Max Krasnyanskyhexagon: HMX quantized matmul rework (llama/23368)
2026-05-25 Andreas KieslingerProgrammatic Dependent Launch (PDL) for more performanc...
2026-05-25 Georgi Gerganovmetal : optimize pad + cpy (llama/23354)
2026-05-25 ravel7524ggml-cuda: tune RDNA3 Q6_K MMVQ nwarps (llama/23349)
2026-05-25 shaofeiqiopencl: add MoE support for q4_k, q5_k, q6_k on Adreno...
2026-05-25 Aparna M Phexagon: add MROPE and IMROPE support in HTP rope op...
2026-05-25 Aparna M Phexagon: enable support for NORM op (llama/23319)
2026-05-25 Reese Levineggml-webgpu : extend GDN for K>1 (llama/23299)
2026-05-25 Intel AI Get... sycl: add GGML_SYCL_USE_ASYNC_MEM_OP env toggle (llama...
2026-05-25 Radoslav Gerganovrpc : keep last_graph_uid in the device context (llama...
2026-05-25 Pranav Dhinakarhexagon: add support for TRI op (llama/22822)
2026-05-25 Pranav Dhinakarggml-hexagon: add PAD op HVX kernel (llama/23078)
2026-05-25 Intel AI Get... sycl: scalar SWAR byte-subtract in Q6_K MMVQ dot produc...
2026-05-25 Intel AI Get... sycl: route small f32 matmuls to oneMKL, bypass oneDNN...
2026-05-25 Gabe Goodhartfeat: Support d_conv=15 for ssm-conv.cu (llama/23017)
2026-05-25 Oliver SimonsCUDA: Continue directly including cuda/iterator (llama...
2026-05-25 Jan Ekströmggml-vulkan/CMakeLists: add a check for SPIRV-Headers...
next