]> git.djapps.eu Git - pkg/ggml/sources/whisper.cpp/shortlog
pkg/ggml/sources/whisper.cpp
2026-06-08 Georgi Gerganovtalk-llama : sync llama.cpp
2026-06-08 Georgi Gerganovsync : ggml
2026-06-08 Georgi Gerganovggml : bump version to 0.14.0 (ggml/1533)
2026-06-08 Georgi Gerganovsync : ggml
2026-06-08 Harkirat GillHIP: add gfx1152 and gfx1153 to RDNA3.5 (llama/24129)
2026-06-08 Xuan-Son Nguyenmetal : fix im2col 1D case (audio models) (llama/24220)
2026-06-08 Ruben Ortlamvulkan: check coopmat2 features before reporting suppor...
2026-06-08 lhezopencl: improve get_rows, cpy, concat and q6_k flat...
2026-06-08 Ruben Ortlamvulkan: add fwht support for Intel with shmem reduction...
2026-06-08 Charles Xukleidiai : dynamic chunck-based scheduling for hybrid...
2026-06-08 Oliver SimonsCUDA: enroll mul_mat_vec_q_moe into pdl (llama/24087)
2026-06-08 Mason Milburnsycl : port multi-column MMVQ from CUDA backend (llama...
2026-06-08 Kartik Sirohiggml: vectorize ggml_vec_dot_q4_1_q8_1 with WASM SIMD12...
2026-06-08 Georgi Gerganovmetal : reduce rset heartbeat from 500ms -> 5ms (llama...
2026-06-08 Reese Levineggml-webgpu: FlashAttention refactor + standardize...
2026-06-08 rehan-10xengineerggml-cpu: extend RVV quantization vec dot to higher...
2026-06-08 Andreas KieslingerAvoid PDL race conditions by disabling __restrict__...
2026-06-08 Charles Xuggml-cpu: use runtime SVE width in FWHT (llama/24059)
2026-06-08 Aman Guptacuda: reserve space for quantize kv-cache at startup...
2026-06-08 lhezopencl: use flat variants of q4_K and q6_K gemv for...
2026-06-08 Max Krasnyanskyhexagon: profiler output fix and script updates (llama...
2026-06-08 Max Krasnyanskyhexagon: MUL_MAT, MUL_MAT_ID, FLASH_ATTN and GDN cleanu...
2026-06-08 Todor Boinovskihexagon: add gelu_quick (llama/24007)
2026-06-08 Anav Prasadclean up unused variables warnings (llama/23975)
2026-06-08 lhezopencl: fix compiler warnings for non-adreno path ...
2026-06-08 Masashi Yoshimurarevert to using global_invocation_id for cpy shader...
2026-06-08 shaofeiqiopencl: add basic support for q5_0 and q5_1 (llama...
2026-06-08 Shrivas Shankarmetal: template GLU kernels to support f16/f32 (llama...
2026-06-08 Jeff Bolzvulkan: don't hold the device mutex while compiling...
2026-06-08 Winston Mavulkan: reduce host memory lock contention (llama/23376)
2026-06-08 Johannes GäßlerTP: quantized KV cache support (llama/23792)
2026-06-08 Matt Corallovulkan: Block-load Q3_K/Q6_K block data and subtract...
2026-06-08 Winston Mavulkan: Removed unused functions (llama/23175)
2026-06-08 Neo ZhangSupport Q4_1, Q5_0, Q5_1 in Flash-attention (llama...
2026-06-08 Neo ZhangAdd more types in GET_ROWS OP (llama/23710)
2026-06-08 Neo Zhangsycl : Optimize Q3_K mul_mat by reorder (llama/23725)
2026-06-08 lhezopencl: support bf16 by converting to f16 (llama/23839)
2026-06-08 Georgi Gerganovmetal : restore im2col implementation for large kernels...
2026-06-08 Jinyang Heggml : add some lsx support (llama/23798)
2026-06-08 Ruben Ortlamvulkan: add Flash Attention support for BFloat16 KV...
2026-06-08 Reese Levineggml-webgpu: Check earlier for WebGPU required features...
2026-06-08 Reese Levineggml-webgpu: add q4_0/q8_0 SET_ROWS (llama/23760)
2026-06-08 Oliver SimonsCUDA: Check PTX version on host side to guard PDL dispa...
2026-06-08 fairydreamingmodel : support for DeepseekV32ForCausalLM with generic...
2026-06-08 Daniel Beveniusci : add ccache to build-sycl [no ci] (#3859)
2026-06-06 Daniel Beveniusci : add HF_TOKEN to docker.yml workflow [no ci] (...
2026-06-06 Daniel Beveniusci : add ccache to quantize, vad, and wasm jobs (#3860)
2026-06-04 Daniel Beveniusci: build-windows action slimming (#3858)
2026-06-04 Daniel Beveniusci : use emscripten-core and pin version (#3857)
2026-06-04 Daniel Beveniusci : pin github actions to commit SHAs (#3856)
2026-06-04 Daniel Beveniusci : use ccache instead of sccache for windows-cublas...
2026-06-04 Daniel Beveniusci : only publish/push docker images daily (#3854)
2026-06-04 Georgi Gerganovci : refactor + optimize (#3847)
2026-06-02 danscMaxwhisper : catch C++ exceptions in whisper_init_with_par...
2026-06-02 Noah Lyonsserver : merge split utf-8 token text in verbose json...
2026-06-02 Patrice Levesquecmake : do not assume /usr/lib library installation...
2026-06-01 Georgi Gerganovrelease : v1.8.6
2026-06-01 Daniel Beveniusci : fix path to whisper.h in examples.yml [no ci]...
2026-05-31 Georgi Gerganovci : fix self-hosted paths to mnt
2026-05-31 Georgi Gerganovpi : add config
2026-05-31 Georgi Gerganovci : remove obsolete self-hosted label
2026-05-31 Georgi Gerganovcommon : pass sample rate to `ffmpeg_decode_audio()`
2026-05-31 Georgi Gerganovcommon : re-implement `ffmpeg-transcode.cpp` + clarify...
2026-05-29 Georgi Gerganovsync : ggml
2026-05-29 Georgi Gerganovggml : bump version to 0.13.1 (ggml/1523)
2026-05-29 Georgi Gerganovtalk-llama : sync llama.cpp
2026-05-29 Georgi Gerganovsync : ggml
2026-05-29 Andreas Kieslingercuda : disables launch_fattn PDL enrollment due to...
2026-05-29 Matt Corallometa : Add missing `buffer` set in allreduce fallback...
2026-05-29 Max Krasnyanskyhexagon: basic/generic op fusion support and RMS_NORM...
2026-05-29 lhezopencl: move backend info printing into its own functio...
2026-05-29 fl0rianrggml: auto apply iGPU flag CUDA/HIP if integrated devic...
2026-05-29 redfoxmmvq Optim: add MMVQ_PARAMETERS_TURING(mmvq_parameter_t...
2026-05-29 Jaden_MachCUDA: route batch>=4 quantized matmul to MMQ on AMD...
2026-05-29 Max Krasnyanskyhexagon: minor refresh for HMX FA and MM (llama/23796)
2026-05-29 Jeff Bolzvulkan: fast path for walsh-hadamard transform (llama...
2026-05-29 Winston Mavulkan: fix wrong index variable in inner loop (llama...
2026-05-29 Winston Mavulkan: Fix memory logger unsafe iterator access (llama...
2026-05-29 fairydreamingcuda : fix KQ mask offset integer overflow in fattn...
2026-05-29 Martin Klacerggml: fixed Arm SVE usage bug in vec.h, vec.cpp (llama...
2026-05-29 ymckiHexagon: OP_GATED_DELTA_NET K>1 support (llama/23531)
2026-05-29 ymckiopencl: OP_GATED_DELTA_NET (llama/23312)
2026-05-29 Reese Levineggml-webgpu: remove legacy constants (llama/23672)
2026-05-29 Max Krasnyanskyhexagon: add support for Q4_1 in MUL_MAT and MUL_MAT_ID...
2026-05-29 Masashi Yoshimuraggml-webgpu: Fix how to dispatch WG to some ops (llama...
2026-05-29 Matt Corallovulkan: Switch MUL_MAT_VEC to 4 K per iteration for...
2026-05-29 Jeff Bolzvulkan: use GL_NV_cooperative_matrix_decode_vector...
2026-05-29 l8bloomvulkan: add REPEAT op support for f16 to f16. (llama...
2026-05-29 Oliver SimonsCUDA: restrict PDL to CTK >= 12.3 due to MSVC issues...
2026-05-29 Winston Mavulkan: avoid preferring transfer queue on AMD UMA...
2026-05-29 Vladislavggml-zendnn : fixed naming of matmul function (llama...
2026-05-29 Jeff Bolzvulkan: optimize conv2d and implement coopmat1 support...
2026-05-29 Max Krasnyanskyhexagon: add support for CONCAT op (llama/23648)
2026-05-29 Alexey KopytkoSYCL: implement ggml_sycl_pool_vmm (llama/22862)
2026-05-29 Masashi Yoshimuraggml-webgpu: Add MMVQ path for Q4/Q8/Q2_K/Q4_K and...
2026-05-29 Nikhil JainCheck batch_compute_passes before sending passes when...
2026-05-29 Johannes GäßlerCUDA: missing PDL sync for FWHT, better fallback (llama...
2026-05-29 forforever73metal : add apple device id (llama/23566)
2026-05-29 Aman GuptaCUDA: add fast walsh-hadamard transform (llama/23615)
2026-05-28 Daniel Beveniusci : add ignore for bindings/{ruby, go} in build.yml...
next