..
cmake
Merge vulkan code from mainline up to commit of 6/28/2025 ( #563 )
2025-07-02 08:49:42 +02:00
ggml-cann
Merge mainline - Aug 12 2024 ( #17 )
2024-08-12 15:14:32 +02:00
ggml-cuda
CUDA: faster prompt processing for 4-bit quants ( #713 )
2025-08-21 15:57:35 +03:00
ggml-sycl
Merge mainline - Aug 12 2024 ( #17 )
2024-08-12 15:14:32 +02:00
iqk
Fix q8_0 repacking issues on AVX2 ( #708 )
2025-08-19 19:49:58 +03:00
kompute @ 4565194ed7
Merge mainline llama.cpp ( #3 )
2024-07-27 07:55:01 +02:00
kompute-shaders
Merge mainline llama.cpp ( #3 )
2024-07-27 07:55:01 +02:00
llamafile
Merge mainline llama.cpp ( #3 )
2024-07-27 07:55:01 +02:00
vulkan-shaders
Vulkan: a fresh start ( #608 )
2025-07-15 08:03:13 +02:00
CMakeLists.txt
Vulkan: add cmake options to build without coopmat(2) support ( #674 )
2025-08-07 17:26:21 +03:00
ggml-aarch64.c
Merge mainline - Aug 12 2024 ( #17 )
2024-08-12 15:14:32 +02:00
ggml-aarch64.h
Merge mainline llama.cpp ( #3 )
2024-07-27 07:55:01 +02:00
ggml-alloc.c
Enable CUDA graphs for MoE models + GPT-OSS support ( #689 )
2025-08-15 09:18:07 +03:00
ggml-backend-impl.h
Merge vulkan code from mainline up to commit of 6/28/2025 ( #563 )
2025-07-02 08:49:42 +02:00
ggml-backend.c
Fix debug build failure with RPC off ( #579 )
2025-07-03 15:26:28 +02:00
ggml-blas.cpp
Merge mainline - Aug 12 2024 ( #17 )
2024-08-12 15:14:32 +02:00
ggml-cann.cpp
Merge vulkan code from mainline up to commit of 6/28/2025 ( #563 )
2025-07-02 08:49:42 +02:00
ggml-common.h
MXFP4 ( #682 )
2025-08-09 08:40:18 +03:00
ggml-cuda.cu
Enable CUDA graphs for MoE models + GPT-OSS support ( #689 )
2025-08-15 09:18:07 +03:00
ggml-impl.h
MXFP4 ( #682 )
2025-08-09 08:40:18 +03:00
ggml-kompute.cpp
Merge vulkan code from mainline up to commit of 6/28/2025 ( #563 )
2025-07-02 08:49:42 +02:00
ggml-metal.m
MXFP4 ( #682 )
2025-08-09 08:40:18 +03:00
ggml-metal.metal
MXFP4 ( #682 )
2025-08-09 08:40:18 +03:00
ggml-quants.c
MXFP4 ( #682 )
2025-08-09 08:40:18 +03:00
ggml-quants.h
IQ1_M_R4: better 1.75 bpw quants ( #187 )
2025-02-06 14:08:52 +02:00
ggml-rpc.cpp
Merge vulkan code from mainline up to commit of 6/28/2025 ( #563 )
2025-07-02 08:49:42 +02:00
ggml-sycl.cpp
Merge vulkan code from mainline up to commit of 6/28/2025 ( #563 )
2025-07-02 08:49:42 +02:00
ggml-vulkan.cpp
Vulkan: a fresh start ( #608 )
2025-07-15 08:03:13 +02:00
ggml.c
Revert "Better CPU prompt processing performance for SWA models ( #696 )" ( #701 )
2025-08-17 15:44:02 +03:00