Nexes the Elder
6198a356a8
Remove deprecated Kompute (Vulkan compute) backend ( #2097 )
...
* Remove broken kompute submodule (ghost - nulled config, corrupted tracking)
The kompute submodule at ggml/src/kompute had its .git/modules/kompute/config
completely zeroed out (null bytes). The submodule was non-functional and is
not used in this fork. Removed:
- .gitmodules entry
- .git/config [submodule kompute] section
- .git/modules/kompute directory
- ggml/src/kompute working tree
* Extensive removal of all Kompute code and references
Removed the entire Kompute Vulkan compute backend which was
unmaintained and superseded by the Vulkan backend:
Files deleted:
- ggml/src/ggml-kompute.cpp (Vulkan compute backend implementation)
- ggml/include/ggml-kompute.h (header)
- ggml/src/kompute-shaders/ (34 SPIR-V shader source files)
Build system:
- ggml/CMakeLists.txt: removed GGML_KOMPUTE option
- ggml/src/CMakeLists.txt: removed compile_shader function, submodule
add, shader compilation, stamp targets, and all KOMPUTE source refs
- CMakeLists.txt: removed LLAMA_KOMPUTE deprecation alias
Source code:
- ggml/src/ggml-backend.cpp: removed kompute reg decl and call
- ggml/include/ggml.h: removed ggml_cpu_has_kompute() declaration
- ggml/src/ggml.c: removed ggml_cpu_has_kompute() implementation
and its reference in ggml_cpu_has_gpublas()
- src/llama.cpp: removed #include, backend init, buffer type, model
loading guard, and GPU offload check for Kompute
- src/llama-model-loader.cpp: removed kompute include
- common/common.cpp: removed cpu_has_kompute print
- tests/test-c.c: removed kompute include guard
- examples/llama-bench/llama-bench.cpp: removed kompute member,
construction, field serialization, and display string
- scripts/compare-llama-bench.py: removed kompute from key props,
bool props, and pretty names
- scripts/sync-ggml.sh: removed kompute file copy lines
- scripts/sync-ggml-am.sh: removed kompute path mappings
Git submodule:
- .gitmodules: removed kompute entry
- .git/config: removed [submodule kompute] section
- .git/modules/kompute: removed
- ggml/src/kompute: removed (working tree)
2026-07-08 10:01:01 +02:00
Kawrakow
1a4cfbcc53
Merge mainline - Aug 12 2024 ( #17 )
...
* Merge mainline
* Fix after merge
* Remove CI check
---------
Co-authored-by: Iwan Kawrakow <iwan.kawrakow@gmail.com>
2024-08-12 15:14:32 +02:00
Kawrakow
0ceeb11721
Merge mainline llama.cpp ( #3 )
...
* Merging mainline - WIP
* Merging mainline - WIP
AVX2 and CUDA appear to work.
CUDA performance seems slightly (~1-2%) lower as it is so often
the case with llama.cpp/ggml after some "improvements" have been made.
* Merging mainline - fix Metal
* Remove check
---------
Co-authored-by: Iwan Kawrakow <iwan.kawrakow@gmail.com>
2024-07-27 07:55:01 +02:00
Georgi Gerganov
8de006f83e
ggml : remove OpenCL ( #7735 )
...
ggml-ci
2024-06-04 21:23:20 +03:00
Georgi Gerganov
556fc986b2
scripts : remove mpi remnants
2024-05-29 14:31:18 +03:00
Georgi Gerganov
fd3a0965b2
script : sync ggml-rpc
2024-05-14 19:14:38 +03:00
Georgi Gerganov
41c01483dd
license : update copyright notice + add AUTHORS ( #6405 )
...
* license : add AUTHORS
* authors : update
* scipts : add LICENSE and gen-authors.sh to sync
2024-04-09 09:23:19 +03:00
Georgi Gerganov
537fc022b8
sync : ggml ( #6351 )
...
* sync : ggml
ggml-ci
* cuda : move GGML_CUDA_DMMV constants to dmmv.cuh
---------
Co-authored-by: slaren <slarengh@gmail.com>
2024-03-29 17:45:46 +02:00
Georgi Gerganov
1cad57678f
ggml : add ggml-common.h to deduplicate shared code ( #5940 )
...
* ggml : add ggml-common.h to shared code
ggml-ci
* scripts : update sync scripts
* sycl : reuse quantum tables
ggml-ci
* ggml : minor
* ggml : minor
* sycl : try to fix build
2024-03-09 12:47:57 +02:00
Georgi Gerganov
71e5730e05
scripts : update sync scripts with new backends
2024-02-10 09:53:05 +02:00
Georgi Gerganov
8a8220f13a
sync : ggml (new ops, tests, backend, etc.) ( #4359 )
...
* sync : ggml (part 1)
* sync : ggml (part 2, CUDA)
* sync : ggml (part 3, Metal)
* ggml : build fixes
ggml-ci
* cuda : restore lost changes
* cuda : restore lost changes (StableLM rope)
* cmake : enable separable compilation for CUDA
ggml-ci
* ggml-cuda : remove device side dequantize
* Revert "cmake : enable separable compilation for CUDA"
This reverts commit 09e35d04b1c4ca67f9685690160b35bc885a89ac.
* cuda : remove assert for rope
* tests : add test-backend-ops
* ggml : fix bug in ggml_concat
* ggml : restore `ggml_get_n_tasks()` logic in `ggml_graph_plan()`
* ci : try to fix macOS
* ggml-backend : remove backend self-registration
* ci : disable Metal for macOS cmake build
ggml-ci
* metal : fix "supports family" call
* metal : fix assert
* metal : print resource path
ggml-ci
---------
Co-authored-by: slaren <slarengh@gmail.com>
2023-12-07 22:26:54 +02:00
Georgi Gerganov
ddea57dbe3
sync : ggml (backend v2) ( #3912 )
...
* sync : ggml (backend v2) (wip)
* sync : migrate examples and llama.cpp to dynamic graphs (wip)
* sync : update tests + fix max op params to 64
ggml-ci
* sync : ggml-cuda
ggml-ci
* llama : fix save/load state context size
ggml-ci
* sync : try to fix build on tvOS
* sync : pass custom graph sizes in training examples
* sync : update graph copies to new ggml API
* sync : update sync-ggml.sh with new files
* scripts : fix header in sync script
* train : fix context size calculations
* llama : increase inference graph size up to 4096 nodes
* train : allocate grads for backward graphs
* train : allocate grads for gb_tmp
2023-11-13 14:16:23 +02:00
Georgi Gerganov
78b3d9b796
sync : ggml (ggml-backend) ( #3548 )
...
* sync : ggml (ggml-backend)
ggml-ci
* zig : add ggml-backend to the build
2023-10-08 20:19:14 +03:00
Georgi Gerganov
cd3ea9a1aa
ggml : sync latest (SAM + SD operators, CUDA alibi) ( #2709 )
...
* ggml : sync latest (SAM + SD operators, CUDA alibi)
ggml-ci
* ggml : fix tabs
2023-08-22 14:22:08 +03:00
Eve
31e5af188f
tests : Fix compilation warnings (Linux/GCC) ( #2451 )
...
* fix hellaswag print format, cast away warning in test-double-float
* c++11 cannot use designated initializers
* add static to test-grad0.c internal functions
* use memcpy in test-double-float.c
* port c tests to c++
* use initializer list for ggml_init_params
2023-08-02 11:06:19 +03:00
Georgi Gerganov
5ac2fe8ccf
tests : fix test-grad0
2023-07-05 20:20:25 +03:00
Georgi Gerganov
40c6a525ea
ggml : sync latest (new ops, macros, refactoring) ( #2106 )
...
- add ggml_argmax()
- add ggml_tanh()
- add ggml_elu()
- refactor ggml_conv_1d() and variants
- refactor ggml_conv_2d() and variants
- add helper macros to reduce code duplication in ggml.c
2023-07-04 21:54:11 +03:00
Georgi Gerganov
3dbec7fb22
scripts : add helper scripts to synch ggml repo
2023-04-23 19:57:09 +03:00