This website requires JavaScript.
Explore
Help
Sign In
nicolas
/
ik_llama_opt
Watch
1
Star
0
Fork
You've already forked ik_llama_opt
0
Code
Issues
Pull Requests
Packages
Projects
Releases
Wiki
Activity
ff141691a1
ik_llama_opt
/
ggml
History
Kawrakow
ff141691a1
Use f32 accumulation in CUDA DSA implementation (
#2311
)
2026-08-13 15:22:01 +02:00
..
cmake
Merge mainline llama.cpp (
#3
)
2024-07-27 07:55:01 +02:00
include
DS4: faster long-context TG (
#2201
)
2026-07-30 13:13:42 +03:00
src
Use f32 accumulation in CUDA DSA implementation (
#2311
)
2026-08-13 15:22:01 +02:00
.gitignore
Merge mainline llama.cpp (
#3
)
2024-07-27 07:55:01 +02:00
CMakeLists.txt
Chunked experts (CPU) (
#2202
)
2026-07-30 13:16:02 +03:00
Powered by
TurnKey Linux
.