ik_llama_opt/src
Nexes the Elder 2b1af6bade CLI - Specify GGML_TYPE to quantize for the main tensors. (#91)
To complement the token_embd.weight and output.weight :

attn_v.weight
attn_k.weight.
attn_q_weight
attn_output.weight
attn_qkv.weight
ffn_gate
ffn_down
ffn_up
2024-10-18 09:48:15 +02:00
..
CMakeLists.txt
llama-grammar.cpp Merge mainline - Aug 12 2024 (#17) 2024-08-12 15:14:32 +02:00
llama-grammar.h
llama-impl.h Time to fix replace_all (#68) 2024-09-28 17:59:47 +03:00
llama-sampling.cpp
llama-sampling.h
llama-vocab.cpp Merge mainline - Aug 12 2024 (#17) 2024-08-12 15:14:32 +02:00
llama-vocab.h Merge mainline - Aug 12 2024 (#17) 2024-08-12 15:14:32 +02:00
llama.cpp CLI - Specify GGML_TYPE to quantize for the main tensors. (#91) 2024-10-18 09:48:15 +02:00
unicode-data.cpp
unicode-data.h
unicode.cpp
unicode.h