This website requires JavaScript.
Explore
Help
Sign In
LLM-Inference
/
ik_llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ikawrakow/ik_llama.cpp.git
synced
2026-08-12 22:29:39 +04:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
8d91d3c3d9495c20d481a39bdbb4d28151be86e5
ik_llama.cpp
/
ggml
T
History
Kawrakow
and
GitHub
e6f8112f3b
Adjust CUDA FA kernel parameters for head size 512 on Turing (
#1942
)
2026-06-10 07:49:21 +02:00
..
cmake
Merge mainline llama.cpp (
#3
)
2024-07-27 07:55:01 +02:00
include
MLA TP -khad: ggml_dequant_hadamard fused op + wv_b/wk_b_pp Hadamard fold (
#1852
)
2026-05-21 07:29:15 +03:00
src
Adjust CUDA FA kernel parameters for head size 512 on Turing (
#1942
)
2026-06-10 07:49:21 +02:00
.gitignore
Merge mainline llama.cpp (
#3
)
2024-07-27 07:55:01 +02:00
CMakeLists.txt
ggml : default GGML_WIN_VER to 0x0A00 (Windows 10) (
#1755
)
2026-05-08 13:23:04 +03:00