This website requires JavaScript.
Explore
Help
Sign In
LLM-Inference
/
llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-08-12 22:31:11 +04:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
10,397
Commits
684
Branches
7,066
Tags
master
Commit Graph
1 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
0
and
GitHub
5988633170
cuda : add warp-per-row wkv7 kernel for single-token decode (
#26111
)
2026-08-11 20:46:23 +03:00