This website requires JavaScript.
Explore
Help
Sign In
LLM-Inference
/
ik_llama.cpp
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ikawrakow/ik_llama.cpp.git
synced
2026-08-12 22:29:39 +04:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
afd9fd274e679c34d9d3417f3017b17c844cc8a0
ik_llama.cpp
/
examples
/
quantize-stats
T
History
Kawrakow
318899c8b7
bitnet: add 2 bpw quantization
...
The scalar dot product already chieves 37 t/s for TG!
2024-06-22 12:02:51 +03:00
..
CMakeLists.txt
build
: rename main → llama-cli, server → llama-server, llava-cli → llama-llava-cli, etc... (
#7809
)
2024-06-13 00:41:52 +01:00
quantize-stats.cpp
bitnet: add 2 bpw quantization
2024-06-22 12:02:51 +03:00