This website requires JavaScript.
Explore
Help
Sign In
mozempk
/
ggml-org-llama-cpp-mirror
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-11 00:52:23 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
2,848
Commits
780
Branches
7,393
Tags
988631335a20d06497f58be0b8ba13adb4323a22
Commit Graph
1 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
Shouzheng Liu
and
GitHub
1cbf561466
metal : new q4_0 matrix-vector kernel (
#2188
)
...
Prefetch data to improve GPU utilization. ~48% faster for 33B model.
2023-07-12 23:10:55 +03:00