This website requires JavaScript.
Explore
Help
Sign In
mozempk
/
ggml-org-llama-cpp-mirror
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-12 09:23:00 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
2,283
Commits
784
Branches
7,403
Tags
7c4263d4261d6ee6f0539d53eb9e1b4d120ba8af
Commit Graph
1 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
Shouzheng Liu
and
GitHub
1cbf561466
metal : new q4_0 matrix-vector kernel (
#2188
)
...
Prefetch data to improve GPU utilization. ~48% faster for 33B model.
2023-07-12 23:10:55 +03:00