This website requires JavaScript.
Explore
Help
Sign In
mozempk
/
ggml-org-llama-cpp-mirror
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-12 09:23:00 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
2,392
Commits
784
Branches
7,403
Tags
bb6d00bbf98476215596b9df3870b5504ef5a29a
Commit Graph
1 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
Shouzheng Liu
and
GitHub
1cbf561466
metal : new q4_0 matrix-vector kernel (
#2188
)
...
Prefetch data to improve GPU utilization. ~48% faster for 33B model.
2023-07-12 23:10:55 +03:00