This website requires JavaScript.
Explore
Help
Sign In
mozempk
/
ggml-org-llama-cpp-mirror
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-17 03:39:45 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
2,853
Commits
790
Branches
7,465
Tags
5a419926b0c4efab0531401aea91522aaea9fd07
Commit Graph
1 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
Shouzheng Liu
and
GitHub
1cbf561466
metal : new q4_0 matrix-vector kernel (
#2188
)
...
Prefetch data to improve GPU utilization. ~48% faster for 33B model.
2023-07-12 23:10:55 +03:00