This website requires JavaScript.
Explore
Help
Sign In
mozempk
/
ggml-org-llama-cpp-mirror
Watch
1
Star
0
Fork
0
mirror of
https://github.com/ggml-org/llama.cpp.git
synced
2026-09-20 21:29:46 +02:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
3,051
Commits
794
Branches
7,510
Tags
5921b8f089d3b7bda86aac5a66825df6a6c10603
Commit Graph
1 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
Shouzheng Liu
and
GitHub
1cbf561466
metal : new q4_0 matrix-vector kernel (
#2188
)
...
Prefetch data to improve GPU utilization. ~48% faster for 33B model.
2023-07-12 23:10:55 +03:00