This website requires JavaScript.
Explore
Help
Register
Sign In
Lumpiasty
/
llama.cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Packages
Projects
Releases
Wiki
Activity
5,608
Commits
4
Branches
0
Tags
91a8ee6a6f1f4c8547ff7b745ef95c6edc1d2af6
Commit Graph
1 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
Ivan
116efee0ee
cuda: add q8_0->f32 cpy operation (
#9571
)
...
llama: enable K-shift for quantized KV cache It will fail on unsupported backends or quant types.
2024-09-24 02:14:24 +02:00