This website requires JavaScript.
Explore
Help
Register
Sign In
Lumpiasty
/
llama.cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Packages
Projects
Releases
Wiki
Activity
5,397
Commits
4
Branches
0
Tags
02cdd2d8b092b5a4bb18e013c6887ce49ba20ac5
Commit Graph
2 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
Gian-Carlo Pascutto
and
Georgi Gerganov
58d07a8043
metal : copy kernels for quant to F32/F16 conversions (
#12017
)
...
metal: use dequantize_q templates --------- Co-authored-by: Georgi Gerganov <
ggerganov@gmail.com
>
2025-02-25 11:27:58 +02:00
Gian-Carlo Pascutto
d70908421f
cuda: Add Q5_1, Q5_0, Q4_1 and Q4_0 to F32 conversion support. (
#12000
)
2025-02-22 09:43:24 +01:00