This website requires JavaScript.
Explore
Help
Register
Sign In
Lumpiasty
/
llama.cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Packages
Projects
Releases
Wiki
Activity
Files
32e41fa5b48e15b93c7a40ce677226b2e773c351
llama.cpp
/
ggml
T
History
Masashi Yoshimura
and
GitHub
32e41fa5b4
ggml-webgpu: tune subgroup split (d_split) in flash_attn_vec (
#25418
)
2026-07-09 08:34:19 +09:00
..
cmake
ggml : Parallelize quant LUT init (
#23595
)
2026-05-25 10:15:46 +03:00
include
Add Q2_0 quantization: type definition and CPU backend (
#24448
)
2026-07-07 12:05:47 -07:00
src
ggml-webgpu: tune subgroup split (d_split) in flash_attn_vec (
#25418
)
2026-07-09 08:34:19 +09:00
.gitignore
vulkan : cmake integration (
#8119
)
2024-07-13 18:12:39 +02:00
CMakeLists.txt
ggml : bump version to 0.15.3 (ggml/1550)
2026-06-26 15:04:42 +03:00