This website requires JavaScript.
Explore
Help
Register
Sign In
Lumpiasty
/
llama.cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Packages
Projects
Releases
Wiki
Activity
Files
683f0c72e5b3c07fab90bfd9ec2ce8661d624228
llama.cpp
/
ggml
/
src
/
ggml-webgpu
T
History
Masashi Yoshimura
32e41fa5b4
ggml-webgpu: tune subgroup split (d_split) in flash_attn_vec (
#25418
)
2026-07-09 08:34:19 +09:00
..
wgsl-shaders
ggml-webgpu: tune subgroup split (d_split) in flash_attn_vec (
#25418
)
2026-07-09 08:34:19 +09:00
CMakeLists.txt
ggml-webgpu: FlashAttention refactor + standardize quantization support (
#23834
)
2026-06-04 08:05:04 +03:00
ggml-webgpu-shader-lib.hpp
ggml-webgpu: tune subgroup split (d_split) in flash_attn_vec (
#25418
)
2026-07-09 08:34:19 +09:00
ggml-webgpu.cpp
ggml-webgpu: add support for NVFP4 (
#25143
)
2026-06-30 17:20:04 +09:00
pre_wgsl.hpp
ggml-webgpu: FlashAttention refactor + standardize quantization support (
#23834
)
2026-06-04 08:05:04 +03:00