This website requires JavaScript.
Explore
Help
Register
Sign In
Lumpiasty
/
llama.cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Packages
Projects
Releases
Wiki
Activity
Files
c9ca51c1f6b18427cde490c7c7eba11d87a96b2d
llama.cpp
/
ggml
/
src
/
ggml-opencl
T
History
Shawn Gu
58546250cf
opencl: add bin kernels
kernel_gemm_moe_q4_0_q8_1_dp4a_bin
,
kernel_gemm_moe_mxfp4_q8_1_dp4a_bin
(
#27768
)
2026-08-27 09:44:05 -07:00
..
kernels
opencl: fold the gpt-oss MoE per-expert bias adds into the epilogue (op/kernel fusion) (
#26431
)
2026-08-21 14:24:33 -07:00
cl-program-cache.cpp
opencl: cache compiled cl_program binaries on disk (
#26050
)
2026-07-24 08:14:33 -07:00
cl-program-cache.h
opencl: cache compiled cl_program binaries on disk (
#26050
)
2026-07-24 08:14:33 -07:00
CMakeLists.txt
opencl: fold the gpt-oss MoE per-expert bias adds into the epilogue (op/kernel fusion) (
#26431
)
2026-08-21 14:24:33 -07:00
fa_tune.h
opencl: general flash attention decode performance optimizations (
#25366
)
2026-07-06 19:57:52 -07:00
ggml-opencl.cpp
opencl: add bin kernels
kernel_gemm_moe_q4_0_q8_1_dp4a_bin
,
kernel_gemm_moe_mxfp4_q8_1_dp4a_bin
(
#27768
)
2026-08-27 09:44:05 -07:00
libdl.h
opencl: allow loading precompiled binary kernels from library (
#23042
)
2026-07-01 10:29:22 -07:00