This website requires JavaScript.
Explore
Help
Register
Sign In
Lumpiasty
/
llama.cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Packages
Projects
Releases
Wiki
Activity
Files
516e8d7a8ab3d231b2c3bf238b3d895970bc769f
llama.cpp
/
ggml
T
History
Rithik Sharma
and
GitHub
434b2a1ff6
ggml-webgpu: add Q1_0 support (
#22374
)
...
* add fast matmul matvec q1_0 kernel * ggml-webgpu: drop redundant zero-fills in Q1_0 shmem init
2026-04-27 15:50:59 -07:00
..
cmake
ggml: backend-agnostic tensor parallelism (experimental) (
#19378
)
2026-04-09 16:42:19 +02:00
include
CUDA: manage NCCL communicators in context (
#21891
)
2026-04-15 15:58:40 +02:00
src
ggml-webgpu: add Q1_0 support (
#22374
)
2026-04-27 15:50:59 -07:00
.gitignore
vulkan : cmake integration (
#8119
)
2024-07-13 18:12:39 +02:00
CMakeLists.txt
HIP: flip GGML_HIP_GRAPHS to default on (
#22254
)
2026-04-23 02:34:31 +02:00