This website requires JavaScript.
Explore
Help
Sign In
nikola
/
llama_cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
30
Packages
Projects
Releases
Wiki
Activity
Files
1e1aca09dab40318c9e0c2c0c47ce239876adc53
llama_cpp
/
ggml
T
History
Masashi Yoshimura
and
GitHub
1e1aca09da
ggml-webgpu: Improve prefill speeds for k-quants + refactor matmul for Q4/Q5/Q8 and k-quants (
#24225
)
...
* ggml-webgpu: Improve prefill speeds + refactor matmul for quants * Fixes for editroconfig checker
2026-06-08 15:19:56 -07:00
..
cmake
ggml : Parallelize quant LUT init (
#23595
)
2026-05-25 10:15:46 +03:00
include
TP: quantized KV cache support (
#23792
)
2026-06-01 12:30:10 +02:00
src
ggml-webgpu: Improve prefill speeds for k-quants + refactor matmul for Q4/Q5/Q8 and k-quants (
#24225
)
2026-06-08 15:19:56 -07:00
.gitignore
vulkan : cmake integration (
#8119
)
2024-07-13 18:12:39 +02:00
CMakeLists.txt
ggml : bump version to 0.14.0 (ggml/1533)
2026-06-08 14:31:33 +03:00