This website requires JavaScript.
Explore
Help
Sign In
nikola
/
llama_cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
22
Packages
Projects
Releases
Wiki
Activity
8,522
Commits
1
Branch
0
Tags
9c600bcd4b3b21f70c9d95cf8a938e43192eb492
Commit Graph
2 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
Ivan Chikish
and
GitHub
cceb1b4e33
common : inline functions (
#18639
)
2026-02-16 17:52:24 +02:00
Ivan
and
GitHub
116efee0ee
cuda: add q8_0->f32 cpy operation (
#9571
)
...
llama: enable K-shift for quantized KV cache It will fail on unsupported backends or quant types.
2024-09-24 02:14:24 +02:00