This website requires JavaScript.
Explore
Help
Sign In
nikola
/
llama_cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
33
Packages
Projects
Releases
Wiki
Activity
5,876
Commits
1
Branch
0
Tags
0c1df14b5f8d992805cb22d0b77b44092a18aeab
Commit Graph
2 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
3 people
Gian-Carlo Pascutto
GitHub
Georgi Gerganov
58d07a8043
metal : copy kernels for quant to F32/F16 conversions (
#12017
)
...
metal: use dequantize_q templates --------- Co-authored-by: Georgi Gerganov <
ggerganov@gmail.com
>
2025-02-25 11:27:58 +02:00
Gian-Carlo Pascutto
and
GitHub
d70908421f
cuda: Add Q5_1, Q5_0, Q4_1 and Q4_0 to F32 conversion support. (
#12000
)
2025-02-22 09:43:24 +01:00