This website requires JavaScript.
Explore
Help
Sign In
nikola
/
llama_cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
34
Packages
Projects
Releases
Wiki
Activity
6,117
Commits
1
Branch
0
Tags
1425f587a82bc303469b5c32759a2746ba4e1e20
Commit Graph
2 Commits
This Branch
This Branch
All Branches
Author
SHA1
Message
Date
3 people
Gian-Carlo Pascutto
GitHub
Georgi Gerganov
58d07a8043
metal : copy kernels for quant to F32/F16 conversions (
#12017
)
...
metal: use dequantize_q templates --------- Co-authored-by: Georgi Gerganov <
ggerganov@gmail.com
>
2025-02-25 11:27:58 +02:00
Gian-Carlo Pascutto
and
GitHub
d70908421f
cuda: Add Q5_1, Q5_0, Q4_1 and Q4_0 to F32 conversion support. (
#12000
)
2025-02-22 09:43:24 +01:00