This website requires JavaScript.
Explore
Help
Sign In
nikola
/
llama_cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
33
Packages
Projects
Releases
Wiki
Activity
Files
e84b71c2c6da6e69c8f815168ea836f9716a325e
llama_cpp
/
gguf-py
/
gguf
T
History
Georgi Gerganov
and
GitHub
e84b71c2c6
ggml : drop support for QK_K=64 (
#7473
)
...
* ggml : drop support for QK_K=64 ggml-ci * opencl : restore QK_K=256 define
2024-05-23 10:00:21 +03:00
..
__init__.py
convert-hf : support direct Q8_0 conversion (
#7234
)
2024-05-13 14:10:51 -04:00
constants.py
ggml : drop support for QK_K=64 (
#7473
)
2024-05-23 10:00:21 +03:00
gguf_reader.py
convert-hf : save memory with lazy evaluation (
#7075
)
2024-05-08 18:16:38 -04:00
gguf_writer.py
llama : add phi3 128K model support (
#7225
)
2024-05-21 23:28:32 +03:00
gguf.py
gguf-py: Refactor and allow reading/modifying existing GGUF files (
#3981
)
2023-11-11 08:04:50 +03:00
lazy.py
convert-hf : support direct Q8_0 conversion (
#7234
)
2024-05-13 14:10:51 -04:00
py.typed
convert : various script cleanups/fixes + merges and special token handling (
#2842
)
2023-08-30 11:25:50 +03:00
quants.py
convert-hf : support direct Q8_0 conversion (
#7234
)
2024-05-13 14:10:51 -04:00
tensor_mapping.py
llama : add Jina Embeddings architecture (
#6826
)
2024-05-11 10:46:09 +03:00
vocab.py
convert-hf : save memory with lazy evaluation (
#7075
)
2024-05-08 18:16:38 -04:00