This website requires JavaScript.
Explore
Help
Sign In
nikola
/
llama_cpp
Watch
1
Star
0
Fork
0
Code
Issues
Pull Requests
Actions
33
Packages
Projects
Releases
Wiki
Activity
Files
b83bab15a5d2a1e7807d09613a9b34309d86cfaa
llama_cpp
/
gguf-py
/
gguf
T
History
compilade
and
GitHub
b83bab15a5
gguf-py : fix and simplify quantized shape round-trip (
#7483
)
...
* gguf-py : fix and simplify quantized shape round-trip * gguf-py : remove unused import
2024-05-25 11:11:48 +10:00
..
__init__.py
convert-hf : support direct Q8_0 conversion (
#7234
)
2024-05-13 14:10:51 -04:00
constants.py
Add support for ArcticForCausalLM (
#7020
)
2024-05-24 14:31:13 +02:00
gguf_reader.py
gguf-py : fix and simplify quantized shape round-trip (
#7483
)
2024-05-25 11:11:48 +10:00
gguf_writer.py
gguf-py : fix and simplify quantized shape round-trip (
#7483
)
2024-05-25 11:11:48 +10:00
gguf.py
gguf-py: Refactor and allow reading/modifying existing GGUF files (
#3981
)
2023-11-11 08:04:50 +03:00
lazy.py
convert-hf : support direct Q8_0 conversion (
#7234
)
2024-05-13 14:10:51 -04:00
py.typed
convert : various script cleanups/fixes + merges and special token handling (
#2842
)
2023-08-30 11:25:50 +03:00
quants.py
gguf-py : fix and simplify quantized shape round-trip (
#7483
)
2024-05-25 11:11:48 +10:00
tensor_mapping.py
Add support for ArcticForCausalLM (
#7020
)
2024-05-24 14:31:13 +02:00
vocab.py
convert-hf : save memory with lazy evaluation (
#7075
)
2024-05-08 18:16:38 -04:00