DAN™ and GitHub
46be942214
llama : add support for the cohere2 model architecture ( #10900 )
2025-01-04 16:33:31 +02:00
DAN™ and GitHub
d62b532c52
Use model->gguf_kv for loading the template instead of using the C API. ( #10868 )
...
* Bump model_template to 16384 bytes to support larger chat templates.
* Use `model->gguf_kv` for efficiency.
2024-12-17 23:24:22 +01:00
DAN™ and GitHub
4cd621c26d
convert : add BPE pre-tokenization for DBRX ( #7132 )
...
* Add BPE pre-tokenization for DBRX.
* Add vocab GGUFs.
* Remove test.
* Remove GGUFs.
2024-05-08 13:43:23 +03:00
889bdd7686
command-r : add BPE pre-tokenization ( #7063 )
...
* Add BPE pre-tokenization for Command-R/R+.
* Bump transformers convert requirement.
* command-r : add individual digits regex
---------
Co-authored-by: Georgi Gerganov <ggerganov@gmail.com >
2024-05-05 08:19:30 +03:00
DAN™ and GitHub
e00b4a8f81
Fix more int overflow during quant (PPL/CUDA). ( #6563 )
...
* Fix more int overflow during quant.
* Fix some more int overflow in softmax.
* Revert back to int64_t.
2024-04-29 00:38:44 +02:00
DAN™ and GitHub
e0717e751e
Add GritLM as supported models. ( #6513 )
2024-04-07 19:33:59 +02:00
fa046eafbc
Fix params underscore convert to dash. ( #6203 )
...
* Fix params underscore convert to dash.
* Update common/common.cpp
---------
Co-authored-by: slaren <slarengh@gmail.com >
2024-03-22 02:32:42 +01:00
DAN™ and GitHub
d8b009a945
Remove undeed header file. ( #6158 )
2024-03-19 17:16:09 +01:00
DAN™ and GitHub
4c28b82529
common : print usage on '-h' and '--help' ( #6145 )
2024-03-19 07:59:36 +02:00
496bc79bc2
common : tidy-up argument parsing ( #6105 )
...
* Tidy-up argument parsing.
* Missing ref.
* common : minor
* common : add static classifier
---------
Co-authored-by: Georgi Gerganov <ggerganov@gmail.com >
2024-03-18 10:27:44 +02:00
DAN™ and GitHub
15961ec04d
common : refactor nested if causing error C1061 on MSVC ( #6101 )
...
* Refactor nested if causing error C1061 on MSVC.
* Revert back and remove else's.
* Add flag to track found arguments.
2024-03-16 17:39:15 +02:00
bcebd7dbf6
llama : add support for GritLM ( #5959 )
...
* add gritlm example
* gritlm results match
* tabs to spaces
* comment out debug printing
* rebase to new embed
* gritlm embeddings are back babeee
* add to gitignore
* allow to toggle embedding mode
* Clean-up GritLM sample code.
* Fix types.
* Flush stdout and output ending newline if streaming.
* mostly style fixes; correct KQ_mask comment
* add causal_attn flag to llama_cparams
* gritml : minor
* llama : minor
---------
Co-authored-by: Douglas Hanley <thesecretaryofwar@gmail.com >
Co-authored-by: Georgi Gerganov <ggerganov@gmail.com >
2024-03-10 17:56:30 +02:00
DAN™ and GitHub
82f3e668ad
common : use LLAMA_DEFAULT_SEED ( #5855 )
2024-03-04 10:08:19 +02:00
5a51cc1bb4
main : support special tokens as reverse/anti prompt ( #5847 )
...
* Support special tokens as reverse/anti prompt.
* Tokenize antiprompts only once.
* main : minor
---------
Co-authored-by: Georgi Gerganov <ggerganov@gmail.com >
2024-03-04 09:57:20 +02:00
DAN™ and GitHub
99115f3fa6
cmake : fix build-info.h on MSVC ( #3309 )
2023-09-25 18:45:33 -04:00