llama : add Qwen support (#4281)
* enable qwen to llama.cpp * llama : do not GPU split bias tensors --------- Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
This commit is contained in:
co-authored by
Georgi Gerganov
parent
880f57973b
commit
37c746d687
@@ -0,0 +1 @@
|
||||
You are a helpful assistant.
|
||||
Reference in New Issue
Block a user