llama : add Qwen support (#4281)

* enable qwen to llama.cpp

* llama : do not GPU split bias tensors

---------

Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>
This commit is contained in:
Shijie
2023-12-01 20:16:31 +02:00
committed by GitHub
co-authored by Georgi Gerganov
parent 880f57973b
commit 37c746d687
5 changed files with 372 additions and 9 deletions
+1
View File
@@ -0,0 +1 @@
You are a helpful assistant.