[Refactor] Integrated some lightllm kernels into token-attention (#4946)

mirror of https://github.com/hpcaitech/ColossalAI.git synced 2025-09-03 18:19:58 +00:00

* add some req for inference

* clean codes

* add codes

* add some lightllm deps

* clean codes

* hello

* delete rms files

* add some comments

* add comments

* add doc

* add lightllm deps

* add lightllm cahtglm2 kernels

* add lightllm cahtglm2 kernels

* replace rotary embedding with lightllm kernel

* add some commnets

* add some comments

* add some comments

* add

* replace fwd kernel att1

* fix a arg

* add

* add

* fix token attention

* add some comments

* clean codes

* modify comments

* fix readme

* fix bug

* fix bug

---------

Co-authored-by: cuiqing.li <lixx336@gmail.com>
Co-authored-by: CjhHa1 <cjh18671720497@outlook.com>

This commit is contained in:

Cuiqing Li

2023-10-19 22:22:47 +08:00

committed by

GitHub

parent 11009103be

commit 3a41e8304e

20 changed files with 160 additions and 1555 deletions

3

requirements/requirements.txt

View File

@@ -11,3 +11,6 @@ ninja
 torch>=1.12
 safetensors
 einops
 sentencepiece
 google
 protobuf

[Refactor] Integrated some lightllm kernels into token-attention (#4946)

3 requirements/requirements.txt Unescape Escape View File

3

requirements/requirements.txt

View File