1
0
mirror of https://github.com/hpcaitech/ColossalAI.git synced 2025-04-30 20:55:17 +00:00
ColossalAI/colossalai/inference/modeling
Hongxin Liu 646b3c5a90
[shardformer] fix linear 1d row and support uneven splits for fused qkv linear ()
* [tp] hotfix linear row

* [tp] support uneven split for fused linear

* [tp] support sp for fused linear

* [tp] fix gpt2 mlp policy

* [tp] fix gather fused and add fused linear row
2024-10-10 14:34:45 +08:00
..
backends [Inference] Fix flash-attn import and add model test () 2024-06-12 14:13:50 +08:00
layers [Feat] Distrifusion Acceleration Support for Diffusion Inference () 2024-07-30 10:43:26 +08:00
models [Feat] Distrifusion Acceleration Support for Diffusion Inference () 2024-07-30 10:43:26 +08:00
policy [shardformer] fix linear 1d row and support uneven splits for fused qkv linear () 2024-10-10 14:34:45 +08:00
__init__.py [doc] updated inference readme () 2024-02-02 14:31:10 +08:00