ColossalAI/requirements.txt at 7b9b86441fbffdd07021f234ec88d0dbc470fa5c - ColossalAI - Gitea: Git with a cup of tea at home

github/ColossalAI

mirror of https://github.com/hpcaitech/ColossalAI.git synced 2025-05-12 10:27:44 +00:00

Wenhao Chen 7b9b86441f

[chat]: update rm, add wandb and fix bugs (#4471 )

* feat: modify forward fn of critic and reward model

* feat: modify calc_action_log_probs

* to: add wandb in sft and rm trainer

* feat: update train_sft

* feat: update train_rm

* style: modify type annotation and add warning

* feat: pass tokenizer to ppo trainer

* to: modify trainer base and maker base

* feat: add wandb in ppo trainer

* feat: pass tokenizer to generate

* test: update generate fn tests

* test: update train tests

* fix: remove action_mask

* feat: remove unused code

* fix: fix wrong ignore_index

* fix: fix mock tokenizer

* chore: update requirements

* revert: modify make_experience

* fix: fix inference

* fix: add padding side

* style: modify _on_learn_batch_end

* test: use mock tokenizer

* fix: use bf16 to avoid overflow

* fix: fix workflow

* [chat] fix gemini strategy

* [chat] fix

* sync: update colossalai strategy

* fix: fix args and model dtype

* fix: fix checkpoint test

* fix: fix requirements

* fix: fix missing import and wrong arg

* fix: temporarily skip gemini test in stage 3

* style: apply pre-commit

* fix: temporarily skip gemini test in stage 1&2

---------

Co-authored-by: Mingyan Jiang <1829166702@qq.com>

2023-09-20 15:53:58 +08:00

15 lines

166 B

Plaintext

Raw Blame History

 transformers>=4.20.1
 tqdm
 datasets
 loralib
 colossalai>=0.3.1
 torch<2.0.0, >=1.12.1
 langchain
 tokenizers
 fastapi
 sse_starlette
 wandb
 sentencepiece
 gpustat
 tensorboard