dataset
|
add prompt template (#6273)
|
2025-04-22 10:39:47 +08:00 |
distributed
|
add uuid to rollout log
|
2025-05-20 09:45:56 +08:00 |
models
|
Add GRPO and Support RLVR for PPO (#6186)
|
2025-02-18 09:43:36 +08:00 |
quant
|
[ColossalChat] Update RLHF V2 (#5286)
|
2024-03-29 14:12:29 +08:00 |
ray
|
[ColossalChat] Update RLHF V2 (#5286)
|
2024-03-29 14:12:29 +08:00 |
trainer
|
[feat] Support prompt level dynamic (#6300)
|
2025-05-14 16:40:35 +08:00 |
utils
|
Add GRPO and Support RLVR for PPO (#6186)
|
2025-02-18 09:43:36 +08:00 |
__init__.py
|
[ColossalChat] Update RLHF V2 (#5286)
|
2024-03-29 14:12:29 +08:00 |