hiyouga
|
095fab58d3
|
tiny fix about badam
|
2024-06-25 01:54:53 +08:00 |
|
Jonery
|
5c2ff1b749
|
Cleaner integration.
|
2024-06-19 12:29:40 +08:00 |
|
Jonery
|
0f72aac8c9
|
Support distributed BAdam.
|
2024-06-18 12:27:47 +08:00 |
|
hiyouga
|
38b6b0f52e
|
tiny fix
|
2024-06-16 01:06:41 +08:00 |
|
hiyouga
|
d87108daa6
|
add license
|
2024-06-15 17:54:33 +08:00 |
|
hiyouga
|
cf9f2d6c42
|
fix #4209
DeepSpeed ZeRO3 has inflight param error when calling model.eval()
|
2024-06-13 02:25:50 +08:00 |
|
hiyouga
|
f9e818d79c
|
fix #4120
|
2024-06-07 04:18:05 +08:00 |
|
hiyouga
|
74f96efef9
|
rename files
|
2024-06-07 00:09:06 +08:00 |
|
hiyouga
|
fad2591e31
|
update trainers
|
2024-06-06 18:45:49 +08:00 |
|
hiyouga
|
f9a206509e
|
remove gc warnings in DPO&KTO
|
2024-06-03 22:53:54 +08:00 |
|
hoshi-hiyouga
|
24499f40dc
|
Update trainer.py
|
2024-06-03 22:08:38 +08:00 |
|
enji.zhou
|
34a2c5087a
|
fix KTO Trainer Sampler
|
2024-06-03 21:32:38 +08:00 |
|
hiyouga
|
7c8e01bb74
|
update dpo, kto trainer
|
2024-05-29 00:14:29 +08:00 |
|
hiyouga
|
900e1ea622
|
clean kto trainer
|
2024-05-28 21:43:26 +08:00 |
|
hiyouga
|
cb63b32986
|
support SimPO #3900
|
2024-05-26 23:46:33 +08:00 |
|
hiyouga
|
3a023bca2a
|
refactor data preprocessing, fix mllm rlhf
|
2024-05-24 04:08:25 +08:00 |
|
hiyouga
|
c450ee87a3
|
improve KTO impl., replace datasets
|
2024-05-18 03:44:56 +08:00 |
|
enji.zhou
|
db1d5a4f51
|
add kto
|
2024-05-17 13:09:17 +08:00 |
|