Commit Graph

31 Commits

Author SHA1 Message Date
hiyouga
5111cac6f8 support report custom args 2024-12-21 21:42:45 +00:00
hoshi-hiyouga
947e22a4a3 Merge pull request #6401 from Zeyi-Lin/hiyouga/swanlab
feat: add swanlab for experiment tracking and visualization.
2024-12-21 14:09:33 +08:00
ZeYi Lin
3a7ea2048a fix: by hiyouga suggestion 2024-12-20 16:43:03 +08:00
hiyouga
c7cedc7569 support disable shuffling 2024-12-19 08:53:21 +00:00
hiyouga
4270f7dfb9 fix dpo metrics 2024-11-02 20:59:01 +08:00
hiyouga
30567a1487 fix incorrect loss value for vlms 2024-10-30 08:56:46 +00:00
hiyouga
ae045c884f fix #5747 2024-10-29 10:47:04 +00:00
hiyouga
21db8ed2f4 use pre-commit 2024-10-29 09:07:46 +00:00
hiyouga
54c6905937 add docstrings, refactor logger 2024-09-08 00:56:56 +08:00
hiyouga
8e49940746 add rlhf-v dataset 2024-09-01 22:57:41 +08:00
hiyouga
3382317e32 refactor mm training 2024-08-30 02:14:31 +08:00
moontidef
40908a36fa fix: rename optimzer to optimizer 2024-08-07 10:05:01 +08:00
hiyouga
8baf3b22b0 refactor pissa, improve llamaboard 2024-06-28 01:04:24 +08:00
hiyouga
095fab58d3 tiny fix about badam 2024-06-25 01:54:53 +08:00
Jonery
5c2ff1b749 Cleaner integration. 2024-06-19 12:29:40 +08:00
Jonery
0f72aac8c9 Support distributed BAdam. 2024-06-18 12:27:47 +08:00
hiyouga
38b6b0f52e tiny fix 2024-06-16 01:06:41 +08:00
hiyouga
d87108daa6 add license 2024-06-15 17:54:33 +08:00
hiyouga
cf9f2d6c42 fix #4209
DeepSpeed ZeRO3 has inflight param error when calling model.eval()
2024-06-13 02:25:50 +08:00
hiyouga
f9e818d79c fix #4120 2024-06-07 04:18:05 +08:00
hiyouga
74f96efef9 rename files 2024-06-07 00:09:06 +08:00
hiyouga
fad2591e31 update trainers 2024-06-06 18:45:49 +08:00
hiyouga
f9a206509e remove gc warnings in DPO&KTO 2024-06-03 22:53:54 +08:00
hoshi-hiyouga
24499f40dc Update trainer.py 2024-06-03 22:08:38 +08:00
enji.zhou
34a2c5087a fix KTO Trainer Sampler 2024-06-03 21:32:38 +08:00
hiyouga
7c8e01bb74 update dpo, kto trainer 2024-05-29 00:14:29 +08:00
hiyouga
900e1ea622 clean kto trainer 2024-05-28 21:43:26 +08:00
hiyouga
cb63b32986 support SimPO #3900 2024-05-26 23:46:33 +08:00
hiyouga
3a023bca2a refactor data preprocessing, fix mllm rlhf 2024-05-24 04:08:25 +08:00
hiyouga
c450ee87a3 improve KTO impl., replace datasets 2024-05-18 03:44:56 +08:00
enji.zhou
db1d5a4f51 add kto 2024-05-17 13:09:17 +08:00