Commit Graph

23 Commits

Author SHA1 Message Date
hiyouga
ae045c884f fix #5747 2024-10-29 10:47:04 +00:00
hiyouga
21db8ed2f4 use pre-commit 2024-10-29 09:07:46 +00:00
hiyouga
54c6905937 add docstrings, refactor logger 2024-09-08 00:56:56 +08:00
hiyouga
8e49940746 add rlhf-v dataset 2024-09-01 22:57:41 +08:00
moontidef
40908a36fa fix: rename optimzer to optimizer 2024-08-07 10:05:01 +08:00
hiyouga
2f09520c0d fix #4742 2024-07-09 23:24:24 +08:00
hiyouga
8baf3b22b0 refactor pissa, improve llamaboard 2024-06-28 01:04:24 +08:00
hiyouga
095fab58d3 tiny fix about badam 2024-06-25 01:54:53 +08:00
Jonery
5c2ff1b749 Cleaner integration. 2024-06-19 12:29:40 +08:00
Jonery
0f72aac8c9 Support distributed BAdam. 2024-06-18 12:27:47 +08:00
hiyouga
8c1046d78a support pissa 2024-06-16 01:08:12 +08:00
hiyouga
d87108daa6 add license 2024-06-15 17:54:33 +08:00
hiyouga
cf9f2d6c42 fix #4209
DeepSpeed ZeRO3 has inflight param error when calling model.eval()
2024-06-13 02:25:50 +08:00
hiyouga
f9e818d79c fix #4120 2024-06-07 04:18:05 +08:00
hiyouga
74f96efef9 rename files 2024-06-07 00:09:06 +08:00
hiyouga
fad2591e31 update trainers 2024-06-06 18:45:49 +08:00
hiyouga
67fe822324 fix #4090 2024-06-06 00:50:32 +08:00
hiyouga
f9a206509e remove gc warnings in DPO&KTO 2024-06-03 22:53:54 +08:00
hiyouga
7c8e01bb74 update dpo, kto trainer 2024-05-29 00:14:29 +08:00
hiyouga
cb63b32986 support SimPO #3900 2024-05-26 23:46:33 +08:00
hiyouga
3a023bca2a refactor data preprocessing, fix mllm rlhf 2024-05-24 04:08:25 +08:00
hiyouga
c450ee87a3 improve KTO impl., replace datasets 2024-05-18 03:44:56 +08:00
hiyouga
308edbc426 rename package 2024-05-16 18:39:08 +08:00