hiyouga
|
c555a83ec9
|
fix #6448
Former-commit-id: 04f78e85af5af14b4c195936623e426a6a128af2
|
2024-12-27 16:54:39 +00:00 |
|
hiyouga
|
c57fbebd55
|
support report custom args
Former-commit-id: d41254c40a1c5cacf9377096adb27efa9bdb79ea
|
2024-12-21 21:42:45 +00:00 |
|
hoshi-hiyouga
|
da8a72d611
|
Merge pull request #6401 from Zeyi-Lin/hiyouga/swanlab
feat: add swanlab for experiment tracking and visualization.
Former-commit-id: e65fe507f7643bf40b0fc462805c7b7f8ef6b738
|
2024-12-21 14:09:33 +08:00 |
|
ZeYi Lin
|
6d13503867
|
fix: by hiyouga suggestion
Former-commit-id: 41195f1bc69e4b5da7a265369d368b06754362cf
|
2024-12-20 16:43:03 +08:00 |
|
hiyouga
|
67479ce5d9
|
support disable shuffling
Former-commit-id: 9d8c35fd6b838ede0bd6827c6c6121f2cba2b11b
|
2024-12-19 08:53:21 +00:00 |
|
hiyouga
|
0ce5b84ccd
|
fix dpo metrics
Former-commit-id: 57029280da825a39fbf5a05097921b861f126669
|
2024-11-02 20:59:01 +08:00 |
|
hiyouga
|
25f00034d5
|
fix incorrect loss value for vlms
Former-commit-id: 0aa29a71ce958343a2086090d647eb63b8f5f5be
|
2024-10-30 08:56:46 +00:00 |
|
hiyouga
|
94f6eef40f
|
fix #5747
Former-commit-id: 26d07de349c98b547cd6a6166ea20616d08ba343
|
2024-10-29 10:47:04 +00:00 |
|
hiyouga
|
dbbfb5f5dc
|
use pre-commit
Former-commit-id: 7cfede95df22a9ff236788f04159b6b16b8d04bb
|
2024-10-29 09:07:46 +00:00 |
|
hiyouga
|
fb90faf19a
|
add docstrings, refactor logger
Former-commit-id: c34e489d71f8f539028543ccf8ee92cecedd6276
|
2024-09-08 00:56:56 +08:00 |
|
hiyouga
|
04db03bdfd
|
add rlhf-v dataset
Former-commit-id: 3fd18fc34a0c994a738504746abfd5548e002437
|
2024-09-01 22:57:41 +08:00 |
|
hiyouga
|
228f745235
|
refactor mm training
Former-commit-id: 179c0558699e287cbf38a2d73bff47e86d589c5a
|
2024-08-30 02:14:31 +08:00 |
|
moontidef
|
5243075bb7
|
fix: rename optimzer to optimizer
Former-commit-id: 186dc1fde822e6a603ac273538741ea3853f243e
|
2024-08-07 10:05:01 +08:00 |
|
hiyouga
|
884a4a33ee
|
refactor pissa, improve llamaboard
Former-commit-id: 619556e46c19718f702c97df5d570a2a4c5fb13a
|
2024-06-28 01:04:24 +08:00 |
|
hiyouga
|
4d2c279083
|
tiny fix about badam
Former-commit-id: 03f49267c7406e36aee35639f86e6e0383897090
|
2024-06-25 01:54:53 +08:00 |
|
Jonery
|
a22e932b4f
|
Cleaner integration.
Former-commit-id: 26d4b05d424bd71f570195dd433258caf6465d92
|
2024-06-19 12:29:40 +08:00 |
|
Jonery
|
b567216702
|
Support distributed BAdam.
Former-commit-id: bdcb986e37975911c190a74d3e60bb77aa2033bd
|
2024-06-18 12:27:47 +08:00 |
|
hiyouga
|
640372cb66
|
tiny fix
Former-commit-id: f7f440986b0ae3b38ea9f2da80789629d4f79ea1
|
2024-06-16 01:06:41 +08:00 |
|
hiyouga
|
acfae2e677
|
add license
Former-commit-id: 69cfc98d7c81756a5ab6bf962240e393e449fef0
|
2024-06-15 17:54:33 +08:00 |
|
hiyouga
|
045cef901e
|
fix #4209
DeepSpeed ZeRO3 has inflight param error when calling model.eval()
Former-commit-id: 4be013f18ea6a35b5a11db98db5f0670ffb41619
|
2024-06-13 02:25:50 +08:00 |
|
hiyouga
|
8cc3bbdc62
|
fix #4120
Former-commit-id: 2a44da678a5e360a9c0f9056397ac9e801329321
|
2024-06-07 04:18:05 +08:00 |
|
hiyouga
|
0b1f4a34f8
|
rename files
Former-commit-id: e1a8431770fc36c0c9ee7fed4abbc3d7fdcc5efd
|
2024-06-07 00:09:06 +08:00 |
|
hiyouga
|
67246f52f2
|
update trainers
Former-commit-id: b7f6c4a171293cf4f3e88f15a811f847342f84ee
|
2024-06-06 18:45:49 +08:00 |
|
hiyouga
|
2dc5743fba
|
remove gc warnings in DPO&KTO
Former-commit-id: b649bdcbafb464a638387429b770fe258b41f8af
|
2024-06-03 22:53:54 +08:00 |
|
hoshi-hiyouga
|
ca60eca259
|
Update trainer.py
Former-commit-id: 8565d4b43db905374c328ae57c71fc226980d14f
|
2024-06-03 22:08:38 +08:00 |
|
enji.zhou
|
59aca304c0
|
fix KTO Trainer Sampler
Former-commit-id: 39eb1bfa272011554322e9bb2534f83b68282a70
|
2024-06-03 21:32:38 +08:00 |
|
hiyouga
|
0de2ab5d16
|
update dpo, kto trainer
Former-commit-id: 4a6cc3c7046f8b27d05ea53ef216bab6fa7ebfaf
|
2024-05-29 00:14:29 +08:00 |
|
hiyouga
|
e15389be7d
|
clean kto trainer
Former-commit-id: 76402bd78cbd3a99a544f0ac019468b569b0e1d1
|
2024-05-28 21:43:26 +08:00 |
|
hiyouga
|
ed2601a909
|
support SimPO #3900
Former-commit-id: 6b954ce60155cf8334150b795cfc4bb63ca74c8b
|
2024-05-26 23:46:33 +08:00 |
|
hiyouga
|
664cba05e3
|
refactor data preprocessing, fix mllm rlhf
Former-commit-id: 53ff2dd24f9121ea30c95063bb72e49a9b31e980
|
2024-05-24 04:08:25 +08:00 |
|
hiyouga
|
d24969bb7e
|
improve KTO impl., replace datasets
Former-commit-id: e56a57ddcf061de6e4acc8679f7dbf0b68364986
|
2024-05-18 03:44:56 +08:00 |
|
enji.zhou
|
d16a1d9ed0
|
add kto
Former-commit-id: ec51986cf70b0bdd79b8141e45916670fb97a08e
|
2024-05-17 13:09:17 +08:00 |
|