hiyouga
|
f23e5b602a
|
fix reward model loading
Former-commit-id: 9709ca501180a1afce32e9043aedb359762b437d
|
2023-11-07 17:20:51 +08:00 |
|
hiyouga
|
857696ed9c
|
fix args
Former-commit-id: 44d0fa2ac6a6423c7ddaf91eb8998c1b9248c04e
|
2023-11-07 16:36:06 +08:00 |
|
hiyouga
|
2eb65d21ac
|
upgrade peft, fix #1088 #1411
Former-commit-id: aa7d104f8e050d12cb8f585bc8a52c850995500f
|
2023-11-07 16:13:36 +08:00 |
|
hiyouga
|
217fde0918
|
fix bug in data loader, support dpo eval
Former-commit-id: f4f3dcff990468a2fa864b7176adcebbcf16dac9
|
2023-11-03 00:34:26 +08:00 |
|
hiyouga
|
e387a50475
|
fix shift short attention
Former-commit-id: 9a49cce8e6f6b222f74a07bdab40efee6a77b0f1
|
2023-10-09 17:07:46 +08:00 |
|
hiyouga
|
a09a7b650d
|
remove PeftTrainer
Former-commit-id: cc0cff3e991f194732d278e627648e528118a719
|
2023-09-10 22:23:23 +08:00 |
|
hiyouga
|
180a05a446
|
fix import error
Former-commit-id: b3207a974a45038591b8cbbcf20d1ca1142d6679
|
2023-08-23 20:45:03 +08:00 |
|
hiyouga
|
eb9ac9ee1f
|
fix #649
Former-commit-id: e6120a937ddb4f3c0b9bcb2466742f5cf4f77f8c
|
2023-08-23 20:21:15 +08:00 |
|
hiyouga
|
d6be98cda6
|
fix #617
Former-commit-id: a7bdaf1c92c7d798caf8438dc42a8972632ec584
|
2023-08-21 18:16:11 +08:00 |
|
hiyouga
|
c2644f939a
|
update training resuming
Former-commit-id: 2ec75c31f609e65116ac3b621eeb7d8ccbf69135
|
2023-08-18 01:41:17 +08:00 |
|
hiyouga
|
ca719a8697
|
support DPO training (2305.18290)
Former-commit-id: 6d98de148e4af63a7028dfaeb6cf86eb56a4488f
|
2023-08-11 03:02:53 +08:00 |
|