hiyouga
|
c0e887c85a
|
update training resuming
Former-commit-id: 2ec75c31f609e65116ac3b621eeb7d8ccbf69135
|
2023-08-18 01:41:17 +08:00 |
|
hiyouga
|
ff5ebb6a86
|
support bf16 ppo #551
Former-commit-id: 092088967de7409a2d51847cfc7afc83a8887320
|
2023-08-18 00:40:32 +08:00 |
|
hiyouga
|
a15a602809
|
support rope scaling, fix #475 #476 #478
Former-commit-id: 337d5f68b72230e545e7a94ca789187c7a2b7187
|
2023-08-12 20:46:27 +08:00 |
|
hiyouga
|
7ada4f5f6f
|
support DPO training (2305.18290)
Former-commit-id: 6d98de148e4af63a7028dfaeb6cf86eb56a4488f
|
2023-08-11 03:02:53 +08:00 |
|
hiyouga
|
6eb4120464
|
support val set in streaming mode
Former-commit-id: faed15b58ed00b1e09bb091e7eee48f5ef7c508b
|
2023-08-09 23:00:26 +08:00 |
|
hiyouga
|
ab3f685330
|
fix RM save model
Former-commit-id: 8104cc2425431eb1cddccf3909855296116f922b
|
2023-08-01 11:56:17 +08:00 |
|
hiyouga
|
5faad2a64c
|
release v0.1.4
Former-commit-id: 81f84aaf2e120e39edb28ef42893939fc9a184e2
|
2023-08-01 10:08:47 +08:00 |
|
hiyouga
|
1c2422df31
|
fix arg check
Former-commit-id: 2c5c73de9ebc88e2d04e80754781c94a571133a0
|
2023-07-31 23:48:57 +08:00 |
|
hiyouga
|
5cab5c1b36
|
update readme
Former-commit-id: d99cda254e5025ff3f968d256197ab031bfabef1
|
2023-07-31 23:42:32 +08:00 |
|
hiyouga
|
63123a9098
|
support streaming data, fix #284 #274 #268
Former-commit-id: 819cc1353599e5fa45658bc56dd0dbe4b258b197
|
2023-07-31 23:33:00 +08:00 |
|
hiyouga
|
1f4a65afea
|
update web UI, support rm predict #210
Former-commit-id: 92cc6b655dc91b94d5bf9d8618c3b57d5cf94333
|
2023-07-21 13:27:27 +08:00 |
|
hiyouga
|
a69b1b1c3a
|
modity code structure
Former-commit-id: 0682ed357210897e0b67c4a6eb31a94b3eb929f1
|
2023-07-15 16:54:28 +08:00 |
|