hiyouga
|
2b191ca776
|
support ppo score norm (trl 0.5.1.dev required)
Former-commit-id: 2b25db6d260ec1532281a592e873579346c7d21c
|
2023-08-18 12:02:42 +08:00 |
|
hiyouga
|
be4d2822ea
|
fix PPO trainer #551 , update readme
Former-commit-id: faead74849470cebae9e37cde5fab2a71b32aa43
|
2023-08-18 11:43:10 +08:00 |
|
hiyouga
|
c2644f939a
|
update training resuming
Former-commit-id: 2ec75c31f609e65116ac3b621eeb7d8ccbf69135
|
2023-08-18 01:41:17 +08:00 |
|
hoshi-hiyouga
|
3126164aa6
|
Merge branch 'main' into main
Former-commit-id: 870d2c7bf74d0da5a927bef4b8b01d15cc66a3e9
|
2023-08-18 01:37:23 +08:00 |
|
hiyouga
|
ed10486cad
|
support bf16 ppo #551
Former-commit-id: 092088967de7409a2d51847cfc7afc83a8887320
|
2023-08-18 00:40:32 +08:00 |
|
hiyouga
|
04fa430c6c
|
fix ChatGLM2 ppo #527 #528
Former-commit-id: 60d6ad64d7c9f6445b0df8de0153c3a311974198
|
2023-08-18 00:34:59 +08:00 |
|
hiyouga
|
fa1893b59c
|
fix generation bug #532
Former-commit-id: c071121e67374e5f09798db57cfc8668617a36ae
|
2023-08-17 22:21:34 +08:00 |
|
hiyouga
|
e993e717a5
|
fix streaming in pt stage #548 #549
Former-commit-id: 050e992bee2a9293cc7399b578de807b5bf9bddc
|
2023-08-17 17:59:26 +08:00 |
|
hiyouga
|
ffa09a01d6
|
fix baichuan and intern template
Former-commit-id: e1fd18fa6ef1009f978aca5210a259251a0b19a6
|
2023-08-17 01:27:20 +08:00 |
|
hiyouga
|
7d04f8567b
|
fix generation
Former-commit-id: 66a0300d312ef91c24fcf80667fa3b0bb8e1a342
|
2023-08-16 22:39:54 +08:00 |
|
hiyouga
|
baa709674f
|
fix system prompt
Former-commit-id: 411e775aa939bdd154a3f1e92921ede90d989f18
|
2023-08-16 01:35:52 +08:00 |
|
hiyouga
|
ca9a494d0c
|
fix baichuan template #481
Former-commit-id: 7608c6c25877d97ef26a1c209c4073c9c42f4535
|
2023-08-15 11:38:21 +08:00 |
|
hiyouga
|
7c046edb7b
|
fix ChatGLM RLHF
Former-commit-id: 4e43e887e432ceb7e9287b4e309b63af3c3ba1bf
|
2023-08-15 11:19:20 +08:00 |
|
hiyouga
|
ef2ca0a827
|
alert pad_token source
Former-commit-id: f26a84e0d927d2554890daf431a93652e18f4235
|
2023-08-15 00:07:56 +08:00 |
|
hiyouga
|
7f0b908de2
|
update webui
Former-commit-id: da30d0fb4abdb825f3383ddd106bb06a84695b7a
|
2023-08-14 22:45:26 +08:00 |
|
hoshi-hiyouga
|
5fc5e776ff
|
Merge pull request #511 from hiyouga/feature-autoTemplate
add template match and stage in webui
Former-commit-id: 413752ecba845cddaff5fb48db7d3d24b960eec1
|
2023-08-14 22:44:04 +08:00 |
|
codemayq
|
93b281c016
|
auto match template when change model_name
Former-commit-id: ab2d7ab0572765ce33a52ac71641062d5d904db4
|
2023-08-14 20:56:05 +08:00 |
|
codemayq
|
9585699918
|
add template match and stage in webui
Former-commit-id: d6283e7f041f08f76d18350cb5f6a6c58ca80e92
|
2023-08-14 20:42:59 +08:00 |
|
hiyouga
|
bceaba551d
|
fix ChatGLM lm_head #494
Former-commit-id: bf0048abdaeb2b9592d38ac991704ad014370b47
|
2023-08-14 14:14:48 +08:00 |
|
hiyouga
|
0bfeed3a7e
|
fix bug in webui
Former-commit-id: c95f0f687689934379b6c24abf872ffcde06073b
|
2023-08-14 11:38:42 +08:00 |
|
hiyouga
|
70a780c3c0
|
fix webui cache
Former-commit-id: 9aba5c197fbc8abaab77f454374f8b497f0310d0
|
2023-08-14 11:37:01 +08:00 |
|
hiyouga
|
688e8601ab
|
web UI integrating RLHF
Former-commit-id: 137fd146b90f89a1164b56e6d507b30b1f5c2437
|
2023-08-14 10:48:47 +08:00 |
|
hiyouga
|
4933ab5956
|
fix #480
Former-commit-id: ec15ca8fffacba2c34e1849c5ce90ca9989d66a2
|
2023-08-14 00:23:56 +08:00 |
|
hiyouga
|
6c7225a5d4
|
fix webui
Former-commit-id: 2c8b7414be9b43e20cc1d0575cc4dc1c7545fd86
|
2023-08-12 23:52:07 +08:00 |
|
hiyouga
|
a22982f2fa
|
tiny fix
Former-commit-id: 50a34c043de6d9e1410291e1d8c1ea9d53754e9e
|
2023-08-12 22:02:43 +08:00 |
|
hiyouga
|
c95479dddb
|
fix rope scaling
Former-commit-id: 2e0dd36700ec5e8294581c1db4b9431f755fc5f8
|
2023-08-12 22:00:01 +08:00 |
|
hiyouga
|
37bcbe8046
|
update readme
Former-commit-id: 6fa381400c21fa249cebcdff8c3afd72f8de20b3
|
2023-08-12 21:00:11 +08:00 |
|
hiyouga
|
fdfb644f0a
|
support rope scaling, fix #475 #476 #478
Former-commit-id: 337d5f68b72230e545e7a94ca789187c7a2b7187
|
2023-08-12 20:46:27 +08:00 |
|
codemayq
|
8bf5a98815
|
add sft script preview in webui
Former-commit-id: 2b72649b404750226aa418b61ef5a6c9ac03938f
|
2023-08-12 13:53:55 +08:00 |
|
hiyouga
|
be566a15a5
|
fix unusual output of 8bit models #278 #391
Former-commit-id: 337ce5272b81f5561162beb08814b0e5abf23703
|
2023-08-12 00:25:29 +08:00 |
|
hiyouga
|
d5f1b99ac4
|
Release v0.1.6
Former-commit-id: 43c8b3c3c8bfb2e32d17fb3e8b194938e37d54bd
|
2023-08-11 23:25:57 +08:00 |
|
hiyouga
|
bc665bacc7
|
add defaults
Former-commit-id: 4636d3bbe6b984ca93e3a80ae5239f3ddda461bd
|
2023-08-11 13:56:26 +08:00 |
|
hiyouga
|
52bfcf4883
|
fix stop word in baichuan template
Former-commit-id: cba5ac9cfc81f11b97831998ea15def5e0b487c2
|
2023-08-11 13:51:46 +08:00 |
|
hiyouga
|
06df3d6fb6
|
fix baichuan template
Former-commit-id: b1681fe35346381cda613297f1cbb710f0a6daa6
|
2023-08-11 13:45:47 +08:00 |
|
hiyouga
|
ca719a8697
|
support DPO training (2305.18290)
Former-commit-id: 6d98de148e4af63a7028dfaeb6cf86eb56a4488f
|
2023-08-11 03:02:53 +08:00 |
|
hoshi-hiyouga
|
72dfd74005
|
Merge pull request #451 from jovialchen/main
huggingface login for projects must login while running
Former-commit-id: 246ac241277908909b81cdf85fec1f24449dbae9
|
2023-08-10 17:25:38 +08:00 |
|
hiyouga
|
69302c4420
|
fix webui val size
Former-commit-id: 490c067d4e0828832e0ebdb704a9207dc974b15b
|
2023-08-10 15:20:44 +08:00 |
|
jiongxuc
|
42d7019b2e
|
huggingface login for projects must login while running
Former-commit-id: 0a4a2a1d3e0ff1f57215512d294d782080bd383c
|
2023-08-10 14:57:12 +08:00 |
|
hiyouga
|
5f0d0d6b9b
|
fix template
Former-commit-id: e3967eb1cdd8d19e8afee9ba52e7eb7d6cd86129
|
2023-08-09 23:14:27 +08:00 |
|
hiyouga
|
76cb63e4f6
|
fix template
Former-commit-id: 907e8cd86fbd4cdfa26dad21ceaf6e01d8fe37e4
|
2023-08-09 23:10:20 +08:00 |
|
hiyouga
|
467d571206
|
support val set in streaming mode
Former-commit-id: faed15b58ed00b1e09bb091e7eee48f5ef7c508b
|
2023-08-09 23:00:26 +08:00 |
|
hiyouga
|
972bfa700a
|
fix tokenizer
Former-commit-id: 7849587cd4e149291d08edef9a528a1bad796c7e
|
2023-08-09 17:52:15 +08:00 |
|
niuba
|
458955d0fb
|
add last_checkpoint support
Former-commit-id: 9f1977e4de00b14a9d1b555c25bcaf12998d5046
|
2023-08-09 16:39:27 +08:00 |
|
hiyouga
|
990eeccf45
|
fix sft trainer
Former-commit-id: 08cc888b1569572d0cd20bcf3f07e20072a0311a
|
2023-08-09 16:35:03 +08:00 |
|
hiyouga
|
a3a7465f00
|
fix rm #420, fix template #426, fix #423
Former-commit-id: 70ea3caaa7a7695c77179cd1bb18707a80a373d7
|
2023-08-09 16:23:31 +08:00 |
|
hoshi-hiyouga
|
031a819257
|
fix llama2 template
Former-commit-id: 6c74f726d4e672f5a1a57df201c27c1f697384f0
|
2023-08-09 00:58:27 +08:00 |
|
hoshi-hiyouga
|
eb4b4e3c8c
|
fix tokenizer
Former-commit-id: fa463ef279b596d5d53cc169831f51b42031fc05
|
2023-08-09 00:54:54 +08:00 |
|
hiyouga
|
d2e1fe9b1d
|
update webui
Former-commit-id: 343a4cd82b07a40f96ba413d1d991419ff07a24a
|
2023-08-09 00:26:11 +08:00 |
|
hiyouga
|
6e27a9e39a
|
fix tokenizer #417
Former-commit-id: 01aa678311bfd213a4b410a4e0ff09f48a0d40a1
|
2023-08-08 23:59:41 +08:00 |
|
hiyouga
|
805478c911
|
fix bug
Former-commit-id: 0dff1d951f1a9fe05a74d334bf477b55c7c64199
|
2023-08-08 21:28:28 +08:00 |
|