hiyouga
|
b71da932eb
|
fix bug in PPO training
Former-commit-id: 856522a3df4bb9ddfaaa137119eceb9574873950
|
2023-11-16 02:32:54 +08:00 |
|
hiyouga
|
f441932bd1
|
support full-parameter PPO
Former-commit-id: ce783036001397a20b0b4c5da2fea6d0c03389d2
|
2023-11-16 02:08:04 +08:00 |
|
hiyouga
|
e30290444a
|
support multiple modules in freeze training #1514
Former-commit-id: 4907452d955367ebe987e6deae4fd4213628f2b2
|
2023-11-15 17:08:18 +08:00 |
|
hiyouga
|
8387f3011c
|
fix #1494
Former-commit-id: d125ef55358837d4d76943739afeb6c70a901cd7
|
2023-11-14 18:07:20 +08:00 |
|
hiyouga
|
125587b187
|
refactor evaluation, upgrade trl to 074
Former-commit-id: 442aefb925c4ff02b98aa30c49c2e01d04f6496a
|
2023-11-13 22:20:35 +08:00 |
|
hiyouga
|
55e097aaac
|
add todo
Former-commit-id: a0c31c68c4909637b86c90c319c321fd887c4910
|
2023-11-10 14:38:18 +08:00 |
|
hiyouga
|
6ee32cf71c
|
tiny fix
Former-commit-id: 415bca900e5cc3afaddd5b06d35f472d9ead3263
|
2023-11-09 17:20:49 +08:00 |
|
Yanqing
|
fc05fd52cf
|
Update finetuning_args.py
更新 chatglm/falcon/bloom 的 lora_target 的名称
Former-commit-id: 3684dffa14ca0551d51027467c0134b884ed1c59
|
2023-11-09 17:04:40 +08:00 |
|
hiyouga
|
91f406cc99
|
fix ppo train and dpo eval
Former-commit-id: 01260d975477ebb8570933a1bd7f547b4dba607f
|
2023-11-07 22:48:51 +08:00 |
|
hiyouga
|
1f2c56bff9
|
delete file
Former-commit-id: 479d0af2dc4ab8282b9d55aba1b03ab3a54f400b
|
2023-11-07 16:20:12 +08:00 |
|
hiyouga
|
3d40bdb600
|
upgrade peft, fix #1088 #1411
Former-commit-id: b2a60905f384ada92618bf21301fe96dac1c10bf
|
2023-11-07 16:13:36 +08:00 |
|
hiyouga
|
a9db89a025
|
update data readme (zh)
Former-commit-id: cc8ffa10d877f5893f3940204e5bec6f3266559f
|
2023-11-02 23:42:49 +08:00 |
|
hiyouga
|
a1b0655457
|
support sharegpt format, add datasets
Former-commit-id: a8371724130db2fbd7273a480e2acb251e382aec
|
2023-11-02 23:10:04 +08:00 |
|
hiyouga
|
15cef791ba
|
fix #1356
Former-commit-id: dff128c7e38dd079a5840ea4e73ee3e9bbd1c3c9
|
2023-11-02 16:51:52 +08:00 |
|
hiyouga
|
22b3c913e9
|
fix #1325
Former-commit-id: 083787dbfe41f58ff59cb16ddde02df98593aef5
|
2023-11-01 23:38:49 +08:00 |
|
hiyouga
|
fcfcac4858
|
support dataset cache
Former-commit-id: 3fe7df628db4093d7b3c121ececff60be0aa3a8a
|
2023-10-26 21:48:45 +08:00 |
|
hiyouga
|
d6c77d9196
|
reimplement neftune
Former-commit-id: 7b4acf7265b04cc4a674b3dcafdb90e76f149e39
|
2023-10-22 16:15:08 +08:00 |
|
anvie
|
3635823fbe
|
add NEFTune optimization
Former-commit-id: 57fb40aa04fec11ca165a97ea463579faeaeebe7
|
2023-10-21 13:24:10 +07:00 |
|
hiyouga
|
95697652f1
|
fix #1232
Former-commit-id: b665e9e133bf2f6f10346c374eb0de8a96dd5c7e
|
2023-10-20 23:28:52 +08:00 |
|
hiyouga
|
4930118761
|
fix #1218
Former-commit-id: 7a11a42dfd414d140cd83b7a74760715d2ae2078
|
2023-10-19 16:17:41 +08:00 |
|
hiyouga
|
f3fa47fa7d
|
refactor export, fix #1190
Former-commit-id: ea82f8a82a7356bbdf204190d596d0b1c8ef1a84
|
2023-10-15 16:01:48 +08:00 |
|
hiyouga
|
e585c789ce
|
fix #1184
Former-commit-id: af18b0dce7a4ef10b30da069d454010eddd269af
|
2023-10-14 19:20:11 +08:00 |
|
hiyouga
|
2562376f84
|
fix ppo args
Former-commit-id: 11bd271364488d523d5117ec2ea26f39853175b7
|
2023-10-11 23:40:50 +08:00 |
|
hiyouga
|
c9d1cd108d
|
refactor model_dtype, fix PPO trainer
Former-commit-id: 2818af0b0967d7695f27658acac0b7e2c2728e5d
|
2023-10-11 23:16:01 +08:00 |
|
hiyouga
|
deb17942ab
|
fix layer norm dtype
Former-commit-id: 84b7486885c600e5e65c5ba9095d56ecc2502977
|
2023-09-28 00:25:55 +08:00 |
|
hiyouga
|
927ff702ff
|
refactor finetuning Args
Former-commit-id: 620efe1d8d2429f4bc3fa8009900ec43e1b5ef4b
|
2023-09-27 22:28:06 +08:00 |
|
hiyouga
|
108c31e1fc
|
support LongLoRA
Former-commit-id: 90375f600d5601866836123597fa3ef52008eeef
|
2023-09-27 21:55:50 +08:00 |
|
hiyouga
|
4581d09fa6
|
fix #944
Former-commit-id: 338b8664edea5ae65192ac657bb013581245ae15
|
2023-09-21 19:51:02 +08:00 |
|
hiyouga
|
8ab5566dc0
|
support FlashAttention2
Former-commit-id: d8aa1404bee9842f3e4cd037ad8d66c85470ac37
|
2023-09-10 20:43:56 +08:00 |
|
hiyouga
|
9ed4bb63d4
|
change to right-padding, update reward score #803
Former-commit-id: 8ea32e4046d75ddfa9517669e9de9f48fea720c6
|
2023-09-08 20:04:31 +08:00 |
|
hiyouga
|
a4fd976048
|
refactor dataset_attr, add eos in pt, fix #757
Former-commit-id: a9d1fb72f791ae57a4d12f4e3a7e2abccf6a7077
|
2023-09-01 19:00:45 +08:00 |
|
codemayq
|
2b979d39f2
|
add stage in DatasetAttr
Former-commit-id: ba94c8729dd8c90bedf7079a6978e150fc92b737
|
2023-08-23 20:54:53 +08:00 |
|
hiyouga
|
802494e20a
|
update template
Former-commit-id: 4318347d3f1982c773dad1074636ec7b550770fd
|
2023-08-22 19:46:09 +08:00 |
|
hiyouga
|
b88f0b396c
|
support ppo score norm (trl 0.5.1.dev required)
Former-commit-id: 53e33418d02ee0f34c783e30ae510b811308c598
|
2023-08-18 12:02:42 +08:00 |
|
hiyouga
|
03edfd07e7
|
fix PPO trainer #551 , update readme
Former-commit-id: 90205244186df558cd6b0000728d638348db3a10
|
2023-08-18 11:43:10 +08:00 |
|
hiyouga
|
edc15c62fa
|
fix system prompt
Former-commit-id: 7407d9daa16bf6b3cd5002e16b2c53e402d2bc39
|
2023-08-16 01:35:52 +08:00 |
|
hiyouga
|
3f0a2d6adc
|
support rope scaling, fix #475 #476 #478
Former-commit-id: fa940c17b8d3e379af08804003f1a522c1cd6ac4
|
2023-08-12 20:46:27 +08:00 |
|
hiyouga
|
79f4ba0d26
|
Release v0.1.6
Former-commit-id: a48cb0d474ef0648a97387daf5f623498b5e3ee6
|
2023-08-11 23:25:57 +08:00 |
|
hiyouga
|
abdfa26d06
|
support DPO training (2305.18290)
Former-commit-id: 3ec4351cfdaf2aefcc7d13345e19d79874ed61d3
|
2023-08-11 03:02:53 +08:00 |
|
jiongxuc
|
7ffd961b8b
|
huggingface login for projects must login while running
Former-commit-id: 3e000c2b60c2e29bcafcf8d39c1a5d567ae2491c
|
2023-08-10 14:57:12 +08:00 |
|
hiyouga
|
6404167ab7
|
support val set in streaming mode
Former-commit-id: d86ea314a197fd821770d895e988c48d46679047
|
2023-08-09 23:00:26 +08:00 |
|
hiyouga
|
921778a7cf
|
update args spec
Former-commit-id: 5453b93db04ee26223b7616dfdf26e749bea7473
|
2023-08-07 15:23:35 +08:00 |
|
hiyouga
|
b32ed1d7be
|
support interleave probs
Former-commit-id: 69744c17e8180e0ad549b57d575454724b820d01
|
2023-08-04 21:27:35 +08:00 |
|
hiyouga
|
9c84c4ed5d
|
support Qwen-7B, fix InternLM-7B inference
Former-commit-id: 87f8f830e20aa839e089559c1d038954742000ef
|
2023-08-03 15:53:32 +08:00 |
|
hiyouga
|
e80b75b560
|
support streaming data, fix #284 #274 #268
Former-commit-id: 0411a4b3e122e7907441bc7a64b004948741a620
|
2023-07-31 23:33:00 +08:00 |
|
hiyouga
|
2a664783c3
|
Update data_args.py
Former-commit-id: 513e1f1ec952dc5901d90dd34e32440230bca3f0
|
2023-07-28 17:42:41 +08:00 |
|
hiyouga
|
2e0342dc54
|
fix #194
Former-commit-id: 8f7819fcaa00f6ddcf59552c2866d8fc5659d0e9
|
2023-07-19 17:07:33 +08:00 |
|
hiyouga
|
b8b38a9ade
|
create chat model
Former-commit-id: 657cf0f55a7f0886bc837bdd44528971dc5e5caa
|
2023-07-15 19:26:20 +08:00 |
|
hiyouga
|
a696148d6b
|
modity code structure
Former-commit-id: f75137661358f9070bc70c341dfa2cc5fd69cf94
|
2023-07-15 16:54:28 +08:00 |
|