hiyouga
|
2f878bde11
|
support ORPO
Former-commit-id: 17bf8a2c3a7bb5b83071c8659cfd8751e894e692
|
2024-03-31 18:29:50 +08:00 |
|
hiyouga
|
e4f3d583df
|
fix #2982
Former-commit-id: 8d603f8820efd1617557f2bc5d9674143abe7c57
|
2024-03-28 20:22:31 +08:00 |
|
hiyouga
|
c311375b50
|
fix bug
Former-commit-id: 3164b4f11b72684c8aa2105037cb36c47b6acfd4
|
2024-03-26 17:30:12 +08:00 |
|
hiyouga
|
ec94e5e876
|
fix #2961
Former-commit-id: 511f6754026fbbf48bd481018015338a6a3ad92f
|
2024-03-26 17:26:14 +08:00 |
|
hiyouga
|
8b8671817f
|
improve lora+ impl.
Former-commit-id: 72367307dfadf936fb989ebe8bc9f0ff229fb933
|
2024-03-13 23:32:51 +08:00 |
|
齐保元
|
24c9277488
|
[FEATURE]: ADD LORA+ ALGORITHM
Former-commit-id: a0965cd62c85545aa2364e244295df2963308354
|
2024-03-13 19:43:27 +08:00 |
|
hiyouga
|
4a4e4b4354
|
support layerwise galore
Former-commit-id: 8664262cde3919e10eaecbd66e8c5d356856362e
|
2024-03-10 00:24:11 +08:00 |
|
hiyouga
|
5c00783697
|
update hardware requirements
Former-commit-id: 393c2de27ce0a2dee793092843ec0afa54f49a6d
|
2024-03-09 03:58:18 +08:00 |
|
hiyouga
|
7443ac3116
|
fix chat engine, update webui
Former-commit-id: 5d956e2a5167201aecdfce2794c25d8a2d84e234
|
2024-03-08 03:01:53 +08:00 |
|
hiyouga
|
2235020cc9
|
update galore args
Former-commit-id: 0ac6b40a4772b61a3476bb74b976d24c408a2c35
|
2024-03-08 01:17:32 +08:00 |
|
hiyouga
|
5b50458acf
|
fix galore
Former-commit-id: 33a4c24a8a3c153bc62edf74b9246699a0ae3233
|
2024-03-08 00:44:51 +08:00 |
|
hiyouga
|
2c010c72b8
|
support galore
Former-commit-id: 28f78621883917425fabe49f5473778111012127
|
2024-03-07 22:41:36 +08:00 |
|
hiyouga
|
34533b2f35
|
support vllm
Former-commit-id: d07ad5cc1cdbc13879afd84f653afdfee03a6933
|
2024-03-07 20:26:31 +08:00 |
|
hiyouga
|
e887aface7
|
fix version checking
Former-commit-id: 3016e6565708637c1d760f2cd5a67cbd8a5a6c26
|
2024-03-06 14:51:51 +08:00 |
|
hiyouga
|
5abbca70d3
|
support DoRA, AWQ, AQLM #2512
Former-commit-id: cfefacaa37453a15c55866d019887f24e886a577
|
2024-02-28 19:53:28 +08:00 |
|
hiyouga
|
2f738a1db6
|
fix #2532
Former-commit-id: 3cc10a01a792a92b99b952a45bb21c25097fccf6
|
2024-02-21 21:55:14 +08:00 |
|
hiyouga
|
0fcb931f18
|
support lora for llama pro
Former-commit-id: 9aeb404a946795d6c4fa3cb45e3e96ffeec13646
|
2024-02-21 02:17:22 +08:00 |
|
hiyouga
|
48ae2110aa
|
update webui
Former-commit-id: ba998c67abb56b2679068572f4eb8398635bbda9
|
2024-02-19 16:49:58 +08:00 |
|
hiyouga
|
96265ec154
|
support llama pro #2338 , add rslora
Former-commit-id: 7924ffc55d98e33bfbfbca303e46c8f476435673
|
2024-02-15 02:27:36 +08:00 |
|
hiyouga
|
75adbfec79
|
add option to disable version check
Former-commit-id: 91d09a01ac3b5da29d284b8d51cdfe4252b391e0
|
2024-02-10 22:31:23 +08:00 |
|
hiyouga
|
23dd337ac2
|
lint
Former-commit-id: 88a1bc97736bf06f292cd768fc8b61503aca1988
|
2024-02-07 01:10:04 +08:00 |
|
hiyouga
|
0fc8612b97
|
add hint for freeze #2412
Former-commit-id: 6545c02790e39395a87d664682ab73e0e3191099
|
2024-02-03 23:38:56 +08:00 |
|
hiyouga
|
b27e91222c
|
format style
Former-commit-id: 638234ceee1b19716e45b6e5f4ea54d9122da4df
|
2024-01-20 20:15:56 +08:00 |
|
hiyouga
|
e199967391
|
add bf16 lora option
Former-commit-id: b6ec112bebcb379caa32617a135df4d5d3cf865b
|
2024-01-19 16:29:03 +08:00 |
|
hiyouga
|
4be704823c
|
fix #2195
Former-commit-id: a83fb6d3ff1e2a9657e926083a926d48b0f3e1a6
|
2024-01-16 23:53:50 +08:00 |
|
hiyouga
|
2a3980d6ba
|
update loader
Former-commit-id: 6629087e12f64f2635f24311234202077814083c
|
2023-12-24 19:10:23 +08:00 |
|
hiyouga
|
f0d405f392
|
support unsloth
Former-commit-id: 7aad0b889d9a316fffd65f32a419078418fc0986
|
2023-12-23 00:14:33 +08:00 |
|
hiyouga
|
ce79528bb1
|
fix param type
Former-commit-id: ba69378841778410f8004385df3fd4c41e5fa573
|
2023-12-21 17:33:01 +08:00 |
|
hiyouga
|
4e75ca1222
|
support dpo-ftx
Former-commit-id: b87c74289d523ef88611b376074199ffd03cf103
|
2023-12-16 19:21:41 +08:00 |
|
hiyouga
|
7dbc670902
|
support quantization in export model
Former-commit-id: 3524aa1e58da94ab00e9a2024952ea1b4119b2af
|
2023-12-15 23:44:50 +08:00 |
|
hiyouga
|
bd03307bbd
|
refactor adapter hparam
Former-commit-id: 0716f5e470afffd2df5a815712b552a4b4797153
|
2023-12-15 20:53:11 +08:00 |
|
hiyouga
|
15b321da8e
|
remove loftq
Former-commit-id: 3a8a50d4d42082b3bdce549653b398e49f2eb554
|
2023-12-13 01:53:46 +08:00 |
|
hiyouga
|
4c69025a83
|
support loftq
Former-commit-id: 6219dfbd9377528bce286c724ae2dd0090881095
|
2023-12-12 22:47:06 +08:00 |
|
hiyouga
|
bd28dd0fe6
|
update readme
Former-commit-id: 8cace7780867dd78760f40c46fd5b6ddd47dea0a
|
2023-12-12 11:44:30 +08:00 |
|
hiyouga
|
1cb390b9b2
|
implement rm server #1543
Former-commit-id: 7df4f3ab206fddb462f6ed865eaf04234fd72ed6
|
2023-12-03 20:52:54 +08:00 |
|
hiyouga
|
ae1048db6d
|
fix #1659
Former-commit-id: 475a3fa0f4c09d4cfd55ec66271a6d3c9eb5f4d2
|
2023-11-28 20:52:28 +08:00 |
|
hiyouga
|
b015ac35d8
|
support export size setting
Former-commit-id: 859a6ea9425a09d7263f6436d05102df8129c248
|
2023-11-26 18:34:09 +08:00 |
|
hiyouga
|
f06c4c8f7a
|
update ppo trainer
Former-commit-id: 5021062493ed63ad1f6133cfb543e4e7f528d2cc
|
2023-11-20 21:39:15 +08:00 |
|
Yuchen Han
|
ec910a87c0
|
Update finetuning_args.py
Former-commit-id: b24635d22b3084ad29217ef55c1dd1fa4f85a1fb
|
2023-11-17 00:15:51 -08:00 |
|
hiyouga
|
678052a7ef
|
fix rlhf callback
Former-commit-id: 1817ffc86fe3463ea91e9359c0e3611979a9d53e
|
2023-11-16 03:26:19 +08:00 |
|
hiyouga
|
b71da932eb
|
fix bug in PPO training
Former-commit-id: 856522a3df4bb9ddfaaa137119eceb9574873950
|
2023-11-16 02:32:54 +08:00 |
|
hiyouga
|
f441932bd1
|
support full-parameter PPO
Former-commit-id: ce783036001397a20b0b4c5da2fea6d0c03389d2
|
2023-11-16 02:08:04 +08:00 |
|
hiyouga
|
e30290444a
|
support multiple modules in freeze training #1514
Former-commit-id: 4907452d955367ebe987e6deae4fd4213628f2b2
|
2023-11-15 17:08:18 +08:00 |
|
hiyouga
|
125587b187
|
refactor evaluation, upgrade trl to 074
Former-commit-id: 442aefb925c4ff02b98aa30c49c2e01d04f6496a
|
2023-11-13 22:20:35 +08:00 |
|
hiyouga
|
6ee32cf71c
|
tiny fix
Former-commit-id: 415bca900e5cc3afaddd5b06d35f472d9ead3263
|
2023-11-09 17:20:49 +08:00 |
|
Yanqing
|
fc05fd52cf
|
Update finetuning_args.py
更新 chatglm/falcon/bloom 的 lora_target 的名称
Former-commit-id: 3684dffa14ca0551d51027467c0134b884ed1c59
|
2023-11-09 17:04:40 +08:00 |
|
hiyouga
|
91f406cc99
|
fix ppo train and dpo eval
Former-commit-id: 01260d975477ebb8570933a1bd7f547b4dba607f
|
2023-11-07 22:48:51 +08:00 |
|
hiyouga
|
3d40bdb600
|
upgrade peft, fix #1088 #1411
Former-commit-id: b2a60905f384ada92618bf21301fe96dac1c10bf
|
2023-11-07 16:13:36 +08:00 |
|
hiyouga
|
d6c77d9196
|
reimplement neftune
Former-commit-id: 7b4acf7265b04cc4a674b3dcafdb90e76f149e39
|
2023-10-22 16:15:08 +08:00 |
|
anvie
|
3635823fbe
|
add NEFTune optimization
Former-commit-id: 57fb40aa04fec11ca165a97ea463579faeaeebe7
|
2023-10-21 13:24:10 +07:00 |
|