Commit Graph

71 Commits

Author SHA1 Message Date
hiyouga
9df7a26e6b video datasets
Former-commit-id: 8cafc7b055
2024-09-05 02:04:17 +08:00
hiyouga
6e98872622 fix #5324
Former-commit-id: a61c8c4890
2024-09-02 23:56:21 +08:00
hiyouga
bfdcc6bacf add rlhf-v dataset
Former-commit-id: 8e49940746
2024-09-01 22:57:41 +08:00
hiyouga
cb776752f6 fix mixed mm inputs and rlhf-v
Former-commit-id: 9967ccb3ae
2024-09-01 20:52:47 +08:00
hiyouga
09a2ecebc4 add test mm plugin
Former-commit-id: a2a8c0b92c
2024-08-31 01:53:38 +08:00
hiyouga
f31e7e0dfc remove visual_inputs, fix qlora
Former-commit-id: a025c3df61
2024-08-31 00:24:51 +08:00
hiyouga
a83756b5e9 refactor mm training
Former-commit-id: 3382317e32
2024-08-30 02:14:31 +08:00
hoshi-hiyouga
98b0c7530c Merge pull request #5290 from simonJJJ/qwen2_vl
support qwen2-vl

Former-commit-id: 727e184840
2024-08-30 02:10:36 +08:00
hiyouga
0e4ee9d9a3 update liger kernel
Former-commit-id: a7dd7d325e
2024-08-29 20:46:08 +08:00
simonJJJ
8a09b1e732 initial-commit
Former-commit-id: aeb85f200b
2024-08-28 16:51:35 +08:00
hiyouga
c765292093 support liger kernel
Former-commit-id: 72bc8f0111
2024-08-27 11:20:14 +08:00
hiyouga
20013e130b fix #5048
Former-commit-id: b7ca6c8dc1
2024-08-05 23:48:19 +08:00
hiyouga
1cddf80a97 tiny fix
Former-commit-id: 5665062ca0
2024-07-22 21:10:15 +08:00
hoshi-hiyouga
37c6a0c6dc fix #4917
Former-commit-id: 26082fc6c9
2024-07-22 11:28:31 +08:00
hiyouga
14bc7b0551 fix up
Former-commit-id: 29ebcd75d5
2024-07-15 01:04:56 +08:00
hiyouga
12e0e5d0d7 tiny fix
Former-commit-id: d3c01552e0
2024-07-14 10:56:45 +08:00
hiyouga
0b26011181 fix gemma2 attention
Former-commit-id: 2f6af73da2
2024-07-13 23:33:45 +08:00
hiyouga
d7657d772d tiny fix
Former-commit-id: 0c699de39d
2024-07-04 03:47:05 +08:00
hiyouga
7b3c1f29ff fix packing for eager/sdpa attn
Former-commit-id: 6fd6aa4530
2024-07-04 01:52:43 +08:00
hoshi-hiyouga
a38ff842d0 Merge pull request #4224 from chuan298/main
Implement efficient packing without cross-contamination attention

Former-commit-id: 87d9b2d005
2024-07-04 01:18:54 +08:00
hiyouga
bfdaadcc40 update packing
Former-commit-id: cce7083024
2024-07-04 01:10:55 +08:00
hoshi-hiyouga
51c75985b8 Update packing.py
Former-commit-id: a36e8f2dd5
2024-07-03 23:36:01 +08:00
hiyouga
13cec0cc2f update func name
Former-commit-id: c346f79f99
2024-07-03 23:29:33 +08:00
hiyouga
e671ed520b update arg name
Former-commit-id: 8a6a7b9c8a
2024-07-03 23:23:24 +08:00
hiyouga
104151d558 tiny fix
Former-commit-id: 8b1172b910
2024-07-03 02:31:50 +08:00
ancv
7f42932957 move efficient_packing from data_args to model_args
Former-commit-id: e8e13b0942
2024-07-02 18:37:55 +07:00
hoshi-hiyouga
2452f57cd7 Merge branch 'main' into main
Former-commit-id: e8e6af2651
2024-07-01 21:01:09 +08:00
hiyouga
de4de5b5ab tiny fix
Former-commit-id: 8c41a0aa6d
2024-07-01 03:55:20 +08:00
hiyouga
bbc37b2880 fix #4398 #4592
Former-commit-id: d74244d568
2024-06-30 21:28:51 +08:00
hiyouga
2b006beab1 loose gemma2 attention
Former-commit-id: 2f4b89ace1
2024-06-29 01:42:14 +08:00
hiyouga
87e60f8bac bf16 by default, gemma2 attns
Gemma2 finetuning cannot work until merging https://github.com/huggingface/transformers/pull/31674


Former-commit-id: 4d35e218b1
2024-06-28 06:00:26 +08:00
hiyouga
58607ec1b0 add quant checks
Former-commit-id: 96a5044394
2024-06-27 01:12:25 +08:00
hiyouga
d2d9fa4abb support HQQ/EETQ #4113
Former-commit-id: ad144c2265
2024-06-27 00:29:42 +08:00
hiyouga
6b2733ce12 improve autogptq integration
Former-commit-id: addca926de
2024-06-26 22:11:44 +08:00
hiyouga
0ae1302e41 fix #4432
Former-commit-id: 1e9d0aa1e4
2024-06-25 02:34:04 +08:00
hiyouga
47651a94a3 fix #4410
Former-commit-id: fca893d73c
2024-06-24 22:34:31 +08:00
stceum
9aa640f27b Bug Fix: off is parsed as False in yaml file, changed to disabled to avoid this.
Former-commit-id: 3ed063f281
2024-06-24 20:39:31 +08:00
ancv
5319447aa5 move configure_packing to llamafactory.model.patcher and fix constants
Former-commit-id: 770f75dc83
2024-06-21 00:45:06 +07:00
hiyouga
0844750bb9 tiny fix
Former-commit-id: 8d4f5093cf
2024-06-20 22:56:05 +08:00
hiyouga
030b4811c7 update patcher
Former-commit-id: 3b040e8e0f
2024-06-19 21:27:00 +08:00
hiyouga
5156114981 fix #4357
Former-commit-id: 4bd77d8563
2024-06-18 22:42:45 +08:00
hiyouga
7ef169ed39 fix #4326
Former-commit-id: e2665e71c7
2024-06-17 18:17:48 +08:00
ancv
988231026a update packing with sdpa and eager attention mode
Former-commit-id: 238f5c3d99
2024-06-16 02:25:47 +07:00
hiyouga
f25b8626bf support pissa
Former-commit-id: 8c1046d78a
2024-06-16 01:08:12 +08:00
hiyouga
c0c6b8075a tiny fix
Former-commit-id: 38b6b0f52e
2024-06-16 01:06:41 +08:00
ancv
9d9f8c6531 remove some unused params
Former-commit-id: 04315c3d92
2024-06-15 23:00:55 +07:00
hiyouga
2946153cea add license
Former-commit-id: d87108daa6
2024-06-15 17:54:33 +08:00
hiyouga
a3f4925c2c add test cases
Former-commit-id: b27269bd2b
2024-06-15 04:05:54 +08:00
hiyouga
833aa324c2 clean code
Former-commit-id: 2ed8270112
2024-06-13 01:58:16 +08:00
ancv
045eb155a2 implement efficient packing without cross-contamination attention
Former-commit-id: b2c367bc61
2024-06-12 11:56:01 +07:00