hoshi-hiyouga
|
37c6a0c6dc
|
fix #4917
Former-commit-id: 26082fc6c9
|
2024-07-22 11:28:31 +08:00 |
|
hiyouga
|
14bc7b0551
|
fix up
Former-commit-id: 29ebcd75d5
|
2024-07-15 01:04:56 +08:00 |
|
hiyouga
|
12e0e5d0d7
|
tiny fix
Former-commit-id: d3c01552e0
|
2024-07-14 10:56:45 +08:00 |
|
hiyouga
|
0b26011181
|
fix gemma2 attention
Former-commit-id: 2f6af73da2
|
2024-07-13 23:33:45 +08:00 |
|
hiyouga
|
d7657d772d
|
tiny fix
Former-commit-id: 0c699de39d
|
2024-07-04 03:47:05 +08:00 |
|
hiyouga
|
7b3c1f29ff
|
fix packing for eager/sdpa attn
Former-commit-id: 6fd6aa4530
|
2024-07-04 01:52:43 +08:00 |
|
hoshi-hiyouga
|
a38ff842d0
|
Merge pull request #4224 from chuan298/main
Implement efficient packing without cross-contamination attention
Former-commit-id: 87d9b2d005
|
2024-07-04 01:18:54 +08:00 |
|
hiyouga
|
bfdaadcc40
|
update packing
Former-commit-id: cce7083024
|
2024-07-04 01:10:55 +08:00 |
|
hoshi-hiyouga
|
51c75985b8
|
Update packing.py
Former-commit-id: a36e8f2dd5
|
2024-07-03 23:36:01 +08:00 |
|
hiyouga
|
13cec0cc2f
|
update func name
Former-commit-id: c346f79f99
|
2024-07-03 23:29:33 +08:00 |
|
hiyouga
|
e671ed520b
|
update arg name
Former-commit-id: 8a6a7b9c8a
|
2024-07-03 23:23:24 +08:00 |
|
hiyouga
|
104151d558
|
tiny fix
Former-commit-id: 8b1172b910
|
2024-07-03 02:31:50 +08:00 |
|
ancv
|
7f42932957
|
move efficient_packing from data_args to model_args
Former-commit-id: e8e13b0942
|
2024-07-02 18:37:55 +07:00 |
|
hoshi-hiyouga
|
2452f57cd7
|
Merge branch 'main' into main
Former-commit-id: e8e6af2651
|
2024-07-01 21:01:09 +08:00 |
|
hiyouga
|
de4de5b5ab
|
tiny fix
Former-commit-id: 8c41a0aa6d
|
2024-07-01 03:55:20 +08:00 |
|
hiyouga
|
bbc37b2880
|
fix #4398 #4592
Former-commit-id: d74244d568
|
2024-06-30 21:28:51 +08:00 |
|
hiyouga
|
2b006beab1
|
loose gemma2 attention
Former-commit-id: 2f4b89ace1
|
2024-06-29 01:42:14 +08:00 |
|
hiyouga
|
87e60f8bac
|
bf16 by default, gemma2 attns
Gemma2 finetuning cannot work until merging https://github.com/huggingface/transformers/pull/31674
Former-commit-id: 4d35e218b1
|
2024-06-28 06:00:26 +08:00 |
|
hiyouga
|
58607ec1b0
|
add quant checks
Former-commit-id: 96a5044394
|
2024-06-27 01:12:25 +08:00 |
|
hiyouga
|
d2d9fa4abb
|
support HQQ/EETQ #4113
Former-commit-id: ad144c2265
|
2024-06-27 00:29:42 +08:00 |
|
hiyouga
|
6b2733ce12
|
improve autogptq integration
Former-commit-id: addca926de
|
2024-06-26 22:11:44 +08:00 |
|
hiyouga
|
0ae1302e41
|
fix #4432
Former-commit-id: 1e9d0aa1e4
|
2024-06-25 02:34:04 +08:00 |
|
hiyouga
|
47651a94a3
|
fix #4410
Former-commit-id: fca893d73c
|
2024-06-24 22:34:31 +08:00 |
|
stceum
|
9aa640f27b
|
Bug Fix: off is parsed as False in yaml file, changed to disabled to avoid this.
Former-commit-id: 3ed063f281
|
2024-06-24 20:39:31 +08:00 |
|
ancv
|
5319447aa5
|
move configure_packing to llamafactory.model.patcher and fix constants
Former-commit-id: 770f75dc83
|
2024-06-21 00:45:06 +07:00 |
|
hiyouga
|
0844750bb9
|
tiny fix
Former-commit-id: 8d4f5093cf
|
2024-06-20 22:56:05 +08:00 |
|
hiyouga
|
030b4811c7
|
update patcher
Former-commit-id: 3b040e8e0f
|
2024-06-19 21:27:00 +08:00 |
|
hiyouga
|
5156114981
|
fix #4357
Former-commit-id: 4bd77d8563
|
2024-06-18 22:42:45 +08:00 |
|
hiyouga
|
7ef169ed39
|
fix #4326
Former-commit-id: e2665e71c7
|
2024-06-17 18:17:48 +08:00 |
|
ancv
|
988231026a
|
update packing with sdpa and eager attention mode
Former-commit-id: 238f5c3d99
|
2024-06-16 02:25:47 +07:00 |
|
hiyouga
|
f25b8626bf
|
support pissa
Former-commit-id: 8c1046d78a
|
2024-06-16 01:08:12 +08:00 |
|
hiyouga
|
c0c6b8075a
|
tiny fix
Former-commit-id: 38b6b0f52e
|
2024-06-16 01:06:41 +08:00 |
|
ancv
|
9d9f8c6531
|
remove some unused params
Former-commit-id: 04315c3d92
|
2024-06-15 23:00:55 +07:00 |
|
hiyouga
|
2946153cea
|
add license
Former-commit-id: d87108daa6
|
2024-06-15 17:54:33 +08:00 |
|
hiyouga
|
a3f4925c2c
|
add test cases
Former-commit-id: b27269bd2b
|
2024-06-15 04:05:54 +08:00 |
|
hiyouga
|
833aa324c2
|
clean code
Former-commit-id: 2ed8270112
|
2024-06-13 01:58:16 +08:00 |
|
ancv
|
045eb155a2
|
implement efficient packing without cross-contamination attention
Former-commit-id: b2c367bc61
|
2024-06-12 11:56:01 +07:00 |
|
hiyouga
|
8c574eb3cb
|
fix deepspeed version
Former-commit-id: cca6f35108
|
2024-06-11 16:52:36 +08:00 |
|
hiyouga
|
5834651c4a
|
fix #4198
Former-commit-id: 89f2bd8c8c
|
2024-06-11 15:38:38 +08:00 |
|
hiyouga
|
e3baa5aa08
|
tiny fix
Former-commit-id: 3f24337a8a
|
2024-06-11 01:04:16 +08:00 |
|
hiyouga
|
2f164c2c41
|
fix #4160
The split heads should be concatenated in dim=2
Former-commit-id: a793e8456b
|
2024-06-11 00:37:17 +08:00 |
|
hiyouga
|
3b244a69dc
|
fix #2666
Former-commit-id: c907d81667
|
2024-06-10 21:24:15 +08:00 |
|
hiyouga
|
4f0ce9be4e
|
reorganize adapter code
Former-commit-id: 54cd743ebf
|
2024-06-08 00:47:23 +08:00 |
|
hoshi-hiyouga
|
bad35d1730
|
fix #4139
Former-commit-id: cfd62283a9
|
2024-06-08 00:45:02 +08:00 |
|
hiyouga
|
a8318723a4
|
add resume args in webui
Former-commit-id: 06e5d136a4
|
2024-06-08 00:22:16 +08:00 |
|
hiyouga
|
8da149ba40
|
rename files
Former-commit-id: 74f96efef9
|
2024-06-07 00:09:06 +08:00 |
|
hiyouga
|
6cbc66a602
|
fix torch gc
Former-commit-id: 451b6693c0
|
2024-06-06 20:30:25 +08:00 |
|
hiyouga
|
3fcb678d00
|
support train from scratch #4033 #4075
Former-commit-id: a12a506c3d
|
2024-06-06 02:43:19 +08:00 |
|
hiyouga
|
b88ecd71fd
|
fix full/freeze tuning for mllm
Former-commit-id: 08564838bd
|
2024-05-27 20:37:57 +08:00 |
|
BUAADreamer
|
606240aec0
|
add regex of only tune lm and mm_proj
Former-commit-id: 57eb13b75d
|
2024-05-27 18:59:00 +08:00 |
|