hiyouga
|
163cf2ba5c
|
update requires
Former-commit-id: 77666bd2278a3cfe5b567f4fe285b0f93871d166
|
2024-10-29 16:10:07 +08:00 |
|
hiyouga
|
4464a6ff5b
|
tiny fix
Former-commit-id: 451d271718a8026056d0f7d7b8ab333391d24ad4
|
2024-10-08 17:48:56 +08:00 |
|
hiyouga
|
995491594d
|
tiny fix
Former-commit-id: 76f2e5950483c669a15a961f0554442b6eb5c4a6
|
2024-09-05 23:41:16 +08:00 |
|
hiyouga
|
a83756b5e9
|
refactor mm training
Former-commit-id: 3382317e32f88ed377d3e7759bdeaf0f2559d22a
|
2024-08-30 02:14:31 +08:00 |
|
hiyouga
|
20013e130b
|
fix #5048
Former-commit-id: b7ca6c8dc14f689d0df16684a6121cc0ec24f8ba
|
2024-08-05 23:48:19 +08:00 |
|
hiyouga
|
0b26011181
|
fix gemma2 attention
Former-commit-id: 2f6af73da28c4f8321b625fd09ddec8bd4977b08
|
2024-07-13 23:33:45 +08:00 |
|
hiyouga
|
d7657d772d
|
tiny fix
Former-commit-id: 0c699de39de06eac96af67e8dd4fc4c53335b17e
|
2024-07-04 03:47:05 +08:00 |
|
hiyouga
|
7b3c1f29ff
|
fix packing for eager/sdpa attn
Former-commit-id: 6fd6aa4530f81a2ed306eeb2a5167607288b62c6
|
2024-07-04 01:52:43 +08:00 |
|
hiyouga
|
bfdaadcc40
|
update packing
Former-commit-id: cce7083024bed4c7429ddc8288d1c9190fde29f5
|
2024-07-04 01:10:55 +08:00 |
|
hoshi-hiyouga
|
51c75985b8
|
Update packing.py
Former-commit-id: a36e8f2dd50e0f1c589457a7e785fdbc905d561d
|
2024-07-03 23:36:01 +08:00 |
|
hiyouga
|
13cec0cc2f
|
update func name
Former-commit-id: c346f79f99db5296000e4d22a65e53c26e85b344
|
2024-07-03 23:29:33 +08:00 |
|
hiyouga
|
e671ed520b
|
update arg name
Former-commit-id: 8a6a7b9c8a876da9c16e5ada7df461eb8cabee21
|
2024-07-03 23:23:24 +08:00 |
|
ancv
|
5319447aa5
|
move configure_packing to llamafactory.model.patcher and fix constants
Former-commit-id: 770f75dc8363bfa284a72159ff8ad25ec9abe4e0
|
2024-06-21 00:45:06 +07:00 |
|
ancv
|
988231026a
|
update packing with sdpa and eager attention mode
Former-commit-id: 238f5c3d99809c6ae2571b59bdce8d8ea3c700b9
|
2024-06-16 02:25:47 +07:00 |
|
ancv
|
9d9f8c6531
|
remove some unused params
Former-commit-id: 04315c3d92ecc25537e45d5807cb38bc290dcb16
|
2024-06-15 23:00:55 +07:00 |
|
ancv
|
045eb155a2
|
implement efficient packing without cross-contamination attention
Former-commit-id: b2c367bc61c2778dc359613dca496d9e134c2743
|
2024-06-12 11:56:01 +07:00 |
|