108 Commits

Author SHA1 Message Date
hiyouga
38505ae9e1 update accelerate ver for schedule_free optimizers
Former-commit-id: bdde35fd2e4a919c1d63ebfc9a0ea8ba0c97e14c
2024-09-09 22:51:08 +08:00
hiyouga
7ccb86b215 add docstrings, refactor logger
Former-commit-id: 54c69059379d77dc9046c144cbe2d0253de3a4da
2024-09-08 00:56:56 +08:00
hiyouga
3aa6a3e45b add e2e tests
Former-commit-id: 94d5b1bd8f49dabeb9e3c53d634cfb3c06b0241d
2024-09-05 21:52:28 +08:00
hoshi-hiyouga
de277a8ab8 Merge pull request #5372 from LDLINGLINGLING/main
增加了对minicpm3.0的适配'

Former-commit-id: 12743562639ccc6eb0caf170e7123d9844e2b4a6
2024-09-05 21:35:42 +08:00
liudan
1797fe50a4 根据代码规范修改了代码
Former-commit-id: 3d3fbaaff98da327e10bdebb4aedbdf1ec9565e8
2024-09-05 20:17:55 +08:00
hiyouga
9df7a26e6b video datasets
Former-commit-id: 8cafc7b055a854f483ad1c67f3d487ffd34b5f89
2024-09-05 02:04:17 +08:00
liudan
09cff03026 增加了对minicpm3.0的适配'
Former-commit-id: d7ba97be484bf781d6fe80252ea29eb505b261bb
2024-09-04 23:10:05 +08:00
hiyouga
d5ea05cfff update get template
Former-commit-id: dabad5570bf4a6b1044c963d8f27717030f373ef
2024-09-04 22:36:20 +08:00
hiyouga
a3d47818b7 fix #5344
Former-commit-id: d41d43a7c37cd10e34c9f399d1a346ffaee641c3
2024-09-04 03:06:06 +08:00
hiyouga
cb776752f6 fix mixed mm inputs and rlhf-v
Former-commit-id: 9967ccb3aef3ca557ad6eafb78c6c99866857008
2024-09-01 20:52:47 +08:00
hiyouga
09a2ecebc4 add test mm plugin
Former-commit-id: a2a8c0b92c49fb1ee65de271aec651e011dcabc4
2024-08-31 01:53:38 +08:00
hiyouga
a83756b5e9 refactor mm training
Former-commit-id: 3382317e32f88ed377d3e7759bdeaf0f2559d22a
2024-08-30 02:14:31 +08:00
simonJJJ
8a09b1e732 initial-commit
Former-commit-id: aeb85f200bd824748008dae6047c2607dfcdf174
2024-08-28 16:51:35 +08:00
hoshi-hiyouga
a7604a95c1 Merge pull request #5156 from YeQiuO/main
fix Llama-template's system prompt bug

Former-commit-id: 15be2963477e7d3a9fe4330d7701d457dc49b583
2024-08-20 00:09:03 +08:00
hoshi-hiyouga
103132aa99 Update template.py
Former-commit-id: ec72eeca521ba4ec71f0c52de9eec49da2cf0feb
2024-08-20 00:03:33 +08:00
hoshi-hiyouga
a921505f59 Update template.py
Former-commit-id: 5f3300ec5de564df23c94ebd9662c86708f37ddb
2024-08-19 23:40:16 +08:00
Huiyu Chen
66a7f4f128 Add SailorLLM template
Former-commit-id: 2502833a7755d653e8492cb7f1215dc0105b6ee0
2024-08-15 15:10:14 +08:00
“Wzw”
3e159a0a83 fix Llama-template's system prompt bug
Former-commit-id: bcbbf4506300fc132e68a39a9a6dfa5e61497c8b
2024-08-12 19:22:12 +08:00
hiyouga
b5146facff follow #5115
Former-commit-id: c87023d539875cd8e622d40212a5627c9c182fb8
2024-08-09 18:03:00 +08:00
hoshi-hiyouga
397e4daa5d Merge pull request #5115 from YeQiuO/main
fix: `Train on the last turn only` truncate bug
Former-commit-id: 51542cb15fea785d445ecf80bbad0364ebc0cb77
2024-08-09 17:58:27 +08:00
hoshi-hiyouga
54f57fb354 Update template.py
Former-commit-id: 4f62e1cb243d996d1764c3c86ca234847ad2c022
2024-08-09 16:27:42 +08:00
“Wzw”
0bd25c3a6b fix mask_history tiny bug
Former-commit-id: b5ca86cc07d38cf342e351aab16cce4319245792
2024-08-08 10:09:33 +08:00
moontidef
733cb9087b fix: fix the deepseekcoder template to avoid repeat problem
Former-commit-id: b82ecbedd0fecd85195217916cba3c21998bd10b
2024-08-05 23:55:45 +08:00
huangpan.foo
ee4c3f32d1 update deepseek template
Former-commit-id: 44e48e2b82929888b0880c00519102da4eb38ca8
2024-07-19 15:02:54 +08:00
hiyouga
7fcffb860d add codegeex4, internlm2.5
Former-commit-id: 53b1002fb74123095e7466c75b941a31a7cfba4d
2024-07-06 16:16:47 +08:00
hiyouga
8379a39776 fix processors
Former-commit-id: 9f33f1edf544807e498f60881f30b00149fe570f
2024-07-05 08:33:22 +08:00
hiyouga
ca7b65439d fix #4402 #4617
Deprecate reserved_label_len arg


Former-commit-id: 1771251ce3f6887b301dac10f3de7a253c5e5884
2024-07-01 01:19:27 +08:00
hiyouga
654116c0b1 fix #4556
Former-commit-id: 59e0b4f616736ede37cc37a13346b547f5a2d4e7
2024-06-26 19:43:16 +08:00
hiyouga
d519c2fde5 tiny fix
Former-commit-id: 41086059b12ecb7827eb390294e315068ff9c2e6
2024-06-25 01:15:19 +08:00
hoshi-hiyouga
b7f5cfde6e Update template.py
Former-commit-id: 1240bd57d8a21540c636a6da839e6b3112d1395a
2024-06-24 23:12:59 +08:00
mMrBun
c0e005e2ea Add tool_format to overwrite tool formatter template
Former-commit-id: 20e2e6fdcb0cd1771906be035745a2d9fcd3e138
2024-06-22 02:13:23 +08:00
hiyouga
98abb5c900 remove dup template
Former-commit-id: db9a1912e3551394039cc57b4913f03e8f9aa29d
2024-06-22 01:31:32 +08:00
hiyouga
3d72b1a856 fix jinja template
Former-commit-id: 2b596fb55ff689d2e488d9a9bbab98f70f356c3c
2024-06-19 20:03:50 +08:00
hiyouga
7735456561 fix templates
Former-commit-id: 4cff6a4ad55b24bf57db6be5cf817180c1ea5626
2024-06-19 17:44:05 +08:00
hiyouga
c9557241f6 fix bug
Former-commit-id: 6d2bf216ac3a48450e861148ce664dad717fd019
2024-06-19 03:49:23 +08:00
hiyouga
e73a235a38 use prefix to replace force system
Former-commit-id: 4f22eae8f405de918237d406e5e9847592925565
2024-06-19 03:39:52 +08:00
hiyouga
bccc852f76 fix tool formatter, allow parallel function #4362
Former-commit-id: cd75b1fe9d91fb52a9ae6de7435302ff06b4d933
2024-06-19 03:23:51 +08:00
hoshi-hiyouga
6db02615d4 Merge pull request #4173 from mMrBun/main
Implemented the tool_formatter and tool_extractor for glm4 and Qwen2 tool_format

Former-commit-id: c0ca42566c6aeccd8d384377510690eafef10995
2024-06-19 03:18:55 +08:00
hiyouga
2946153cea add license
Former-commit-id: d87108daa68bd40174b262be1ca65fe6e1b7ab56
2024-06-15 17:54:33 +08:00
mMrBun
daf472994d Merge branch 'hiyouga:main' into main
Former-commit-id: 0f2609ce19492f0bab9b4880ded228b5513e5907
2024-06-09 18:17:24 +08:00
mMrBun
18a86ea104 Implemented the tool_formatter and tool_extractor for glm4 tool_format
Former-commit-id: cb1cbcb293917e960cad8f0eac7a11a122ab644a
2024-06-09 18:16:15 +08:00
hiyouga
ce40d12692 release v0.8.0
Former-commit-id: 5aa4ce47567146cd97c61623018153b41d7c1278
2024-06-08 05:20:54 +08:00
hiyouga
8da149ba40 rename files
Former-commit-id: 74f96efef9bcd63f65d0190c901ff9be54ccd350
2024-06-07 00:09:06 +08:00
hiyouga
94c37490d1 support glm-4
Former-commit-id: f48f5e646e2da9e02333d027033141b0e75dfcf8
2024-06-05 15:16:38 +08:00
hiyouga
19a3262387 fix cohere system
Former-commit-id: d0aa36b8ad02287d97930101958456c523e699d3
2024-05-29 20:58:23 +08:00
hiyouga
c05cb3769f fix #3965
Former-commit-id: 0930f5869929634baa0881167d3d6c714afc63d9
2024-05-29 20:55:51 +08:00
hiyouga
a71a6a05c3 update readme
Former-commit-id: 89ca832740731dfb121175aa5c16b13bd4944011
2024-05-29 18:39:11 +08:00
hzhaoy
ce1be3da4b add TeleChat-12B/TeleChat-12B-v2 models
Former-commit-id: 0dd632fe9e5bbf08605d4b9c6887208b7a127317
2024-05-29 15:00:37 +08:00
Yimi81
7324984127 fix yi template
Former-commit-id: dc07413e7d0b138c89eacaef17596e83ef226540
2024-05-27 13:11:25 +00:00
hiyouga
0706dbf7e6 tiny fix
Former-commit-id: c1fdf81df6ade5da7be4eb66b715f0efd171d5aa
2024-05-27 20:54:26 +08:00