hiyouga
|
ff6fc666c1
|
update hparams
Former-commit-id: 575a02a23d
|
2024-07-03 23:18:58 +08:00 |
|
ancv
|
7f42932957
|
move efficient_packing from data_args to model_args
Former-commit-id: e8e13b0942
|
2024-07-02 18:37:55 +07:00 |
|
hoshi-hiyouga
|
2452f57cd7
|
Merge branch 'main' into main
Former-commit-id: e8e6af2651
|
2024-07-01 21:01:09 +08:00 |
|
hiyouga
|
ca7b65439d
|
fix #4402 #4617
Deprecate reserved_label_len arg
Former-commit-id: 1771251ce3
|
2024-07-01 01:19:27 +08:00 |
|
hiyouga
|
654116c0b1
|
fix #4556
Former-commit-id: 59e0b4f616
|
2024-06-26 19:43:16 +08:00 |
|
hiyouga
|
d519c2fde5
|
tiny fix
Former-commit-id: 41086059b1
|
2024-06-25 01:15:19 +08:00 |
|
hoshi-hiyouga
|
709bbc1d92
|
Merge pull request #4417 from mMrBun/main
Add tool_format parameter to rewrite templates for different function call formats.
Former-commit-id: def6d280db
|
2024-06-24 23:17:55 +08:00 |
|
hoshi-hiyouga
|
b7f5cfde6e
|
Update template.py
Former-commit-id: 1240bd57d8
|
2024-06-24 23:12:59 +08:00 |
|
hoshi-hiyouga
|
673f27a59e
|
Update loader.py
Former-commit-id: dddfd516ee
|
2024-06-24 23:06:18 +08:00 |
|
hiyouga
|
47651a94a3
|
fix #4410
Former-commit-id: fca893d73c
|
2024-06-24 22:34:31 +08:00 |
|
mMrBun
|
c0e005e2ea
|
Add tool_format to overwrite tool formatter template
Former-commit-id: 20e2e6fdcb
|
2024-06-22 02:13:23 +08:00 |
|
hiyouga
|
98abb5c900
|
remove dup template
Former-commit-id: db9a1912e3
|
2024-06-22 01:31:32 +08:00 |
|
hiyouga
|
3d72b1a856
|
fix jinja template
Former-commit-id: 2b596fb55f
|
2024-06-19 20:03:50 +08:00 |
|
hiyouga
|
7735456561
|
fix templates
Former-commit-id: 4cff6a4ad5
|
2024-06-19 17:44:05 +08:00 |
|
hiyouga
|
c9557241f6
|
fix bug
Former-commit-id: 6d2bf216ac
|
2024-06-19 03:49:23 +08:00 |
|
hiyouga
|
e73a235a38
|
use prefix to replace force system
Former-commit-id: 4f22eae8f4
|
2024-06-19 03:39:52 +08:00 |
|
hiyouga
|
bccc852f76
|
fix tool formatter, allow parallel function #4362
Former-commit-id: cd75b1fe9d
|
2024-06-19 03:23:51 +08:00 |
|
hoshi-hiyouga
|
6db02615d4
|
Merge pull request #4173 from mMrBun/main
Implemented the tool_formatter and tool_extractor for glm4 and Qwen2 tool_format
Former-commit-id: c0ca42566c
|
2024-06-19 03:18:55 +08:00 |
|
hiyouga
|
c0c6b8075a
|
tiny fix
Former-commit-id: 38b6b0f52e
|
2024-06-16 01:06:41 +08:00 |
|
ancv
|
9d9f8c6531
|
remove some unused params
Former-commit-id: 04315c3d92
|
2024-06-15 23:00:55 +07:00 |
|
hiyouga
|
2946153cea
|
add license
Former-commit-id: d87108daa6
|
2024-06-15 17:54:33 +08:00 |
|
hiyouga
|
8fccaf20c5
|
fix #4221
Former-commit-id: 6baafd4eb3
|
2024-06-13 02:48:21 +08:00 |
|
ancv
|
045eb155a2
|
implement efficient packing without cross-contamination attention
Former-commit-id: b2c367bc61
|
2024-06-12 11:56:01 +07:00 |
|
hoshi-hiyouga
|
bf3de9bfe8
|
Update pretrain.py
Former-commit-id: 0c29233237
|
2024-06-11 17:02:14 +08:00 |
|
d
|
da39715085
|
经过大量的增量预训练,进行对比试验,发现这个bug:llama3在预训练时使用的tokenizer.eos_toke是'<|end_of_text|>' ,这里在每条数据后面也得用这个,而不是'<|eot_id|>',否则很容易导致严重的性能下降
Former-commit-id: 6979f3f848
|
2024-06-11 16:23:40 +08:00 |
|
mMrBun
|
b6d63b3324
|
Optimize the handling of QWEN2 in scenarios involving multiple tool calls.
Former-commit-id: 950e360ca0
|
2024-06-10 02:00:14 +08:00 |
|
mMrBun
|
3f11ab800f
|
Removed unnecessary comments.
Former-commit-id: 6ed0b0c800
|
2024-06-09 18:25:22 +08:00 |
|
mMrBun
|
daf472994d
|
Merge branch 'hiyouga:main' into main
Former-commit-id: 0f2609ce19
|
2024-06-09 18:17:24 +08:00 |
|
mMrBun
|
18a86ea104
|
Implemented the tool_formatter and tool_extractor for glm4 tool_format
Former-commit-id: cb1cbcb293
|
2024-06-09 18:16:15 +08:00 |
|
hiyouga
|
ce40d12692
|
release v0.8.0
Former-commit-id: 5aa4ce4756
|
2024-06-08 05:20:54 +08:00 |
|
hiyouga
|
c6f5f69644
|
update data processors
Former-commit-id: ccc8b64cc2
|
2024-06-07 04:15:40 +08:00 |
|
hoshi-hiyouga
|
4953ded639
|
Merge pull request #4009 from AlongWY/main
supervised packing with greedy knapsack algorithm
Former-commit-id: 181dbb0d05
|
2024-06-07 03:48:46 +08:00 |
|
hoshi-hiyouga
|
e3ef239bc0
|
Update supervised.py
Former-commit-id: c09ad8bab3
|
2024-06-07 03:42:08 +08:00 |
|
hoshi-hiyouga
|
fd7bd911a6
|
Update supervised.py
Former-commit-id: 788e8232fc
|
2024-06-07 03:38:23 +08:00 |
|
hoshi-hiyouga
|
21df5f0bd0
|
Update supervised.py
Former-commit-id: 8cecade708
|
2024-06-07 03:38:04 +08:00 |
|
hiyouga
|
8da149ba40
|
rename files
Former-commit-id: 74f96efef9
|
2024-06-07 00:09:06 +08:00 |
|
hiyouga
|
e0aadd4b34
|
fix ppo dataset bug #4012
Former-commit-id: 149610c636
|
2024-06-06 19:03:20 +08:00 |
|
hiyouga
|
94c37490d1
|
support glm-4
Former-commit-id: f48f5e646e
|
2024-06-05 15:16:38 +08:00 |
|
hiyouga
|
0eff6a66d5
|
tiny fix
Former-commit-id: 5a13b3baa6
|
2024-06-04 00:31:10 +08:00 |
|
hiyouga
|
8ecf606230
|
fix #3992
Former-commit-id: a18acf2abe
|
2024-06-04 00:17:36 +08:00 |
|
hiyouga
|
64d24842fe
|
fix data loader hint
Former-commit-id: 49b1e88e3d
|
2024-06-03 18:28:27 +08:00 |
|
ylfeng
|
62d55b71a3
|
remove empty line
Former-commit-id: b47e317447
|
2024-05-31 21:43:08 +08:00 |
|
ylfeng
|
0feb2ad35c
|
fix eos
Former-commit-id: 84aee57901
|
2024-05-31 21:40:41 +08:00 |
|
ylfeng
|
8350e508d3
|
supervised packing with greedy knapsack algorithm
Former-commit-id: f9db439cb7
|
2024-05-31 15:33:54 +08:00 |
|
hoshi-hiyouga
|
9b6bdf9449
|
Merge pull request #3829 from seanzhang-zhichen/add_dataset_sample_num
Add dataset sample num
Former-commit-id: 483eb47e5d
|
2024-05-30 00:25:45 +08:00 |
|
hoshi-hiyouga
|
7b83c550ab
|
Update loader.py
Former-commit-id: ca5dd7c6c1
|
2024-05-30 00:20:20 +08:00 |
|
hoshi-hiyouga
|
9fc713da89
|
Update loader.py
Former-commit-id: f9a88b89ca
|
2024-05-30 00:17:21 +08:00 |
|
hoshi-hiyouga
|
c0f11a280e
|
Update loader.py
Former-commit-id: b55fb611c5
|
2024-05-30 00:12:12 +08:00 |
|
hoshi-hiyouga
|
69a51cacb1
|
Update parser.py
Former-commit-id: 51dd454337
|
2024-05-30 00:05:20 +08:00 |
|
hiyouga
|
19a3262387
|
fix cohere system
Former-commit-id: d0aa36b8ad
|
2024-05-29 20:58:23 +08:00 |
|