Erich Schubert
|
05abe47c8b
|
Print help if no arguments given
Former-commit-id: 08dfb7ec636fd5bfbb30dac9d5fba6e32bfc6728
|
2024-06-21 09:14:21 +02:00 |
|
ancv
|
6c185a2c57
|
move configure_packing to llamafactory.model.patcher and fix constants
Former-commit-id: 9c5e972c9c81957f2e9e30bf284ef1c076de9fd0
|
2024-06-21 00:45:06 +07:00 |
|
hiyouga
|
af2cb33bb2
|
tiny fix
Former-commit-id: 2d8d47f6126d68db1701ed18fc31310c6f14dd49
|
2024-06-20 22:56:05 +08:00 |
|
hiyouga
|
0edccc11a5
|
improve llamaboard
Former-commit-id: e606ab35c0eced667dde7137c2d72848f264c96c
|
2024-06-19 23:46:03 +08:00 |
|
hiyouga
|
b2f5c0e0db
|
fix llamaboard abort
Former-commit-id: 9ef609a2c0185040e531dea3829a6f481539cdea
|
2024-06-19 23:22:28 +08:00 |
|
hiyouga
|
5f5d4c1923
|
update patcher
Former-commit-id: afb365e515d615dd62f791622450debab60ce5cc
|
2024-06-19 21:27:00 +08:00 |
|
hiyouga
|
a7d7f79855
|
set dev version
Former-commit-id: 221665345d97f839ce4ba8d54643da30c71b6083
|
2024-06-19 21:08:16 +08:00 |
|
hiyouga
|
b631bdc5b7
|
release v0.8.2
Former-commit-id: 3050bbe51d46acd8473275d2713fc28932e4a3d3
|
2024-06-19 20:42:09 +08:00 |
|
hiyouga
|
c65f7e9bd5
|
fix jinja template
Former-commit-id: 0ebf2e2ee23918d28b0cbb20ba456732d6eedfbb
|
2024-06-19 20:03:50 +08:00 |
|
hiyouga
|
3e0fa4a8da
|
fix templates
Former-commit-id: 6f357d59b73309c5955683008632e7f320e7dcb1
|
2024-06-19 17:44:05 +08:00 |
|
Jonery
|
fa3150548e
|
Cleaner integration.
Former-commit-id: 26d4b05d424bd71f570195dd433258caf6465d92
|
2024-06-19 12:29:40 +08:00 |
|
hiyouga
|
235ed85b0f
|
fix bug
Former-commit-id: 412139eaa2fde98ba19e1257d21144382a59f0d6
|
2024-06-19 03:49:23 +08:00 |
|
hiyouga
|
1ca639a777
|
use prefix to replace force system
Former-commit-id: 731d9a964f1c3dbfb83825524d697831e691fb9d
|
2024-06-19 03:39:52 +08:00 |
|
hiyouga
|
e36a994fe6
|
fix tool formatter, allow parallel function #4362
Former-commit-id: b8f16c976db4ecec1cc8558851c8cbfb6a5b7e9c
|
2024-06-19 03:23:51 +08:00 |
|
hoshi-hiyouga
|
19ffcfea76
|
Merge pull request #4173 from mMrBun/main
Implemented the tool_formatter and tool_extractor for glm4 and Qwen2 tool_format
Former-commit-id: 36b02ceed40198ecd5d559ee4ebef9205442ded2
|
2024-06-19 03:18:55 +08:00 |
|
hiyouga
|
665df5d733
|
add deepseek coder v2 #4346
Former-commit-id: d83d3846d8e3bf5c40d4b90c24e2c5909ec61864
|
2024-06-18 22:53:54 +08:00 |
|
hiyouga
|
4bc0bea0e9
|
fix #4357
Former-commit-id: a6741bba8cebd16a6a3f97a2dc81057d0e27eb39
|
2024-06-18 22:42:45 +08:00 |
|
hiyouga
|
372da52d4a
|
fix #4335
Former-commit-id: 2ab449adbb160f339a0586edeb846fa311ad8382
|
2024-06-18 22:08:56 +08:00 |
|
Jonery
|
870a54ac84
|
fix typo
Former-commit-id: d4bee3716dbf8a84564d5bcc2059172604819f3e
|
2024-06-18 12:39:26 +08:00 |
|
Jonery
|
12fcfc2b72
|
Support distributed BAdam.
Former-commit-id: bdcb986e37975911c190a74d3e60bb77aa2033bd
|
2024-06-18 12:27:47 +08:00 |
|
hiyouga
|
875270b851
|
lint
Former-commit-id: a19a7ac99af62b6715c96274f6350b124a784331
|
2024-06-17 22:35:56 +08:00 |
|
hiyouga
|
43fab306b6
|
update chat engine #4335
Former-commit-id: b163df7de48777e4319c9ccc736b0acdd5f473ed
|
2024-06-17 19:07:17 +08:00 |
|
Jonery
|
95ae30f678
|
Merge remote-tracking branch 'upstream/main'
Former-commit-id: 37834a7e79473ccf50ad7f67745b97c274c326d9
|
2024-06-17 18:44:51 +08:00 |
|
Jonery
|
ba303fd1aa
|
adapt for badam with ds zero3
Former-commit-id: fff2a020ec8713022bd8145f4a7168168ea07ca4
|
2024-06-17 18:18:10 +08:00 |
|
hiyouga
|
60d9896a70
|
fix #4326
Former-commit-id: 3c2c45812a720d92f7f5b15b9f03370fe6bf069e
|
2024-06-17 18:17:48 +08:00 |
|
ancv
|
dd7a1dbfae
|
update packing with sdpa and eager attention mode
Former-commit-id: 285636ba3a57a1038b2f2fd4cf909a1ca07708d4
|
2024-06-16 02:25:47 +07:00 |
|
hoshi-hiyouga
|
ca67b7a568
|
Update parser.py
Former-commit-id: d10c97193d08bd368aca1a72f0d1d8a96c76765d
|
2024-06-16 02:57:00 +08:00 |
|
hiyouga
|
727943f078
|
fix tol
Former-commit-id: bdb54bcb477126687db789bd89f2df84e424a2a3
|
2024-06-16 01:38:44 +08:00 |
|
hiyouga
|
32f45c9e91
|
support pissa
Former-commit-id: ef8e45f2eaf466c54e9a671512a2974575677b08
|
2024-06-16 01:08:12 +08:00 |
|
hiyouga
|
05f3a3c944
|
tiny fix
Former-commit-id: f7f440986b0ae3b38ea9f2da80789629d4f79ea1
|
2024-06-16 01:06:41 +08:00 |
|
ancv
|
f91fe10985
|
remove some unused params
Former-commit-id: fef8132c50505a5fb6a246bd024491bd31798a3c
|
2024-06-15 23:00:55 +07:00 |
|
hiyouga
|
14f7bfc545
|
use fixture
Former-commit-id: 10761985691b9f934f7689c1f82aa6dd68febcca
|
2024-06-15 20:06:17 +08:00 |
|
hiyouga
|
7f90b0cd20
|
add tests
Former-commit-id: 484634ee9c982e82e919ff67d507e0210345182d
|
2024-06-15 19:51:20 +08:00 |
|
hiyouga
|
308abfec6c
|
add minicpm #4227
Former-commit-id: e1bb18ce60be9a1b203989def30f1b9194286325
|
2024-06-15 17:58:52 +08:00 |
|
hiyouga
|
bb88536166
|
add license
Former-commit-id: 69cfc98d7c81756a5ab6bf962240e393e449fef0
|
2024-06-15 17:54:33 +08:00 |
|
hiyouga
|
2af932d969
|
disable DP
Former-commit-id: c18fd609d268389f3e65274992045a6c9f8e6c1f
|
2024-06-15 04:57:19 +08:00 |
|
hiyouga
|
c29fa61a9c
|
fix #4292
Former-commit-id: 4cd4c179d24eab0fcaec2b29b9dd71970f877fe8
|
2024-06-15 04:47:13 +08:00 |
|
hiyouga
|
a30931fe0f
|
fix #4295
Former-commit-id: 08f657868f9d605b837c5d8c2946a25cc05c8735
|
2024-06-15 04:34:55 +08:00 |
|
hiyouga
|
3ff9b87012
|
add test cases
Former-commit-id: 731176ff34cdf0cbf6b41c40c69f4ceb54c2daf6
|
2024-06-15 04:05:54 +08:00 |
|
hiyouga
|
dbd1458adf
|
add quant check in webui export tab
Former-commit-id: 6455ca07061ae9858cd7bc996b28be1fde697a3d
|
2024-06-13 03:19:18 +08:00 |
|
hiyouga
|
49b58fd6af
|
fix #4221
Former-commit-id: 05a3be4853b941909e7d193c31e8d62c8c5f879b
|
2024-06-13 02:48:21 +08:00 |
|
hiyouga
|
103a507b39
|
fix #4209
DeepSpeed ZeRO3 has inflight param error when calling model.eval()
Former-commit-id: 4be013f18ea6a35b5a11db98db5f0670ffb41619
|
2024-06-13 02:25:50 +08:00 |
|
hiyouga
|
0a75224f62
|
clean code
Former-commit-id: f54cafd5c7f0383370d1a2f357834a61a97397ce
|
2024-06-13 01:58:16 +08:00 |
|
hoshi-hiyouga
|
04d7629abf
|
Merge pull request #4246 from hzhaoy/adapt-vllm-v0.5.0
adapt vllm==0.5.0
Former-commit-id: 1068e25fc8b89f11cc79b164ee4aef9ce137ad4c
|
2024-06-13 01:54:02 +08:00 |
|
hiyouga
|
5080f2314c
|
fix lint
Former-commit-id: b170165679317af2b3f03633afac27661b3deb06
|
2024-06-13 00:48:44 +08:00 |
|
hzhaoy
|
799873aa14
|
adapt vllm==0.5.0
Former-commit-id: 02afd9ff64f23e6707ac739ae1269f41bd70c340
|
2024-06-12 18:29:03 +08:00 |
|
hiyouga
|
6392d45ea7
|
fix #4242
Former-commit-id: cf260e7af03f49aa5e3d6daf3b27738ff9b9bcb8
|
2024-06-12 16:50:11 +08:00 |
|
Arthur Kim
|
16c7c92396
|
Support vllm==0.5.0
Former-commit-id: e7a8ffd7af21bc3759f055033ba2209fa7a1be0e
|
2024-06-12 16:49:12 +09:00 |
|
ancv
|
c7ab302c69
|
implement efficient packing without cross-contamination attention
Former-commit-id: a64a5305c0da5ef092d4cc26faf829bb44de65d1
|
2024-06-12 11:56:01 +07:00 |
|
hoshi-hiyouga
|
7598b37543
|
Merge pull request #4204 from dignfei/main
fixbug:llama3在增量预训练时应该使用<|end_of_text|>标识文本的结束
Former-commit-id: e566342636faf0031a0ba5d5dd4fcff8401a2b76
|
2024-06-11 17:06:10 +08:00 |
|