hoshi-hiyouga
|
ac32264bfe
|
Update sft.sh
Former-commit-id: 2b4b1562e91bbb02e345e71b7721da9333c0791b
|
2024-04-16 17:25:40 +08:00 |
|
hoshi-hiyouga
|
bea83bfa18
|
Update requirements.txt
Former-commit-id: 1e45537ca0bb4d49b4147df01122e365b3d617e4
|
2024-04-16 17:10:17 +08:00 |
|
hoshi-hiyouga
|
fc93f9f33c
|
Update setup.py
Former-commit-id: 5df30ea166aff29d48ff83a22ac6ef1611ce3e35
|
2024-04-16 17:10:02 +08:00 |
|
Jonery
|
3ba7e96f66
|
remove badam from core requirements
Former-commit-id: fa5898944a3867ac5108dd0d579ca0677c87d3d6
|
2024-04-16 12:25:50 +08:00 |
|
Jonery
|
2ba03e6ef3
|
resolve gradient checkpointing issue.
Former-commit-id: 6df9135d063bb6102f0cbcdf0d702076f5febbae
|
2024-04-16 12:05:27 +08:00 |
|
Jonery
|
22188f1fa3
|
Feature BAdam
Former-commit-id: d8d2807fbcf587c37f7fd34a23e9397d2775ceed
|
2024-04-15 23:15:27 +08:00 |
|
hiyouga
|
be206df674
|
update examples
Former-commit-id: 369294b31c8a03a1cafcee83eb31a817007d3c49
|
2024-04-15 22:14:34 +08:00 |
|
hoshi-hiyouga
|
29e9b6fc76
|
Merge pull request #3261 from khazic/main
Added specimens for single-card full parameter prediction
Former-commit-id: 60df2a9519fbd8215c3afacc831b0cc89006457a
|
2024-04-15 16:30:57 +08:00 |
|
hoshi-hiyouga
|
740d89e9df
|
Merge pull request #3276 from liu-zichen/fix_mixtral
fix: turn on output_router_logits of mixtral
Former-commit-id: 07bbaf5c67d00a152e5304e81b15fd9189e7bb99
|
2024-04-15 15:38:16 +08:00 |
|
hiyouga
|
506276c9cb
|
fix #3273
Former-commit-id: 3b20c89b342a068356ffc29c3724b645775c65db
|
2024-04-15 15:32:58 +08:00 |
|
liuzc
|
44c86150c9
|
fix: mixtral output_router_logits
Former-commit-id: ab3171ea97ec968b972287287ef9ee2502c6d37c
|
2024-04-15 12:11:49 +08:00 |
|
khazic
|
f8378037a8
|
Upgrade README.md
Former-commit-id: 697f768d7185789ee054c94f4f161a65b8a505bc
|
2024-04-13 20:50:49 +08:00 |
|
khazic
|
8404b94bf6
|
Added specimens for single-card full parameter prediction
Former-commit-id: d8d4fb9fa4b0e1950a453682e5e186f34f085dee
|
2024-04-13 20:45:19 +08:00 |
|
hiyouga
|
b1ae554c83
|
fix #3247
Former-commit-id: bb67c66f80627805b585d157ba807c0ce378d3f2
|
2024-04-12 17:41:33 +08:00 |
|
hiyouga
|
64da8145bf
|
fix model card
Former-commit-id: 920e7149bf2b559c9829aa4b11cfb6d00bbb2f9e
|
2024-04-12 17:11:59 +08:00 |
|
hiyouga
|
4cb8cc563d
|
fix #3238
Former-commit-id: 4d7e81ab4722d13bec6ca1af141f94bdc74d0883
|
2024-04-12 14:28:11 +08:00 |
|
hiyouga
|
5d7887fd5d
|
set dev version
Former-commit-id: f6cc76571d2c789675883a18e0db3d0c61f33808
|
2024-04-11 20:27:34 +08:00 |
|
hiyouga
|
a97f8d1fa8
|
release v0.6.2
Former-commit-id: f92ad0a62d957b595f6a76a5403216b163eb3d17
|
2024-04-11 20:08:51 +08:00 |
|
hiyouga
|
3283cc3b31
|
Merge branch 'main' of https://github.com/hiyouga/LLaMA-Factory
Former-commit-id: 23ff02c1fd3787daf0bc6ac237c8897d02f726e4
|
2024-04-10 23:58:18 +08:00 |
|
hiyouga
|
8f6f06ceb5
|
fix #3225
Former-commit-id: 94110ecf27c32e263f1f2ee61842a3a301b9e089
|
2024-04-10 23:57:59 +08:00 |
|
hoshi-hiyouga
|
72602ed3cf
|
Merge pull request #3201 from kno10/patch-1 and fix #3200
Pass additional_target to unsloth
Former-commit-id: 080a96c52f489fda0d315a77e26c4f6f5d69784a
|
2024-04-10 00:58:48 +08:00 |
|
hoshi-hiyouga
|
db51e05205
|
Update adapter.py
Former-commit-id: 720fde3683529ed7e08ac27c7c4598c6bdc30d44
|
2024-04-10 00:57:51 +08:00 |
|
hoshi-hiyouga
|
bfb090ed7a
|
Update adapter.py
Former-commit-id: a84b8d17dbf221259212e81931d80bcdd6284ad7
|
2024-04-10 00:57:30 +08:00 |
|
Erich Schubert
|
cc2ff3065f
|
Pass additional_target to unsloth
Fixes #3200
Former-commit-id: f8f87f5b0549cba6a011749c42064047f82ba577
|
2024-04-09 17:53:40 +02:00 |
|
hiyouga
|
f8609236ab
|
fix quant infer and qwen2moe
Former-commit-id: b75d16767f35c36e2cf2aaab8a3844135085bccf
|
2024-04-09 17:12:59 +08:00 |
|
hiyouga
|
4739c45e94
|
tiny fix
Former-commit-id: d8f1ff51d4c920d4d0aeb9d53db29d1efb733c85
|
2024-04-08 21:28:39 +08:00 |
|
hoshi-hiyouga
|
cdd24a2f2d
|
Merge pull request #3161 from hiyouga/feature/add-mediatek-model
support Breeze-7B
Former-commit-id: af92ac8b62b919a75673011a1c56832e67882ee8
|
2024-04-08 20:56:51 +08:00 |
|
codingma
|
b7cc559649
|
add empty line
Former-commit-id: 1c6c2e611d10e9fa662e3f4e1e7d23b80ae496cb
|
2024-04-07 18:28:08 +08:00 |
|
codingma
|
7497f4f05b
|
rename template to breeze
Former-commit-id: 1223e6358dab52b4e1505057f1b16fd9d527c79e
|
2024-04-07 18:27:20 +08:00 |
|
hoshi-hiyouga
|
47dc54b9ff
|
Merge pull request #3160 from sliderSun/main
support Qwen1.5-32B
Former-commit-id: 1e5a5882dd494c3e9cf5eae2e0a485ce49d1863c
|
2024-04-07 18:00:40 +08:00 |
|
codingma
|
a61e69c0d9
|
rename template to breeze
Former-commit-id: 1d894e7cfb73b8a29dababb554d051bd50e4f01d
|
2024-04-07 11:39:54 +08:00 |
|
codingma
|
34d061f963
|
support https://github.com/hiyouga/LLaMA-Factory/issues/3152
Former-commit-id: 708f0ab4b0aa72e2c73ca36eb9ed058910e43092
|
2024-04-07 11:34:01 +08:00 |
|
sliderSun
|
6c9e8e89cf
|
fix spell error
Former-commit-id: e6d36a2e593ebc1193b1735075c4ddb5d9f54990
|
2024-04-07 10:59:15 +08:00 |
|
sliderSun
|
3bdbe17702
|
support Qwen1.5-32B
Former-commit-id: c419adf1697b92520342f4ffa697c84bf19ca37d
|
2024-04-07 10:56:03 +08:00 |
|
sliderSun
|
6eb424adbe
|
support Qwen1.5-32B
Former-commit-id: 8f2c67b95a8e177eb4096382417a70cacba38e90
|
2024-04-07 10:26:13 +08:00 |
|
hiyouga
|
ff005dc66e
|
update readme
Former-commit-id: 1cf15547e2420a3e5f7a969c21c10c7fbdfc71fe
|
2024-04-07 00:48:24 +08:00 |
|
hiyouga
|
46a76afd44
|
update examples
Former-commit-id: de40ad62ba3d4c74c69de97b39cc79786ac28f0f
|
2024-04-04 14:48:21 +08:00 |
|
hiyouga
|
9e2af508bb
|
tiny fix
Former-commit-id: 70aceecb27e72095c05462d01f956061669b267e
|
2024-04-04 02:19:03 +08:00 |
|
hiyouga
|
1afadd9975
|
back to gradio 4.21 and fix chat
Former-commit-id: 695734a40a702ea059d855da54080cc8d161e41a
|
2024-04-04 02:07:20 +08:00 |
|
hiyouga
|
b112959309
|
fix bug in latest gradio
Former-commit-id: 44a962862b4a74e50ef5786c8d5719faaa65f63f
|
2024-04-04 00:55:31 +08:00 |
|
hiyouga
|
f939dce8d7
|
fix requires for windows
Former-commit-id: 5e25fae40b7ea9cfa72717efbe3677199ca9608f
|
2024-04-03 21:56:43 +08:00 |
|
hiyouga
|
d97150c571
|
fix resize vocab at inference #3022
Former-commit-id: c243720b89eec0af2872fa3c7980a0026d893f4d
|
2024-04-03 18:14:24 +08:00 |
|
hiyouga
|
99ed1657db
|
fix #3116
Former-commit-id: b7256aa33d761280751518c20f29f9b8ea3fb025
|
2024-04-03 14:47:59 +08:00 |
|
hiyouga
|
2e2ebc0986
|
update vllm example
Former-commit-id: 2df6d2eacfa27ebc69455696b93649624c1facbe
|
2024-04-02 22:45:20 +08:00 |
|
hiyouga
|
54773b6116
|
update readme
Former-commit-id: 7ea7333b51be6b1120fc0b13675f5a0ac3c5a12b
|
2024-04-02 22:17:48 +08:00 |
|
hiyouga
|
aef4b0d4a3
|
update examples
Former-commit-id: 2715cfe20f6f4532bebaa47b80ccd5df43d6a490
|
2024-04-02 21:09:25 +08:00 |
|
hiyouga
|
1317dd28a3
|
add zh readme
Former-commit-id: 389a170a4d42c56c71c0e17bbe018c4cb1983b5a
|
2024-04-02 20:58:45 +08:00 |
|
hiyouga
|
037e8518d5
|
update examples
Former-commit-id: c078582a759f6bce6e760cd39a05883f7eb194fe
|
2024-04-02 20:51:21 +08:00 |
|
hiyouga
|
b1401c70be
|
update examples
Former-commit-id: bf36b16e48d6438de6d0b2f2bfe33f7895699b9d
|
2024-04-02 20:41:49 +08:00 |
|
hiyouga
|
06af770c99
|
update readme
Former-commit-id: 9b8e7ccdab167f53fb897e1940562682324e8ff0
|
2024-04-02 20:37:37 +08:00 |
|