hiyouga
|
dff77004f2
|
support olmo
Former-commit-id: 2719510e8c6baa591c74458b773e4e47215e6052
|
2024-03-12 18:30:38 +08:00 |
|
hiyouga
|
6c1b4aec75
|
fix #2802
Former-commit-id: 1370db270d7ba1a20468abdb29193ce7534d1b4f
|
2024-03-12 17:08:34 +08:00 |
|
hiyouga
|
c9ed3fc3a4
|
fix #2782 #2798
Former-commit-id: eb3ab610610a0964bc8a1c9fa015805353f04c31
|
2024-03-12 15:53:29 +08:00 |
|
hiyouga
|
4f9a47c026
|
fix #2775
Former-commit-id: a5c7feb3e8089f4deff760b00a9f84425957c419
|
2024-03-11 00:42:54 +08:00 |
|
hiyouga
|
3fcb1c6d09
|
tiny fix
Former-commit-id: 1d22c87db2449c7d9915842b70fbd59ce9c2dd70
|
2024-03-11 00:17:18 +08:00 |
|
hiyouga
|
7c492864e9
|
update parser
Former-commit-id: d98258aa08d93494ad50d7786064e7fda15f6ca9
|
2024-03-10 13:35:20 +08:00 |
|
hiyouga
|
7ff8a064f3
|
support layerwise galore
Former-commit-id: d43a4da0947897d0be3f62fad3107754d4c89f2b
|
2024-03-10 00:24:11 +08:00 |
|
hiyouga
|
c635bbe465
|
fix #2732
Former-commit-id: bc39ad1d102b91d5417daa38b8a581e1e1ab2af9
|
2024-03-09 22:37:16 +08:00 |
|
hiyouga
|
4881f4e631
|
allow non-packing pretraining
Former-commit-id: 3fee5cc5a3db9ce874ad90f2500ec092d904bd4e
|
2024-03-09 22:21:46 +08:00 |
|
hiyouga
|
c631799f5d
|
fix #2766
Former-commit-id: a8cd556230c1d0bc4e090acc2276c035910ce6f6
|
2024-03-09 21:35:24 +08:00 |
|
hiyouga
|
48846676d8
|
use default arg for freeze tuning
Former-commit-id: a38fd7c8b39cb59fb61c26fdf80aaa6f2d0623b9
|
2024-03-09 06:08:48 +08:00 |
|
hiyouga
|
5d7d8bd55c
|
update hardware requirements
Former-commit-id: 604b3d10fc1448f702943114b66b97bded21e080
|
2024-03-09 03:58:18 +08:00 |
|
hiyouga
|
43b2ede0f8
|
fix #2756 , patch #2746
Former-commit-id: 627d1c91e675f1d9ebf47bad123cbbf29821da4d
|
2024-03-09 02:01:26 +08:00 |
|
hoshi-hiyouga
|
2f095e2017
|
Merge pull request #2746 from stephen-nju/main
fix deepspeed ppo RuntimeError
Former-commit-id: 656c653f0c628f9494b4d7ae12e60c8eeec1ea7a
|
2024-03-09 01:37:00 +08:00 |
|
hiyouga
|
9b97b23ce7
|
fix aqlm version
Former-commit-id: 05673f81f0295c76957f3247c62f95fda322a63e
|
2024-03-09 00:09:09 +08:00 |
|
stephen_zhu
|
940c00e7ae
|
update
Former-commit-id: 295f9ef2eff2e8b5d7a21d3da8dd3e6eb2a42006
|
2024-03-08 12:47:44 +08:00 |
|
stephen
|
18cfd5f349
|
fix ppo runtime error
Former-commit-id: 14e2f221e3e720075e59065a3dc42aa4d993a8b6
|
2024-03-08 11:48:26 +08:00 |
|
hiyouga
|
48d4364586
|
fix chat engine, update webui
Former-commit-id: 8b32dddd7d883bae07735796a517927c79d1c33b
|
2024-03-08 03:01:53 +08:00 |
|
hiyouga
|
3879d79b89
|
update galore args
Former-commit-id: c7479a7976f773feb36aab4fdb0500be53d83b6a
|
2024-03-08 01:17:32 +08:00 |
|
hiyouga
|
e416cecf62
|
fix galore
Former-commit-id: 62a3ceeef8f60caef43ccc7f971a0c9184e21296
|
2024-03-08 00:44:51 +08:00 |
|
hiyouga
|
81fcb80466
|
add Yi-9B model
Former-commit-id: bfcb0245b832242eefb84de6f70bd75544f3ceb7
|
2024-03-07 23:11:57 +08:00 |
|
hiyouga
|
1e6fb6c8aa
|
support galore
Former-commit-id: b67a4a46a88d83bb2a3459b3317b66cda15e0171
|
2024-03-07 22:41:36 +08:00 |
|
hiyouga
|
056d2d956a
|
support vllm
Former-commit-id: 889f6e910e654d8ec3922c2185042d737ffbf1c3
|
2024-03-07 20:26:31 +08:00 |
|
hiyouga
|
9a69cadab3
|
fix #2735
Former-commit-id: 416f6333f66b6afd70a3a936d82593efca583235
|
2024-03-07 16:15:53 +08:00 |
|
hoshi-hiyouga
|
3de642bffd
|
Merge pull request #2730 from cx2333-gt/main
fix flash_attn in train_web
Former-commit-id: eff0b774fc8e1a5a07a2554d611cb85bef439dec
|
2024-03-07 14:37:18 +08:00 |
|
cx2333
|
286b9d9849
|
revert choice name
Former-commit-id: 7832e68072219c7d1f562aee868812a4d655f4e0
|
2024-03-07 14:28:55 +08:00 |
|
hiyouga
|
cef1ede826
|
fix chatglm3 template
Former-commit-id: 9be0aa70fdd2e9ec208aa1850ace5c287efc8c3a
|
2024-03-07 14:26:16 +08:00 |
|
cx2333
|
5007566588
|
fix flash_attn in train_web
Former-commit-id: 5f340e362b0e91fec76c19c77c5705bba1db481a
|
2024-03-07 10:13:55 +08:00 |
|
hiyouga
|
e93fb3cc6c
|
tiny fix
Former-commit-id: c3145afa4164dd28888f17599a154f7dddbe9326
|
2024-03-06 17:25:08 +08:00 |
|
hiyouga
|
7578209735
|
export use balanced gpu
Former-commit-id: 710487dc694489bf3dfe54f8d32df80ce46439e4
|
2024-03-06 16:33:14 +08:00 |
|
hiyouga
|
67f02f75d0
|
fix add tokens
Former-commit-id: ff5353681a87d033903bf8cf6133c6bdb3fa9e5a
|
2024-03-06 15:04:02 +08:00 |
|
hiyouga
|
73d9dfc7ab
|
fix version checking
Former-commit-id: 5780da8d640609cca388f55983d0251e5547209a
|
2024-03-06 14:51:51 +08:00 |
|
hiyouga
|
3168abc0a1
|
fix arg dtype
Former-commit-id: 999ae05655815ac04ababddae55d9343f5d39f84
|
2024-03-05 20:53:30 +08:00 |
|
hiyouga
|
46ee267cfc
|
improve aqlm optim
Former-commit-id: 81be999b407e988c2f42764d827ac859d079ed3e
|
2024-03-05 20:49:50 +08:00 |
|
hiyouga
|
a10bead9b5
|
optimize aqlm training
Former-commit-id: 8b42660e4039b3d6475f502f397686ba6b140627
|
2024-03-05 18:35:41 +08:00 |
|
hiyouga
|
3553e301dd
|
fix dora inference
Former-commit-id: 21b3597b0a05169afe51e1609b532787a65ca8ea
|
2024-03-05 11:51:41 +08:00 |
|
hiyouga
|
02b838b9b0
|
fix export model
Former-commit-id: 7ba2f7bf8da3c559e05d8dde20e93cd1d3d4e8ef
|
2024-03-05 11:05:41 +08:00 |
|
hiyouga
|
0229fffde5
|
auto set chat template
Former-commit-id: d8bf2f0efe6919990c7032aaa06010980cfde019
|
2024-03-05 02:41:20 +08:00 |
|
hiyouga
|
3555b87363
|
update readme
Former-commit-id: c95bc2774800ed2e6d54a6099a466bdacc0cfb78
|
2024-03-04 19:29:26 +08:00 |
|
hiyouga
|
2dca53962e
|
fix export on cpu device
Former-commit-id: e4722a9a627ea4e9a1341cc00a3108dd06a6b550
|
2024-03-04 17:35:09 +08:00 |
|
hiyouga
|
f4f71f2797
|
fix sub-process error in thread
Former-commit-id: 3448ad43d05301b12a19a02c1cc23d7b0ee525c3
|
2024-03-03 15:04:35 +08:00 |
|
hiyouga
|
4fa53b6282
|
update readme, add starcoder2, cosmopedia
Former-commit-id: 1ae7c183640146bb9b06c98942985a1721d2b9c9
|
2024-03-03 01:01:46 +08:00 |
|
hiyouga
|
59a9a5994e
|
fix #2649
Former-commit-id: 1c850de660c671d92f0bc63f230d338b60b7c0bd
|
2024-03-01 13:02:41 +08:00 |
|
hiyouga
|
5306a71b42
|
tiny fix
Former-commit-id: 59116aa07fa5fc608f8b801dd3c89e53b117033e
|
2024-02-29 21:03:48 +08:00 |
|
hiyouga
|
3eafa2dd9e
|
fix webui
Former-commit-id: 730377a818a7ff5e45bf4ac9ee4364c4f123a598
|
2024-02-29 20:09:09 +08:00 |
|
hiyouga
|
88fddb879d
|
fix #2642
Former-commit-id: d8435e7f1850532310e1bee069b45f38cd666e48
|
2024-02-29 18:32:54 +08:00 |
|
hiyouga
|
30855b924a
|
tiny fix
Former-commit-id: 3b6e1132c4d203e6d5376cf97e81cc160697c822
|
2024-02-29 17:28:50 +08:00 |
|
hiyouga
|
48d2e6d7fe
|
tiny fix and release
Former-commit-id: 79ae5f2e06c151cd8f71a96a5ee099f034043ffd
|
2024-02-29 00:46:47 +08:00 |
|
hoshi-hiyouga
|
041c83ea03
|
Merge pull request #2575 from lungothrin/feature/chatter-with-role
support on fly test of tools
Former-commit-id: c49af47d97ef2bae2c57dd03333752321ad6d483
|
2024-02-29 00:39:47 +08:00 |
|
hiyouga
|
0e621c2dc9
|
fix #2629
Former-commit-id: c18822669568327d4fbf480a80c5fe5b8fc95e7a
|
2024-02-29 00:37:29 +08:00 |
|