hiyouga
|
fe17ad1ecd
|
update readme
Former-commit-id: 3a8c17907c71f46b1b37501e2afdc99ad89fb4bc
|
2024-04-22 00:21:01 +08:00 |
|
hiyouga
|
366c0eb1c5
|
fix mod stuff
Former-commit-id: cf3988226e6398c67bb2955578e436fc505aa5c5
|
2024-04-21 18:11:10 +08:00 |
|
Marco
|
68dbd5d220
|
Added Mixture of Depths
Former-commit-id: 75dd98b9abc847e22cb263c17ebcd2ca5dd98345
|
2024-04-18 20:31:24 +02:00 |
|
hoshi-hiyouga
|
39f0cc7d8b
|
support llama3
Former-commit-id: c1eabb751a5fd73b710714451b146732e0ed4558
|
2024-04-19 01:13:50 +08:00 |
|
hiyouga
|
dcc34ab729
|
tiny fix
Former-commit-id: 86399ca8c06273c42c2b184664ae25d3405b3bf6
|
2024-04-18 00:22:17 +08:00 |
|
hiyouga
|
953bb50ab1
|
update readme
Former-commit-id: a49112a74339ba77bfec53f7870e821fe148db2c
|
2024-04-17 23:40:49 +08:00 |
|
hiyouga
|
ca2c480736
|
add mixtral 8x22B models
Former-commit-id: eccbeecff0909e1fa124b5439ffbbfbc5607e1d6
|
2024-04-17 23:35:59 +08:00 |
|
hiyouga
|
16e20ffa8f
|
update readme and gradio version
Former-commit-id: 4029b60ddcbd15b5354503c51178f0f5e7e9aedf
|
2024-04-16 18:09:16 +08:00 |
|
hiyouga
|
d41793228e
|
support badam for all stages
Former-commit-id: 7a1380646119bfe6855f73dd90570defcea05281
|
2024-04-16 17:44:48 +08:00 |
|
hiyouga
|
1d4c0491dc
|
update readme
Former-commit-id: 8f233745c3aa7a6ef57f275bec80ee731ff76de3
|
2024-04-16 02:36:54 +08:00 |
|
hiyouga
|
480227e04e
|
update readme
Former-commit-id: f9a246572c1ec0e4b36bff237c6523ce629b7000
|
2024-04-16 02:35:36 +08:00 |
|
hiyouga
|
2aa1d1476e
|
add codegemma
Former-commit-id: 9324176525c2eda22962b0ca1895009b6237e6e3
|
2024-04-16 00:11:15 +08:00 |
|
hiyouga
|
19874e39ee
|
support cohere commandR #3184
Former-commit-id: e077c36872740f6b2ac255aee9da6c4c70f28977
|
2024-04-15 23:26:42 +08:00 |
|
hiyouga
|
a97f8d1fa8
|
release v0.6.2
Former-commit-id: f92ad0a62d957b595f6a76a5403216b163eb3d17
|
2024-04-11 20:08:51 +08:00 |
|
hiyouga
|
ff005dc66e
|
update readme
Former-commit-id: 1cf15547e2420a3e5f7a969c21c10c7fbdfc71fe
|
2024-04-07 00:48:24 +08:00 |
|
hiyouga
|
f939dce8d7
|
fix requires for windows
Former-commit-id: 5e25fae40b7ea9cfa72717efbe3677199ca9608f
|
2024-04-03 21:56:43 +08:00 |
|
hiyouga
|
2e2ebc0986
|
update vllm example
Former-commit-id: 2df6d2eacfa27ebc69455696b93649624c1facbe
|
2024-04-02 22:45:20 +08:00 |
|
hiyouga
|
54773b6116
|
update readme
Former-commit-id: 7ea7333b51be6b1120fc0b13675f5a0ac3c5a12b
|
2024-04-02 22:17:48 +08:00 |
|
hiyouga
|
1317dd28a3
|
add zh readme
Former-commit-id: 389a170a4d42c56c71c0e17bbe018c4cb1983b5a
|
2024-04-02 20:58:45 +08:00 |
|
hiyouga
|
06af770c99
|
update readme
Former-commit-id: 9b8e7ccdab167f53fb897e1940562682324e8ff0
|
2024-04-02 20:37:37 +08:00 |
|
hiyouga
|
c95e06b34a
|
update readme
Former-commit-id: 0c73d3c8a5762a8f119b27322ffd52a61de6fe38
|
2024-04-02 20:22:11 +08:00 |
|
hiyouga
|
75819c1220
|
simplify readme
Former-commit-id: 0da6ec2d516326fe9c7583ba71cd1778eb838178
|
2024-04-02 20:07:43 +08:00 |
|
hiyouga
|
ff301c08d6
|
add qwen1.5 moe
Former-commit-id: 3ea94f0d12cec25ac694a2c4ae8971c356990b61
|
2024-04-01 21:49:40 +08:00 |
|
hiyouga
|
8365522ce2
|
fix #3077
Former-commit-id: d0340391e8075cff0d84b3ef879c2101b66ca1dc
|
2024-04-01 21:35:18 +08:00 |
|
hiyouga
|
cd05e1bb8b
|
update readme
Former-commit-id: 297b01f16ac78cde15a5d85a9a5b82ea20bfaf23
|
2024-03-31 18:46:34 +08:00 |
|
hiyouga
|
e6c7e6e667
|
support ORPO
Former-commit-id: f44a4c27e2461cdaa1b16865f597a31033c0e6d9
|
2024-03-31 18:29:50 +08:00 |
|
hiyouga
|
293e372ad1
|
update readme
Former-commit-id: 312d4f90784800dc8db4eaa7d908e6761115bc51
|
2024-03-28 22:02:32 +08:00 |
|
hiyouga
|
c92d02af7a
|
add project
Former-commit-id: 0418e9fecb2337b5d1b72e8358adb8aa10803c4b
|
2024-03-28 20:24:27 +08:00 |
|
hiyouga
|
af6bd1b0c8
|
update readme
Former-commit-id: 6b634b5c2dbad827e8cc9850b8d7697c2056532a
|
2024-03-28 18:35:11 +08:00 |
|
hiyouga
|
27f5c967e4
|
update trainers
Former-commit-id: d0dd6eefed0b86895ed00a7cafb331e5193db645
|
2024-03-28 18:16:27 +08:00 |
|
hiyouga
|
dc8214c01b
|
update readme
Former-commit-id: 32e6a7f10fdc28106e3b086eb79304943c6e8fab
|
2024-03-25 23:06:13 +08:00 |
|
hoshi-hiyouga
|
8b21475453
|
Merge pull request #2967 from Tsumugii24/main
Update README_zh.md
Former-commit-id: 4c3b8da2caf74e9d6819bdb1a4e30ca3c549a2d8
|
2024-03-25 23:02:22 +08:00 |
|
Tsumugii24
|
7c2e29ec9e
|
Update README_zh.md
Former-commit-id: 34141ee0515c3e765ca0cb82a0625fb0abfba6f9
|
2024-03-25 22:54:26 +08:00 |
|
hiyouga
|
80861ca109
|
release v0.6.0
Former-commit-id: 51910d5803eb718e4976da0b3bfcdc5eeeea48eb
|
2024-03-25 22:38:56 +08:00 |
|
Tsumugii24
|
22c44611fe
|
Update README_zh.md
Former-commit-id: deec57ec009ef6c08a90ad8e5800d6d5a936b337
|
2024-03-25 22:31:03 +08:00 |
|
hiyouga
|
06019b7ee3
|
fix #2941
Former-commit-id: 3775ab52017f0b610ddd8199cccfb8c001eda507
|
2024-03-24 00:28:44 +08:00 |
|
0xez
|
5f97b01820
|
Update README_zh.md, fix the release date of the paper
Former-commit-id: 6ea16156b6456216cefab59265dae1edc9dc938f
|
2024-03-22 10:41:17 +08:00 |
|
hiyouga
|
ec6d4ac99a
|
add citation
Former-commit-id: 54199205f2000c0500d29822387646133e06e8b2
|
2024-03-21 17:04:10 +08:00 |
|
hiyouga
|
8d77bc35ac
|
paper release
Former-commit-id: 7bd384655244ce6a8c1f34aa6fed54122d0e9da5
|
2024-03-21 13:49:17 +08:00 |
|
hiyouga
|
85b1fe92ae
|
update readme
Former-commit-id: ab98d4d617b7193c474f58a29ca9475fea7564aa
|
2024-03-21 00:48:42 +08:00 |
|
hiyouga
|
b590e82d41
|
support fsdp + qlora
Former-commit-id: b894bf8e84be689db258021f0638e9ac939abcbc
|
2024-03-21 00:36:06 +08:00 |
|
hiyouga
|
2d12e88c23
|
fix #2777 #2895
Former-commit-id: 54d5f62d29456a8d9d0c0dd3d0bbfffe48935803
|
2024-03-20 17:59:45 +08:00 |
|
khazic
|
9e99af39c2
|
Updated README with new information
Former-commit-id: 90a81c2e52bd44beb3b7feb5d2517b073f7f6ef9
|
2024-03-20 14:21:16 +08:00 |
|
刘一博
|
3b80637eb5
|
Updated README with new information
Former-commit-id: fddbc29ca1bd9b13372087e6a349f21240abc013
|
2024-03-20 14:11:28 +08:00 |
|
hiyouga
|
4ef67ed4dd
|
improve lora+ impl.
Former-commit-id: 332bad25455a70ad9204e7dd384bb086d789aa39
|
2024-03-13 23:32:51 +08:00 |
|
hiyouga
|
28acb02e80
|
support olmo
Former-commit-id: 2719510e8c6baa591c74458b773e4e47215e6052
|
2024-03-12 18:30:38 +08:00 |
|
hiyouga
|
0aacc41252
|
support layerwise galore
Former-commit-id: d43a4da0947897d0be3f62fad3107754d4c89f2b
|
2024-03-10 00:24:11 +08:00 |
|
hiyouga
|
5944e04a21
|
add GaLore results
Former-commit-id: ac05b9bba62924693bdede85917d21b844849b8c
|
2024-03-09 04:11:55 +08:00 |
|
hiyouga
|
72608bbcb3
|
update hardware requirements
Former-commit-id: 604b3d10fc1448f702943114b66b97bded21e080
|
2024-03-09 03:58:18 +08:00 |
|
hiyouga
|
1dd3f17f79
|
fix aqlm version
Former-commit-id: 05673f81f0295c76957f3247c62f95fda322a63e
|
2024-03-09 00:09:09 +08:00 |
|