hoshi-hiyouga
|
3365cc8cf0
|
Merge pull request #3338 from astramind-ai/main
Adding Mixture of Depth
Former-commit-id: 4da2ece53353b63e672ff529d6beba41ff710c14
|
2024-04-21 18:05:52 +08:00 |
|
hoshi-hiyouga
|
3a5e68b7d9
|
fix #3348
Former-commit-id: aa5e921c00f60074eceb2f9d4d8837cc713edba6
|
2024-04-20 10:34:09 +08:00 |
|
hiyouga
|
0cb596fee1
|
add dpo mix dataset
Former-commit-id: 6def3f8bfa51b2d9d73af112352ce07db972e4c9
|
2024-04-20 01:31:38 +08:00 |
|
hiyouga
|
b3b5b530d1
|
fix #3352
Former-commit-id: f315f8e8ec916b82bac94a159e55839ff155c6b5
|
2024-04-19 22:40:01 +08:00 |
|
hiyouga
|
9225c15c88
|
fix llama3 template
Former-commit-id: 20e95250168fbe081c779b2e1ff23f5df3ce02f7
|
2024-04-19 15:46:51 +08:00 |
|
Marco
|
abd9fed445
|
fix small typo
Former-commit-id: 5638a03cd0cf8119ff366b3b3e303b5a2351b065
|
2024-04-18 20:33:29 +02:00 |
|
Marco
|
44cda2eece
|
Added Mixture of Depths
Former-commit-id: 75dd98b9abc847e22cb263c17ebcd2ca5dd98345
|
2024-04-18 20:31:24 +02:00 |
|
hoshi-hiyouga
|
8397808d1d
|
support llama3
Former-commit-id: c1eabb751a5fd73b710714451b146732e0ed4558
|
2024-04-19 01:13:50 +08:00 |
|
hiyouga
|
9e1bd6420d
|
fix #3324
Former-commit-id: 5e710c4ac331f3400534d33b2646c4108c898d98
|
2024-04-18 15:34:45 +08:00 |
|
hiyouga
|
619264c854
|
tiny fix
Former-commit-id: 86399ca8c06273c42c2b184664ae25d3405b3bf6
|
2024-04-18 00:22:17 +08:00 |
|
hiyouga
|
1ebac62e3d
|
update readme
Former-commit-id: a49112a74339ba77bfec53f7870e821fe148db2c
|
2024-04-17 23:40:49 +08:00 |
|
hiyouga
|
ce9bdb3509
|
add mixtral 8x22B models
Former-commit-id: eccbeecff0909e1fa124b5439ffbbfbc5607e1d6
|
2024-04-17 23:35:59 +08:00 |
|
hiyouga
|
0c8d6369ac
|
add CodeQwen models
Former-commit-id: 9f6094241391f8f717818c8ba94e11d1791b4a5c
|
2024-04-17 23:27:22 +08:00 |
|
hiyouga
|
bee796f6b5
|
fix #3316
Former-commit-id: 7395e9e90a209228ff563ab54319955608850fc3
|
2024-04-17 22:54:34 +08:00 |
|
hiyouga
|
9f6349a333
|
fix #3317
Former-commit-id: 7dce1763be4374cf616d96db95ae964ff510a9d6
|
2024-04-17 22:17:19 +08:00 |
|
hiyouga
|
171a029c5e
|
lint
Former-commit-id: 917d65ce65024d17a5030bc57083a427cfae16d7
|
2024-04-16 18:21:09 +08:00 |
|
hoshi-hiyouga
|
eaefaa0fe0
|
Merge pull request #3291 from codemayq/main
support for previewing custom dataset in directory format
Former-commit-id: 40d89152282101a7c08f53e72c2ad7124a0595f3
|
2024-04-16 18:12:09 +08:00 |
|
hiyouga
|
d301f0a64b
|
Update parser.py
Former-commit-id: 92c2133896c20054db86dd53508c982e39bd5ca0
|
2024-04-16 18:09:31 +08:00 |
|
hiyouga
|
0a1578e4e3
|
update readme and gradio version
Former-commit-id: 4029b60ddcbd15b5354503c51178f0f5e7e9aedf
|
2024-04-16 18:09:16 +08:00 |
|
hiyouga
|
a4167fd925
|
support badam for all stages
Former-commit-id: 7a1380646119bfe6855f73dd90570defcea05281
|
2024-04-16 17:44:48 +08:00 |
|
hoshi-hiyouga
|
42084e08ae
|
Merge pull request #3287 from Ledzy/badam
[Feature] Add BAdam algorithm
Former-commit-id: 10a5e1e65b34b03e5ca2a41bf6ded09a3fb25f0c
|
2024-04-16 17:32:16 +08:00 |
|
hoshi-hiyouga
|
9d23f5dc89
|
Update utils.py
Former-commit-id: 01147536b2bb507e87e033fa696e9eb39fe96bbe
|
2024-04-16 17:30:12 +08:00 |
|
hoshi-hiyouga
|
5978427ae0
|
Update trainer.py
Former-commit-id: c6163be1444c00dd000f288e2f834968bd932981
|
2024-04-16 17:29:52 +08:00 |
|
hoshi-hiyouga
|
c7c216069c
|
Update utils.py
Former-commit-id: 7edf4dbed88b8034282f14fd6e0cb6f7f9e5f805
|
2024-04-16 17:29:30 +08:00 |
|
hoshi-hiyouga
|
cde9d1b917
|
Update patcher.py
Former-commit-id: 494e6a1e05b38f5ff61d83327303614f53c92e64
|
2024-04-16 17:29:19 +08:00 |
|
hoshi-hiyouga
|
96213f04b0
|
Update adapter.py
Former-commit-id: 8f7b75b26f020d8ae85baab7b082475c3bfeb512
|
2024-04-16 17:28:12 +08:00 |
|
hoshi-hiyouga
|
7ecea08b9b
|
Update parser.py
Former-commit-id: 898239883afc79f03abd0dc276eef901662a9591
|
2024-04-16 17:27:25 +08:00 |
|
hoshi-hiyouga
|
191971865d
|
Update parser.py
Former-commit-id: 2f3da8169d18b026760cc0ac7dd6141bdd08c932
|
2024-04-16 17:27:02 +08:00 |
|
hoshi-hiyouga
|
ff4f587dd9
|
Update finetuning_args.py
Former-commit-id: 3a23d900aea74078f0bc8cf73fac860a4ce3df67
|
2024-04-16 17:26:30 +08:00 |
|
hoshi-hiyouga
|
de728d0371
|
Update sft.sh
Former-commit-id: 2b4b1562e91bbb02e345e71b7721da9333c0791b
|
2024-04-16 17:25:40 +08:00 |
|
hoshi-hiyouga
|
d08e09642d
|
Update requirements.txt
Former-commit-id: 1e45537ca0bb4d49b4147df01122e365b3d617e4
|
2024-04-16 17:10:17 +08:00 |
|
hoshi-hiyouga
|
351493b183
|
Update setup.py
Former-commit-id: 5df30ea166aff29d48ff83a22ac6ef1611ce3e35
|
2024-04-16 17:10:02 +08:00 |
|
Jonery
|
86ab47e121
|
remove badam from core requirements
Former-commit-id: fa5898944a3867ac5108dd0d579ca0677c87d3d6
|
2024-04-16 12:25:50 +08:00 |
|
Jonery
|
6dd6b3e396
|
resolve gradient checkpointing issue.
Former-commit-id: 6df9135d063bb6102f0cbcdf0d702076f5febbae
|
2024-04-16 12:05:27 +08:00 |
|
codingma
|
5f1418a68b
|
add check
Former-commit-id: 008f6498977c243c80e87242f05c9cf9573541ac
|
2024-04-16 10:56:39 +08:00 |
|
codingma
|
7b97a79efc
|
support for previewing custom dataset in directory format
Former-commit-id: 501cff38c819f06f15194907ce7e052d5f28025a
|
2024-04-16 10:43:14 +08:00 |
|
hiyouga
|
ce4f653121
|
add empty template
Former-commit-id: a325ffa8a668bec354d2636683806acef105e196
|
2024-04-16 03:10:02 +08:00 |
|
hiyouga
|
b053c6454e
|
update readme
Former-commit-id: 8f233745c3aa7a6ef57f275bec80ee731ff76de3
|
2024-04-16 02:36:54 +08:00 |
|
hiyouga
|
ebf0f4a77c
|
update readme
Former-commit-id: f9a246572c1ec0e4b36bff237c6523ce629b7000
|
2024-04-16 02:35:36 +08:00 |
|
hiyouga
|
efa808069a
|
support unsloth 2024.4
Former-commit-id: 14a83f8bc4fe44783252378fce59198194a96bb8
|
2024-04-16 00:25:03 +08:00 |
|
hiyouga
|
b5c5283dd6
|
add codegemma
Former-commit-id: 9324176525c2eda22962b0ca1895009b6237e6e3
|
2024-04-16 00:11:15 +08:00 |
|
hiyouga
|
b638c65519
|
support cohere commandR #3184
Former-commit-id: e077c36872740f6b2ac255aee9da6c4c70f28977
|
2024-04-15 23:26:42 +08:00 |
|
Jonery
|
d4d471450f
|
Feature BAdam
Former-commit-id: d8d2807fbcf587c37f7fd34a23e9397d2775ceed
|
2024-04-15 23:15:27 +08:00 |
|
hoshi-hiyouga
|
3144bdec2c
|
Merge pull request #3254 from marko1616/feature/Add-support-for-CohereForAI/c4ai-command-r-plus
Add template&support for c4ai-command-r/plus (tested)
Former-commit-id: 41d39ec4889abad050820bf153133ac3a11228a3
|
2024-04-15 22:59:35 +08:00 |
|
hoshi-hiyouga
|
c6d6c4c209
|
Update template.py
Former-commit-id: 00b8be7dafa65e13b344724a8d3855919ee4f631
|
2024-04-15 22:58:01 +08:00 |
|
hoshi-hiyouga
|
f5f1589662
|
Update constants.py
Former-commit-id: 39199f712aa7b7a1c66080d9c84651fd2eb0b425
|
2024-04-15 22:56:55 +08:00 |
|
hiyouga
|
276f2cb24e
|
update examples
Former-commit-id: 369294b31c8a03a1cafcee83eb31a817007d3c49
|
2024-04-15 22:14:34 +08:00 |
|
marko1616
|
952b785bb3
|
change default_system accroding to official template
Former-commit-id: 7ad9029c5e77a87a7c324b8f90b4f80a31a5c78b
|
2024-04-15 20:45:46 +08:00 |
|
marko1616
|
72dd676208
|
Revert "Add support for function call(Not strictly following origin)"
This reverts commit dfaa31e99173cc96be65a78c5931a16e002dd929 [formerly 44f3ada4e394c06b0d972329ed2a62d2be2ea0c6].
Former-commit-id: fac9cc6e01dd8f3bc449b656804476e1871326f0
|
2024-04-15 20:27:09 +08:00 |
|
marko1616
|
dfaa31e991
|
Add support for function call(Not strictly following origin)
Former-commit-id: 44f3ada4e394c06b0d972329ed2a62d2be2ea0c6
|
2024-04-15 20:16:52 +08:00 |
|