423A35C7
  • Joined on 2024-05-11
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-31 03:00:35 +08:00
4e1d69579a [data] add DLR-Web dataset for supervised fine-tuning (#9696)
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-30 02:30:35 +08:00
1857fbdd6b [ci] add cuda workflow (#9682)
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-29 18:20:35 +08:00
bb1ba31005 [misc] lint mca code (#9692)
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-29 02:00:36 +08:00
e97d0474fb [ci] Fix NPU device condition in docker workflow (#9688)
3f0c3dc84d [assets] fix installation (#9687)
c107cc22d0 [model] support MiniMax-M1&M2 series (#9680)
Compare 3 commits »
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-28 09:40:34 +08:00
7ef1fba34a [version] fix gradio (#9685)
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-27 09:10:35 +08:00
eceec8ab69 [deps] goodbye python 3.9 (#9677)
b44f651e09 [ci] fix docker (#9678)
55590f5ece [misc] fix ci with uv (#9676)
Compare 3 commits »
423A35C7 synced and deleted reference refs/tags/hiyouga/misc at 423A35C7/LLaMA-Factory from mirror 2025-12-27 09:10:35 +08:00
423A35C7 synced commits to hiyouga/misc at 423A35C7/LLaMA-Factory from mirror 2025-12-27 01:00:34 +08:00
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-27 01:00:34 +08:00
a1b1931b4a [breaking] migrate from setuptools to uv (#9673)
3c17f2722c [model] Update ernie_vl to adapt new version (#9665)
a882e2d5fc [assets] Add GitHub Copilot instructions for repository (#9675)
Compare 3 commits »
423A35C7 synced new reference hiyouga/misc to 423A35C7/LLaMA-Factory from mirror 2025-12-27 01:00:34 +08:00
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-25 08:10:36 +08:00
a754604c11 [misc] fix accelerator (#9661)
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-24 07:40:35 +08:00
6a2eafbae3 [feat] Models trained and inferred with Mxfp4 are dequantized by default (#9652)
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-23 23:30:36 +08:00
84485406b7 [ci] disable pip cache for ci (#9654)
1c8a42d2f8 [v1&WIP] dataloader init (#9645)
7901b2f32e [model] efficient tuning for gpt-oss (#9354)
Compare 3 commits »
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-22 06:40:35 +08:00
1f1f5a7d1b [ci] remove docker cache (#9640)
6ef9854713 [misc] fix cache & pin transformers to 4.57.1 (#9638)
Compare 2 commits »
423A35C7 synced and deleted reference refs/tags/hiyouga/cache at 423A35C7/LLaMA-Factory from mirror 2025-12-22 06:40:35 +08:00
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-21 22:30:35 +08:00
4923f52a28 [model] support MiMo-V2-Flash model (#9637)
423A35C7 synced new reference hiyouga/cache to 423A35C7/LLaMA-Factory from mirror 2025-12-21 22:30:35 +08:00
423A35C7 synced commits to hiyouga/cache at 423A35C7/LLaMA-Factory from mirror 2025-12-21 22:30:35 +08:00
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-20 22:00:36 +08:00
0894b4f37e [misc] lint (#9636)
423A35C7 synced commits to main at 423A35C7/LLaMA-Factory from mirror 2025-12-20 05:40:34 +08:00
b0d49e137f [misc] Support split eval_dataset when explict set "predict_with_generate" (#9604)