LLaMA-Factory

mirror of https://github.com/hiyouga/LLaMA-Factory.git synced 2026-06-21 14:48:54 +08:00

Author	SHA1	Message	Date
hoshi-hiyouga	e76eba051d	[data] fix qwen2.5 omni collator (#7553 )	2025-04-01 00:15:12 +08:00
Kingsley	7eed496336	[model] add Qwen2.5-Omni model (#7537 ) * preserve image_sizes * preserve image_sizes * init plugin * support audio-text2text lora * nit * support image/video-text2text, audio-text2text * remove args * remove lines * add docs && nit * remove some comments * fix && add merge part script * add license	2025-03-31 20:39:35 +08:00
Kingsley	8da1d2fa71	[data] fix pixtral plugin (#7505 ) * preserve `image_sizes` * add comments	2025-03-27 17:06:40 +08:00
Xu-pixel	b578a7d5b6	[3rdparty] support swanlab lark notification (#7481 )	2025-03-27 01:52:01 +08:00
Kdump	24afceddb7	[trainer] fix wsd scheduler (#7304 ) * [trainer] Warmup_stable_decay supports setting the number of stable and decay steps according to the warmup_ratio ratio * Update trainer_utils.py --------- Co-authored-by: hoshi-hiyouga <hiyouga@buaa.edu.cn>	2025-03-26 15:25:02 +08:00
hoshi-hiyouga	0583d06676	[model] add qwen2vl 32b & upgrade peft (#7469 ) * add qwen2vl 32b * fix ci * upgrade peft to 0.15 * fix ci * fix ci	2025-03-25 12:15:58 +08:00
GuoCoder	ec6a261568	[model] fix lora on quant models (#7456 ) Co-authored-by: root <root@ai>	2025-03-25 11:59:46 +08:00
Xiaosu Zhu	6b3b97c738	[misc] update liger-kernel's monkey patch (#7453 ) * Update liger_kernel.py * Update setup.py	2025-03-25 11:58:52 +08:00
AbdelKarim ELJANDOUBI	6d3748f727	[misc] enable liger kernel for gemma3 text and paligemma (#7466 ) * add gemma3 text * add paligemma (1,2 and 2 mix)	2025-03-25 09:27:43 +08:00
Kenny Lam	7c890170e3	[misc] enable liger kernel for gemma3 (#7462 )	2025-03-24 19:09:59 +08:00
hoshi-hiyouga	7203365b80	[trainer] fix vlm loss for transformers 4.49 (#7448 )	2025-03-24 10:24:05 +08:00
hoshi-hiyouga	05b19d6952	[deps] upgrade transformers to 4.50.0 (#7437 ) * upgrade transformers * fix hf cache * fix dpo trainer	2025-03-23 17:44:27 +08:00
hoshi-hiyouga	919415dba9	[deps] upgrade vllm to 0.8 (#7436 )	2025-03-23 14:32:22 +08:00
Eric Tang	db0a08db6f	[3rdparty] fix redundant process group destroy for ray (#7395 ) * fix redundant process group destroy for ray * Update tuner.py --------- Co-authored-by: hoshi-hiyouga <hiyouga@buaa.edu.cn>	2025-03-21 10:56:47 +08:00
hoshi-hiyouga	63752fccf7	[assets] update wechat (#7361 )	2025-03-18 21:31:09 +08:00
hoshi-hiyouga	1f9773395b	[misc] set dev version (#7351 )	2025-03-18 00:10:53 +08:00
hoshi-hiyouga	128b5b12b3	[data] fix template (#7349 )	2025-03-17 23:45:20 +08:00
Hertz	ec1154662b	[model] support hunyuan 7b (#7317 ) * [Model]supported tencent-hunyuan model * [Model]supported tencent-hunyuan model(fix) * [Model]supported tencent-hunyuan model(fix)	2025-03-15 20:55:24 +08:00
Qiaolin Yu	a44a53ebec	[inference] support sglang backend (#7278 ) * Mimic SGLang offline Engine * Add more tests and args * Pass all current tests * Clean Code * fix sample_params * clean code * Fix Stream Chat * change sglang from engine mode to server mode * fix * Fix Review Issues * Use SGLang Built-In Utilities * Fix test SGLang * Some Doc Issue * fix sglang engine * add readme --------- Co-authored-by: Jin Pan <jpan236@wisc.edu> Co-authored-by: hiyouga <hiyouga@buaa.edu.cn>	2025-03-15 04:37:58 +08:00
hoshi-hiyouga	93e6184cbe	[data] gemma3 plugin pan and scan (#7294 ) * gemma3 pan and scan * add test case * fix test	2025-03-13 23:29:23 +08:00
Ritesh Goru	480369a9f2	[data] efficient 4d_attention_mask creation in neat_packing (#7272 )	2025-03-13 03:31:12 +08:00
hoshi-hiyouga	650a9a9057	[misc] update format (#7277 )	2025-03-13 02:53:08 +08:00
hoshi-hiyouga	4b9d8da5a4	[model] support gemma3 (#7273 )	2025-03-13 01:35:23 +08:00
hoshi-hiyouga	e6159ad730	[misc] upgrade deps (#7257 )	2025-03-12 00:33:47 +08:00
hoshi-hiyouga	264538cb26	[misc] upgrade format to py39 (#7256 )	2025-03-12 00:08:41 +08:00
hoshi-hiyouga	e2299e261b	Merge pull request #7242 from hiyouga/hiyouga/release [release] release v0.9.2 Former-commit-id: 6b25268990bf225d84e29d4067595cf720fa12d8	2025-03-11 15:28:45 +08:00
hoshi-hiyouga	8a44dce326	Merge pull request #7247 from hiyouga/hiyouga/commit [misc] support print commit info Former-commit-id: 0f7ec4f8529a5d7ea2153b881335821038307bb7	2025-03-11 15:28:04 +08:00
hiyouga	d019603835	support commit info Former-commit-id: a7d89a6dc10579deaf9f45825cc18405a27cade6	2025-03-11 15:13:59 +08:00
hiyouga	478e8194d9	remove exit in preprocess Former-commit-id: f369b6ef41ffd9586ba568b88c5ff32a1af4bace	2025-03-11 15:08:25 +08:00
hiyouga	1890d3dafe	release v0.9.2 Former-commit-id: e7ed1782d4a006400de6fc0f864abd01f7fadeea	2025-03-11 14:49:13 +08:00
hoshi-hiyouga	522a3e8493	[infer] fix vllm args (#7235 ) Former-commit-id: 999be5b4512890b8cf4f45874a77e35cf35626f5	2025-03-11 01:15:35 +08:00
Ze-Yi LIN	18968405d0	[tracking] add swanlab_logdir param (#7219 ) * feat: add swanlab_logdir param * fix Former-commit-id: 9215ad488b6ac6cd57fe8fa4acdacceb63f68ca5	2025-03-11 00:53:07 +08:00
hoshi-hiyouga	71a1c1321a	[config] update args (#7231 ) Former-commit-id: f71a901840811bf560df671ec63a146ff99140c6	2025-03-10 23:04:43 +08:00
hoshi-hiyouga	cf58a6d860	[config] fix export max len (#7230 ) Former-commit-id: 211c0b3e8f3340acd2fae1762d9152a09f19ba34	2025-03-10 16:46:08 +08:00
hoshi-hiyouga	16419b2834	[data] fix loader (#7207 ) * fix dataloader * add test case * fix type * fix ci * fix ci * fix ci * disable overwrite cache in ci Former-commit-id: e84af0e140b1aafd1a6d6fe185a8e41c8fc5f831	2025-03-07 17:20:46 +08:00
ZhangChuanhui	151ef48b40	[data] fix function formatter (#7201 ) Co-authored-by: zhangchuanhui <zhangchal@digitalchina.com> Former-commit-id: 3efb32b986170d2839e526640f85ba230715879a	2025-03-07 15:17:23 +08:00
hoshi-hiyouga	a255c3a476	[misc] fix cli (#7204 ) Former-commit-id: 999f57133ca163c7108d2d5ee8194eca9b2109b4	2025-03-07 15:01:18 +08:00
hoshi-hiyouga	2635794727	[webui] support escape html (#7190 ) Former-commit-id: cf9840374f171359c828b0d6f7a2aa9893c8f701	2025-03-06 16:52:21 +08:00
hoshi-hiyouga	d2f845d70d	[deps] upgrade vllm (#7183 ) Former-commit-id: 37678a3d64668c3b4a4bfefc054e3b9b40427c1a	2025-03-06 15:25:08 +08:00
hoshi-hiyouga	bb8aba5abf	[data] fix mm template (#7181 ) Former-commit-id: 648616d473c81d393592806307e3e25b159cb278	2025-03-06 15:18:32 +08:00
hoshi-hiyouga	9f16c50155	[model] add QwQ 32b (#7179 ) Former-commit-id: 8897e48b8cd55407812453ddd4ff98ac7bdc4e91	2025-03-06 11:58:36 +08:00
Ze-Yi LIN	25bb9f5ad9	[trainer] fix swanlab callback (#7176 ) Former-commit-id: 6d9acf4bd30db24499118aee16bd19cb19ba9e3d	2025-03-06 00:33:37 +08:00
hoshi-hiyouga	7b985f55db	[trainer] update config (#7174 ) Former-commit-id: 9f535d0e3c4ee3cd0f1b65218c2eee5d03f43c6f	2025-03-05 23:32:54 +08:00
sirui.li	fd0357a26d	[data] fix qwen2audio plugin (#7166 ) * Update pairwise.py [data]Repair multimodal model dpo training * Update pairwise.py [data]repair multimodal model dpo training using deepcopy * Update pairwise.py * Update mm_plugin.py Former-commit-id: 86763dfdb8e9e5668c1ddd7e924e4be76bf78368	2025-03-05 18:03:36 +08:00
hoshi-hiyouga	31f9daa362	[data] use bicubic resampler (#7143 ) Former-commit-id: c708f19ab0ab57526134952afddaa90aae8decbf	2025-03-04 00:17:06 +08:00
hoshi-hiyouga	15ea576246	[webui] fix webui (#7142 ) Former-commit-id: d07281f8a45ad8a38d390181d01dcadbcf9aa1b9	2025-03-04 00:01:49 +08:00
rabbit	19a6916d80	[data] bailing template (#7117 ) * add bailing template * add bailing template * add bailing template --------- Co-authored-by: chengshiwen.csw@antgroup.com <chengshiwen.csw@antgroup.com> Former-commit-id: 4a36f5e0abb5a63f4b3b81560bb1ad0e6832d379	2025-03-03 15:33:22 +08:00
hoshi-hiyouga	585c475f71	[inference] fix hf_engine (#7120 ) Former-commit-id: f8cf5319cb5d6e06a1b0d8b8db2b678627f2271e	2025-03-01 05:22:49 +08:00
Ze-Yi LIN	11672f760d	[webui] display swanlab exp link (#7089 ) * webui add swanlab link * change callback name * update --------- Co-authored-by: hiyouga <hiyouga@buaa.edu.cn> Former-commit-id: 27a4b93871c63b839c92940766bd7e0177972c9b	2025-02-27 19:40:54 +08:00
hoshi-hiyouga	5f65558088	[misc] fix project toml (#7067 ) Former-commit-id: 28a668ff4e0beebfe5387362f5518c1d9343666f	2025-02-25 23:22:48 +08:00

1 2 3 4 5 ...

895 Commits