LLaMA-Factory

mirror of https://github.com/hiyouga/LLaMA-Factory.git synced 2025-10-15 16:18:10 +08:00

Author	SHA1	Message	Date
Yaowei Zheng	7c223c432b	[model] add qwen3 2507 model (#8783 )	2025-07-30 17:19:19 +08:00
Kingsley	52882d01c3	[model] support keye-vl-8b (#8776 )	2025-07-29 21:24:08 +08:00
Kingsley	4e0bf35eb4	[model] update glm4.5 (#8770 )	2025-07-29 19:57:29 +08:00
Yaowei Zheng	8efa506c16	[model] add qwen3 2507 models (#8750 )	2025-07-25 20:21:47 +08:00
Steven sun	9d6565d1a8	[model] support granite4 (#8680 )	2025-07-21 14:15:36 +08:00
Kingsley	9c9b307d33	[model] add Devstral-Small-2507 (#8614 )	2025-07-11 18:59:53 +08:00
Yaowei Zheng	906b31fd47	[assets] update readme (#8529 )	2025-07-02 17:42:27 +08:00
Kingsley	bede213da7	[assets] update readme (#8519 )	2025-07-02 15:38:38 +08:00
Kingsley	e9f70daabe	[model] add gemma3n (#8509 )	2025-07-01 22:37:24 +08:00
Kingsley	d17a672251	[model] add GLM-4.1V (#8462 )	2025-06-30 01:09:41 +08:00
Liu Jiajun	4f0da0aec9	[data] fix gemma2 eos token (#8480 ) Co-authored-by: Yaowei Zheng <hiyouga@buaa.edu.cn>	2025-06-27 18:19:15 +08:00
Kingsley	ecbccb4c5d	[model] Add mistral-small 3.2 & kimi-dev (#8433 )	2025-06-24 14:59:47 +08:00
Yaowei Zheng	9af7915f7b	[model] add kimi vl 2506 (#8432 )	2025-06-23 17:56:48 +08:00
Dhia Eddine Rhaiem	88a92be808	[model] add support for Falcon H1 (#8403 )	2025-06-18 16:51:23 +08:00
Yaowei Zheng	3a3bae1cfe	[data] fix qwen2vl pos ids (#8387 )	2025-06-17 00:48:54 +08:00
Yaowei Zheng	44f1b9b5ad	[misc] tiny fixes (#8348 )	2025-06-10 15:30:58 +08:00
阿丹(adan)	b41697c9b6	[model] support MiniCPM4 (#8314 )	2025-06-10 14:38:39 +08:00
Kingsley	31bca4d172	[model] support Mistral3.1 small 2503 (#8335 )	2025-06-09 10:37:42 +08:00
Yaowei Zheng	c0710be6d7	[assets] update readme (#8303 )	2025-06-05 23:23:15 +08:00
Kingsley	554e89ff02	[model] add MIMO_VL (#8249 )	2025-06-01 03:54:54 +08:00
Akshat Sehgal	c7e63bead7	[model] add smollm2 support (#8220 )	2025-05-31 16:29:01 +08:00
hoshi-hiyouga	42bebc341d	[model] add deepseek 0528 models (#8215 )	2025-05-29 21:37:07 +08:00
hoshi-hiyouga	3c7dc66a92	[model] add smollm2 and medgemma (#8161 )	2025-05-26 23:19:58 +08:00
Akshat Sehgal	501e7d8a8f	feat: add smollm support (#8050 )	2025-05-26 19:47:54 +08:00
hoshi-hiyouga	9ae17cd173	[deps] update to transformers 4.52 (#8125 )	2025-05-21 05:16:18 +08:00
hoshi-hiyouga	dc080399c6	[model] add seed coder and qwen3 quant models (#8039 )	2025-05-13 15:59:55 +08:00
Kingsley	ef86a53063	[model] add mimo7b (#7946 )	2025-05-06 17:10:30 +02:00
hoshi-hiyouga	ce7032e1b3	[model] add qwen2 omni 3b (#7945 )	2025-05-03 16:36:51 +08:00
hoshi-hiyouga	052ca871bd	[data] optimize qwen3 loss computation (#7923 )	2025-04-30 16:18:00 +08:00
hoshi-hiyouga	98f23c6584	[model] add qwen3 (#7885 )	2025-04-29 09:34:05 +08:00
Kingsley	7500e761d3	[misc] update internvl constants (#7801 )	2025-04-22 15:53:08 +08:00
hoshi-hiyouga	86ebb219d6	[breaking] bump transformers to 4.45.0 & improve ci (#7746 ) * update ci * fix * fix * fix * fix * fix	2025-04-17 02:36:48 +08:00
Kingsley	2e518f255f	[model] support intern-VL 2.5-3 series (#7258 ) * add internvl and rebase * fix for internvl2&3 * remove lines * fix video_inputs & lint * nit * add constants * remove lines * fix * fix error * pass ci * pass ci * skip internvl & nit	2025-04-17 00:31:30 +08:00
hoshi-hiyouga	1134baeedd	[assets] update model readme (#7724 )	2025-04-15 00:41:09 +08:00
Kingsley	2101399c94	[model] Support Kimi_VL thinking/instruct (#7719 ) * add kimi_vl * patch config * check version * Update mm_plugin.py * Update mm_plugin.py --------- Co-authored-by: hoshi-hiyouga <hiyouga@buaa.edu.cn>	2025-04-15 00:21:58 +08:00
hoshi-hiyouga	7c61b35106	[misc] upgrade cli (#7714 )	2025-04-14 15:41:22 +08:00
hoshi-hiyouga	f518bfba5b	[deps] upgrade transformers (#7704 )	2025-04-13 18:11:34 +08:00
Yuxuan Zhang	8162f94db5	[model] add GLM-4-0414 (#7695 ) * Update README_zh.md * update	2025-04-13 17:10:45 +08:00
hoshi-hiyouga	c3c0efbaa0	[misc] fix packing and eval plot (#7623 )	2025-04-07 18:20:57 +08:00
hoshi-hiyouga	831e7f1cfd	[model] add llama4 (#7611 )	2025-04-06 13:42:31 +08:00
Kingsley	7eed496336	[model] add Qwen2.5-Omni model (#7537 ) * preserve image_sizes * preserve image_sizes * init plugin * support audio-text2text lora * nit * support image/video-text2text, audio-text2text * remove args * remove lines * add docs && nit * remove some comments * fix && add merge part script * add license	2025-03-31 20:39:35 +08:00
hoshi-hiyouga	0583d06676	[model] add qwen2vl 32b & upgrade peft (#7469 ) * add qwen2vl 32b * fix ci * upgrade peft to 0.15 * fix ci * fix ci	2025-03-25 12:15:58 +08:00
Hertz	ec1154662b	[model] support hunyuan 7b (#7317 ) * [Model]supported tencent-hunyuan model * [Model]supported tencent-hunyuan model(fix) * [Model]supported tencent-hunyuan model(fix)	2025-03-15 20:55:24 +08:00
Qiaolin Yu	a44a53ebec	[inference] support sglang backend (#7278 ) * Mimic SGLang offline Engine * Add more tests and args * Pass all current tests * Clean Code * fix sample_params * clean code * Fix Stream Chat * change sglang from engine mode to server mode * fix * Fix Review Issues * Use SGLang Built-In Utilities * Fix test SGLang * Some Doc Issue * fix sglang engine * add readme --------- Co-authored-by: Jin Pan <jpan236@wisc.edu> Co-authored-by: hiyouga <hiyouga@buaa.edu.cn>	2025-03-15 04:37:58 +08:00
hoshi-hiyouga	4b9d8da5a4	[model] support gemma3 (#7273 )	2025-03-13 01:35:23 +08:00
hoshi-hiyouga	264538cb26	[misc] upgrade format to py39 (#7256 )	2025-03-12 00:08:41 +08:00
hoshi-hiyouga	71a1c1321a	[config] update args (#7231 ) Former-commit-id: f71a901840811bf560df671ec63a146ff99140c6	2025-03-10 23:04:43 +08:00
hoshi-hiyouga	9f16c50155	[model] add QwQ 32b (#7179 ) Former-commit-id: 8897e48b8cd55407812453ddd4ff98ac7bdc4e91	2025-03-06 11:58:36 +08:00
Ze-Yi LIN	11672f760d	[webui] display swanlab exp link (#7089 ) * webui add swanlab link * change callback name * update --------- Co-authored-by: hiyouga <hiyouga@buaa.edu.cn> Former-commit-id: 27a4b93871c63b839c92940766bd7e0177972c9b	2025-02-27 19:40:54 +08:00
Kingsley	2986bef530	[model] add paligemma2-mix series (#7060 ) Former-commit-id: 0c0196306d343242ee5e6f22c55562f9a74aa782	2025-02-25 18:51:16 +08:00

1 2 3 4

167 Commits