LLaMA-Factory

mirror of https://github.com/hiyouga/LLaMA-Factory.git synced 2025-12-16 20:00:36 +08:00

Author	SHA1	Message	Date
hiyouga	21db8ed2f4	use pre-commit	2024-10-29 09:07:46 +00:00
hiyouga	77666bd227	update requires	2024-10-29 16:10:07 +08:00
hiyouga	451d271718	tiny fix	2024-10-08 17:48:56 +08:00
hiyouga	fe7ffccdb9	fix #5542	2024-09-30 23:28:55 +08:00
BUAADreamer	485fc04716	fix constants	2024-09-29 22:00:01 +08:00
BUAADreamer	23916d57c1	fix style	2024-09-29 21:39:37 +08:00
Zhangchi Feng	83abf86657	Merge branch 'main' into main	2024-09-29 21:32:54 +08:00
BUAADreamer	65a8923f5a	add more llava-next series template	2024-09-29 21:29:29 +08:00
BUAADreamer	bf0bcbc5ec	fix readme	2024-09-29 20:45:02 +08:00
BUAADreamer	d5c69400cd	fix some	2024-09-29 17:55:40 +08:00
hoshi-hiyouga	fe7057a8a3	Update attention.py	2024-09-29 10:47:41 +08:00
Amirreza A	94ee105526	made a small change to a warning about fa2 for gemma2 models.	2024-09-28 19:03:36 +03:30
BUAADreamer	7f3f81009e	fix some	2024-09-28 01:15:33 +08:00
BUAADreamer	6642cd501d	add llava-next/llava-next-video/video-llava	2024-09-28 00:57:03 +08:00
Zhangchi Feng	900631755b	Merge branch 'hiyouga:main' into main	2024-09-27 18:14:39 +08:00
Billy Cao	7a2958a44f	Add qwen_vl to liger kernel supported list	2024-09-14 19:28:20 +08:00
BUAADreamer	7b4ba0efb6	try to past test	2024-09-10 13:12:51 +08:00
BUAADreamer	31259e7e0c	support llava-next(video)	2024-09-10 12:31:53 +08:00
hiyouga	c9b3870adb	tiny fix	2024-09-08 23:18:08 +08:00
hiyouga	b332908ab4	fix test case	2024-09-08 01:50:51 +08:00
hiyouga	52a06efaf8	add test case	2024-09-08 01:40:49 +08:00
hiyouga	fb72a3adb0	support activation offloading via unsloth gc	2024-09-08 01:22:19 +08:00
hiyouga	76f2e59504	tiny fix	2024-09-05 23:41:16 +08:00
hiyouga	a61c8c4890	fix #5324	2024-09-02 23:56:21 +08:00
hiyouga	9967ccb3ae	fix mixed mm inputs and rlhf-v	2024-09-01 20:52:47 +08:00
hiyouga	a025c3df61	remove visual_inputs, fix qlora	2024-08-31 00:24:51 +08:00
hiyouga	3382317e32	refactor mm training	2024-08-30 02:14:31 +08:00
hoshi-hiyouga	727e184840	Merge pull request #5290 from simonJJJ/qwen2_vl support qwen2-vl	2024-08-30 02:10:36 +08:00
hiyouga	a7dd7d325e	update liger kernel	2024-08-29 20:46:08 +08:00
simonJJJ	aeb85f200b	initial-commit	2024-08-28 16:51:35 +08:00
hiyouga	72bc8f0111	support liger kernel	2024-08-27 11:20:14 +08:00
hiyouga	b7ca6c8dc1	fix #5048	2024-08-05 23:48:19 +08:00
hiyouga	29ebcd75d5	fix up	2024-07-15 01:04:56 +08:00
hiyouga	d3c01552e0	tiny fix	2024-07-14 10:56:45 +08:00
hiyouga	2f6af73da2	fix gemma2 attention	2024-07-13 23:33:45 +08:00
hiyouga	0c699de39d	tiny fix	2024-07-04 03:47:05 +08:00
hiyouga	6fd6aa4530	fix packing for eager/sdpa attn	2024-07-04 01:52:43 +08:00
hoshi-hiyouga	87d9b2d005	Merge pull request #4224 from chuan298/main Implement efficient packing without cross-contamination attention	2024-07-04 01:18:54 +08:00
hiyouga	cce7083024	update packing	2024-07-04 01:10:55 +08:00
hoshi-hiyouga	a36e8f2dd5	Update packing.py	2024-07-03 23:36:01 +08:00
hiyouga	c346f79f99	update func name	2024-07-03 23:29:33 +08:00
hiyouga	8a6a7b9c8a	update arg name	2024-07-03 23:23:24 +08:00
hiyouga	8b1172b910	tiny fix	2024-07-03 02:31:50 +08:00
hoshi-hiyouga	e8e6af2651	Merge branch 'main' into main	2024-07-01 21:01:09 +08:00
hiyouga	8c41a0aa6d	tiny fix	2024-07-01 03:55:20 +08:00
hiyouga	2f4b89ace1	loose gemma2 attention	2024-06-29 01:42:14 +08:00
hiyouga	4d35e218b1	bf16 by default, gemma2 attns Gemma2 finetuning cannot work until merging https://github.com/huggingface/transformers/pull/31674	2024-06-28 06:00:26 +08:00
hiyouga	96a5044394	add quant checks	2024-06-27 01:12:25 +08:00
hiyouga	ad144c2265	support HQQ/EETQ #4113	2024-06-27 00:29:42 +08:00
hiyouga	addca926de	improve autogptq integration	2024-06-26 22:11:44 +08:00

1 2

65 Commits