hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							3c7dc66a92
							
						
					 | 
					
						
						
							
							[model] add smollm2 and medgemma (#8161)
						
						
						
						
						
						
					 | 
					
						2025-05-26 23:19:58 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Akshat Sehgal
							
						 
					 | 
					
						
						
						
						
							
						
						
							501e7d8a8f
							
						
					 | 
					
						
						
							
							feat: add smollm support (#8050)
						
						
						
						
						
						
					 | 
					
						2025-05-26 19:47:54 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							9ae17cd173
							
						
					 | 
					
						
						
							
							[deps] update to transformers 4.52 (#8125)
						
						
						
						
						
						
					 | 
					
						2025-05-21 05:16:18 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							dc080399c6
							
						
					 | 
					
						
						
							
							[model] add seed coder and qwen3 quant models (#8039)
						
						
						
						
						
						
					 | 
					
						2025-05-13 15:59:55 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Kingsley
							
						 
					 | 
					
						
						
						
						
							
						
						
							ef86a53063
							
						
					 | 
					
						
						
							
							[model] add mimo7b (#7946)
						
						
						
						
						
						
					 | 
					
						2025-05-06 17:10:30 +02:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							ce7032e1b3
							
						
					 | 
					
						
						
							
							[model] add qwen2 omni 3b (#7945)
						
						
						
						
						
						
					 | 
					
						2025-05-03 16:36:51 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							052ca871bd
							
						
					 | 
					
						
						
							
							[data] optimize qwen3 loss computation (#7923)
						
						
						
						
						
						
					 | 
					
						2025-04-30 16:18:00 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							98f23c6584
							
						
					 | 
					
						
						
							
							[model] add qwen3 (#7885)
						
						
						
						
						
						
					 | 
					
						2025-04-29 09:34:05 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Kingsley
							
						 
					 | 
					
						
						
						
						
							
						
						
							7500e761d3
							
						
					 | 
					
						
						
							
							[misc] update internvl constants (#7801)
						
						
						
						
						
						
					 | 
					
						2025-04-22 15:53:08 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							86ebb219d6
							
						
					 | 
					
						
						
							
							[breaking] bump transformers to 4.45.0 & improve ci (#7746)
						
						
						
						
						
						
						
						* update ci
* fix
* fix
* fix
* fix
* fix 
						
						
					 | 
					
						2025-04-17 02:36:48 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Kingsley
							
						 
					 | 
					
						
						
						
						
							
						
						
							2e518f255f
							
						
					 | 
					
						
						
							
							[model] support intern-VL 2.5-3 series (#7258)
						
						
						
						
						
						
						
						* add internvl and rebase
* fix for internvl2&3
* remove lines
* fix video_inputs & lint
* nit
* add constants
* remove lines
* fix
* fix error
* pass ci
* pass ci
* skip internvl & nit 
						
						
					 | 
					
						2025-04-17 00:31:30 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							1134baeedd
							
						
					 | 
					
						
						
							
							[assets] update model readme (#7724)
						
						
						
						
						
						
					 | 
					
						2025-04-15 00:41:09 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Kingsley
							
						 
					 | 
					
						
						
						
						
							
						
						
							2101399c94
							
						
					 | 
					
						
						
							
							[model] Support Kimi_VL thinking/instruct (#7719)
						
						
						
						
						
						
						
						* add kimi_vl
* patch config
* check version
* Update mm_plugin.py
* Update mm_plugin.py
---------
Co-authored-by: hoshi-hiyouga <hiyouga@buaa.edu.cn> 
						
						
					 | 
					
						2025-04-15 00:21:58 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							7c61b35106
							
						
					 | 
					
						
						
							
							[misc] upgrade cli (#7714)
						
						
						
						
						
						
					 | 
					
						2025-04-14 15:41:22 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							f518bfba5b
							
						
					 | 
					
						
						
							
							[deps] upgrade transformers (#7704)
						
						
						
						
						
						
					 | 
					
						2025-04-13 18:11:34 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Yuxuan Zhang
							
						 
					 | 
					
						
						
						
						
							
						
						
							8162f94db5
							
						
					 | 
					
						
						
							
							[model] add GLM-4-0414 (#7695)
						
						
						
						
						
						
						
						* Update README_zh.md
* update 
						
						
					 | 
					
						2025-04-13 17:10:45 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							c3c0efbaa0
							
						
					 | 
					
						
						
							
							[misc] fix packing and eval plot (#7623)
						
						
						
						
						
						
					 | 
					
						2025-04-07 18:20:57 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							831e7f1cfd
							
						
					 | 
					
						
						
							
							[model] add llama4 (#7611)
						
						
						
						
						
						
					 | 
					
						2025-04-06 13:42:31 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Kingsley
							
						 
					 | 
					
						
						
						
						
							
						
						
							7eed496336
							
						
					 | 
					
						
						
							
							[model] add Qwen2.5-Omni model (#7537)
						
						
						
						
						
						
						
						* preserve image_sizes
* preserve image_sizes
* init plugin
* support audio-text2text lora
* nit
* support image/video-text2text, audio-text2text
* remove args
* remove lines
* add docs && nit
* remove some comments
* fix && add merge part script
* add license 
						
						
					 | 
					
						2025-03-31 20:39:35 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							0583d06676
							
						
					 | 
					
						
						
							
							[model] add qwen2vl 32b & upgrade peft (#7469)
						
						
						
						
						
						
						
						* add qwen2vl 32b
* fix ci
* upgrade peft to 0.15
* fix ci
* fix ci 
						
						
					 | 
					
						2025-03-25 12:15:58 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Hertz
							
						 
					 | 
					
						
						
						
						
							
						
						
							ec1154662b
							
						
					 | 
					
						
						
							
							[model] support hunyuan 7b (#7317)
						
						
						
						
						
						
						
						* [Model]supported tencent-hunyuan model
* [Model]supported tencent-hunyuan model(fix)
* [Model]supported tencent-hunyuan model(fix) 
						
						
					 | 
					
						2025-03-15 20:55:24 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Qiaolin Yu
							
						 
					 | 
					
						
						
						
						
							
						
						
							a44a53ebec
							
						
					 | 
					
						
						
							
							[inference] support sglang backend (#7278)
						
						
						
						
						
						
						
						* Mimic SGLang offline Engine
* Add more tests and args
* Pass all current tests
* Clean Code
* fix sample_params
* clean code
* Fix Stream Chat
* change sglang from engine mode to server mode
* fix
* Fix Review Issues
* Use SGLang Built-In Utilities
* Fix test SGLang
* Some Doc Issue
* fix sglang engine
* add readme
---------
Co-authored-by: Jin Pan <jpan236@wisc.edu>
Co-authored-by: hiyouga <hiyouga@buaa.edu.cn> 
						
						
					 | 
					
						2025-03-15 04:37:58 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							4b9d8da5a4
							
						
					 | 
					
						
						
							
							[model] support gemma3 (#7273)
						
						
						
						
						
						
					 | 
					
						2025-03-13 01:35:23 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							264538cb26
							
						
					 | 
					
						
						
							
							[misc] upgrade format to py39 (#7256)
						
						
						
						
						
						
					 | 
					
						2025-03-12 00:08:41 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							71a1c1321a
							
						
					 | 
					
						
						
							
							[config] update args (#7231)
						
						
						
						
						
						
						
						Former-commit-id: f71a901840811bf560df671ec63a146ff99140c6 
						
						
					 | 
					
						2025-03-10 23:04:43 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							9f16c50155
							
						
					 | 
					
						
						
							
							[model] add QwQ 32b (#7179)
						
						
						
						
						
						
						
						Former-commit-id: 8897e48b8cd55407812453ddd4ff98ac7bdc4e91 
						
						
					 | 
					
						2025-03-06 11:58:36 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Ze-Yi LIN
							
						 
					 | 
					
						
						
						
						
							
						
						
							11672f760d
							
						
					 | 
					
						
						
							
							[webui] display swanlab exp link (#7089)
						
						
						
						
						
						
						
						* webui add swanlab link
* change callback name
* update
---------
Co-authored-by: hiyouga <hiyouga@buaa.edu.cn>
Former-commit-id: 27a4b93871c63b839c92940766bd7e0177972c9b 
						
						
					 | 
					
						2025-02-27 19:40:54 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Kingsley
							
						 
					 | 
					
						
						
						
						
							
						
						
							2986bef530
							
						
					 | 
					
						
						
							
							[model] add paligemma2-mix series (#7060)
						
						
						
						
						
						
						
						Former-commit-id: 0c0196306d343242ee5e6f22c55562f9a74aa782 
						
						
					 | 
					
						2025-02-25 18:51:16 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							c1d5073bd3
							
						
					 | 
					
						
						
							
							[model] add models (#7054)
						
						
						
						
						
						
						
						* add qwen25vl awq models
* add moonlight
Former-commit-id: ae3be2970fea8a35907202a313ab767381c44916 
						
						
					 | 
					
						2025-02-24 22:05:13 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							a893505924
							
						
					 | 
					
						
						
							
							[misc] fix lora regex (#6944)
						
						
						
						
						
						
						
						* fix lora regex
* fix
Former-commit-id: 1d0ecbaee1b72f1e03154ddd4fcc8b7876e01f89 
						
						
					 | 
					
						2025-02-14 21:38:43 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							46203856fc
							
						
					 | 
					
						
						
							
							[breaking change] refactor data pipeline (#6901)
						
						
						
						
						
						
						
						* refactor data
* rename file
Former-commit-id: 7a1a4ce6451cb782573d0bd9dd27a5e443e3a18b 
						
						
					 | 
					
						2025-02-13 00:39:20 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							528e06ccaa
							
						
					 | 
					
						
						
							
							fix qwen2vl plugin (#6855)
						
						
						
						
						
						
						
						Former-commit-id: fd13b7138ab3f4da0a429a327b9d076bcb70b944 
						
						
					 | 
					
						2025-02-08 10:59:10 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Zhangchi Feng
							
						 
					 | 
					
						
						
						
						
							
						
						
							8f401e37f8
							
						
					 | 
					
						
						
							
							[model] support audio (#6701)
						
						
						
						
						
						
						
						* support qwen2_audio
* improve code
* lint
* fix
* fix
* fix
---------
Co-authored-by: hiyouga <hiyouga@buaa.edu.cn>
Former-commit-id: 5eacb5629e4d7733cd992a63747a1335f2c6a929 
						
						
					 | 
					
						2025-02-05 04:59:09 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							c2022431aa
							
						
					 | 
					
						
						
							
							[misc] update license year & fix llama pro (#6814)
						
						
						
						
						
						
						
						* fix llamapro script
* change year
Former-commit-id: d9ae594178796994d400a5f207d6499712816f89 
						
						
					 | 
					
						2025-02-05 01:53:33 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							a28261a866
							
						
					 | 
					
						
						
							
							[model] add mistral small models (#6786)
						
						
						
						
						
						
						
						Former-commit-id: e5e95c39bc4199fa89c67e34f9adaaa987058744 
						
						
					 | 
					
						2025-02-01 04:31:38 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							800de98dc8
							
						
					 | 
					
						
						
							
							[model] add qwen2.5 vl models (#6779)
						
						
						
						
						
						
						
						Former-commit-id: ed46fb4f6194c30060b908092464dded12e5787c 
						
						
					 | 
					
						2025-01-31 03:00:29 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hoshi-hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							e71737351f
							
						
					 | 
					
						
						
							
							[webui] improve webui & reasoning mode (#6778)
						
						
						
						
						
						
						
						Former-commit-id: 3f17fc0d7163372e0446f1a38792ff761e99b739 
						
						
					 | 
					
						2025-01-31 00:09:21 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								qvlehao
							
						 
					 | 
					
						
						
						
						
							
						
						
							4f298894da
							
						
					 | 
					
						
						
							
							[model] add deepseek-R1 & show think process (#6767)
						
						
						
						
						
						
						
						Former-commit-id: 4dccb724af51208a001c96fefbdbf226be09e50c 
						
						
					 | 
					
						2025-01-29 12:16:26 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Haian Huang(深度眸)
							
						 
					 | 
					
						
						
						
						
							
						
						
							1bb06e06df
							
						
					 | 
					
						
						
							
							Support InternLM3 Dense 8B Model (#6640)
						
						
						
						
						
						
						
						* support internlm3
* update
* update
* update
* add hint
Former-commit-id: 24ab7ae0944c5f373e9cac60f0332e704824a057 
						
						
					 | 
					
						2025-01-14 18:07:27 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Zhangchi Feng
							
						 
					 | 
					
						
						
						
						
							
						
						
							ae32c148d1
							
						
					 | 
					
						
						
							
							Support new features of MiniCPM-V (#6626)
						
						
						
						
						
						
						
						* fix template name
* tiny fix
* support minicpm-o-2.6
Former-commit-id: 53034a61c7654358f46916cbc370910fb2aeff3b 
						
						
					 | 
					
						2025-01-14 00:26:19 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Zhangchi Feng
							
						 
					 | 
					
						
						
						
						
							
						
						
							73c1c15b62
							
						
					 | 
					
						
						
							
							Fix template name of MiniCPM-V (#6620)
						
						
						
						
						
						
						
						* fix template name
* tiny fix
Former-commit-id: 94dea52cef709a7e6f1cdc0b78e83e0422bd65d3 
						
						
					 | 
					
						2025-01-13 16:46:48 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Zhangchi Feng
							
						 
					 | 
					
						
						
						
						
							
						
						
							8c0a721c4c
							
						
					 | 
					
						
						
							
							Merge branch 'main' into minicpmv
						
						
						
						
						
						
						
						Former-commit-id: d8840ae416660e23f1d615ffd404f519360151d9 
						
						
					 | 
					
						2025-01-10 20:12:07 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							867980196e
							
						
					 | 
					
						
						
							
							improve template, add phi4 model
						
						
						
						
						
						
						
						Former-commit-id: a785b6796e445a3adba45c5b6947166a2ff99871 
						
						
					 | 
					
						2025-01-09 18:27:54 +00:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								Zhangchi Feng
							
						 
					 | 
					
						
						
						
						
							
						
						
							08729dbefc
							
						
					 | 
					
						
						
							
							Merge branch 'hiyouga:main' into minicpmv
						
						
						
						
						
						
						
						Former-commit-id: 873b2d5888038e2328a12a6eb7c84099ba7ca1f3 
						
						
					 | 
					
						2025-01-04 11:20:33 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								fzc8578
							
						 
					 | 
					
						
						
						
						
							
						
						
							2c120aa0df
							
						
					 | 
					
						
						
							
							add some
						
						
						
						
						
						
						
						Former-commit-id: 81176fe226da89eace89cb202bad68e73b7c2a02 
						
						
					 | 
					
						2025-01-04 11:11:15 +08:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							8a5b4bdfd4
							
						
					 | 
					
						
						
							
							update model name
						
						
						
						
						
						
						
						Former-commit-id: bf627d9f1ac117f040adbfd7630b5283f0db556a 
						
						
					 | 
					
						2025-01-02 12:19:21 +00:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							18a1a4b9da
							
						
					 | 
					
						
						
							
							add gpt2 model
						
						
						
						
						
						
						
						Former-commit-id: 37d5e3639fcf5ae6e58cc435e0fa9dee0d6e4ead 
						
						
					 | 
					
						2025-01-02 12:07:38 +00:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							b2e4f11602
							
						
					 | 
					
						
						
							
							add deepseek3 model
						
						
						
						
						
						
						
						Former-commit-id: 611779d412f31e25b1ed38049050eee2da61dde5 
						
						
					 | 
					
						2024-12-30 13:39:20 +00:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							3c55976a0e
							
						
					 | 
					
						
						
							
							add qvq #6439
						
						
						
						
						
						
						
						Former-commit-id: 4dbfa142d899dd6e4d1a9d4db125765af5580a4f 
						
						
					 | 
					
						2024-12-25 07:52:41 +00:00 | 
					
					
						
						
							
							
							
						
					 | 
				
			
				
					
						
							
							
								 
								hiyouga
							
						 
					 | 
					
						
						
						
						
							
						
						
							a5346041bb
							
						
					 | 
					
						
						
							
							update readme
						
						
						
						
						
						
						
						Former-commit-id: 1deda4750e0df6c46aeb33cf3f8b35baa537cc1d 
						
						
					 | 
					
						2024-12-23 14:08:59 +00:00 | 
					
					
						
						
							
							
							
						
					 |