hoshi-hiyouga
|
7203365b80
|
[trainer] fix vlm loss for transformers 4.49 (#7448)
|
2025-03-24 10:24:05 +08:00 |
|
hoshi-hiyouga
|
650a9a9057
|
[misc] update format (#7277)
|
2025-03-13 02:53:08 +08:00 |
|
hoshi-hiyouga
|
264538cb26
|
[misc] upgrade format to py39 (#7256)
|
2025-03-12 00:08:41 +08:00 |
|
Billy Cao
|
58e9ca8aa0
|
[trainer] fix gen_kwarg to eval during training (#5451)
* Correctly pass gen_kwarg to eval during model runs
* fix
* fix
---------
Co-authored-by: hiyouga <hiyouga@buaa.edu.cn>
Former-commit-id: 845d16122496311e08263610a6a922f82604de7b
|
2025-02-13 02:35:06 +08:00 |
|
hoshi-hiyouga
|
222423bcef
|
[breaking] support transformers 4.48 (#6628)
Former-commit-id: f154ab175c513a4d7bb866bf2cffc34b77b50508
|
2025-01-31 01:36:33 +08:00 |
|
hoshi-hiyouga
|
2a05941b14
|
[inference] fix stop token for object detection (#6624)
* fix stop token
* update minicpm data pipeline
* fix npu qlora examples
Former-commit-id: 844919fadaa8a61dfae47020971ea80730b2346f
|
2025-01-13 21:34:20 +08:00 |
|
hiyouga
|
2aaf3697d7
|
fix #6499
Former-commit-id: dffc607220ff6dac15cf501ac9a3cdbe80c25211
|
2025-01-02 11:28:54 +00:00 |
|
hiyouga
|
88b1874c04
|
fix #6448
Former-commit-id: 04f78e85af5af14b4c195936623e426a6a128af2
|
2024-12-27 16:54:39 +00:00 |
|
hiyouga
|
a897d46049
|
support report custom args
Former-commit-id: d41254c40a1c5cacf9377096adb27efa9bdb79ea
|
2024-12-21 21:42:45 +00:00 |
|
hoshi-hiyouga
|
0a869c4ed4
|
Merge pull request #6401 from Zeyi-Lin/hiyouga/swanlab
feat: add swanlab for experiment tracking and visualization.
Former-commit-id: e65fe507f7643bf40b0fc462805c7b7f8ef6b738
|
2024-12-21 14:09:33 +08:00 |
|
hiyouga
|
0385c60177
|
fix #6391
Former-commit-id: 067ba6e6cb4d8a1d95bba0a108f73008416a2865
|
2024-12-19 12:16:38 +00:00 |
|
hiyouga
|
01eeae50b5
|
support disable shuffling
Former-commit-id: 9d8c35fd6b838ede0bd6827c6c6121f2cba2b11b
|
2024-12-19 08:53:21 +00:00 |
|
hiyouga
|
7eeeffdb8a
|
add swanlab
Former-commit-id: c85a77c8a8824a56a67d56b97b4877fcd6edeb3d
|
2024-12-19 07:12:31 +00:00 |
|
hiyouga
|
19ebc0e7a2
|
support control eos, fix #6345
Former-commit-id: cb0f8399356bf372f3b7963f2565c3d504be0923
|
2024-12-17 10:42:05 +00:00 |
|
hiyouga
|
aacd9642f5
|
fix #6348
Former-commit-id: 83e552320909f4775377889f1512994b7e638a7e
|
2024-12-17 10:06:46 +00:00 |
|
hiyouga
|
fb22651faf
|
fix mrope
Former-commit-id: 55bee1d333549ca19858b3f5c1b7b86926e5fb09
|
2024-12-12 15:08:17 +00:00 |
|
hiyouga
|
c1768cfb14
|
support batch infer in vllm
Former-commit-id: 3ef5ed3b9a44eed2f7e3ff221dfc343d0a97c0b5
|
2024-12-04 13:50:00 +00:00 |
|
hiyouga
|
2bb3255e74
|
fix dpo metrics
Former-commit-id: 57029280da825a39fbf5a05097921b861f126669
|
2024-11-02 20:59:01 +08:00 |
|
hiyouga
|
093eda2ad6
|
support rank0 logger
Former-commit-id: 84528eabe560091bfd866b6a0ca864085af7529b
|
2024-11-02 18:31:04 +08:00 |
|
hiyouga
|
8185eb1890
|
fix incorrect loss value for vlms
Former-commit-id: 0aa29a71ce958343a2086090d647eb63b8f5f5be
|
2024-10-30 08:56:46 +00:00 |
|
hiyouga
|
7f71276ad8
|
add docstrings, refactor logger
Former-commit-id: c34e489d71f8f539028543ccf8ee92cecedd6276
|
2024-09-08 00:56:56 +08:00 |
|
hoshi-hiyouga
|
7367c6ec21
|
fix trainer predict
Former-commit-id: 2790790cd26c6743105555a60523b89f367ebce3
|
2024-09-02 10:15:29 +08:00 |
|
hiyouga
|
d789b667d7
|
optimize predict vram
Former-commit-id: a577e44eee351b3ed8011a33ae01cd713354ff97
|
2024-08-30 23:08:45 +08:00 |
|
moontidef
|
33a90b9026
|
fix: rename optimzer to optimizer
Former-commit-id: 186dc1fde822e6a603ac273538741ea3853f243e
|
2024-08-07 10:05:01 +08:00 |
|
hiyouga
|
884b49e662
|
add eval acc
Former-commit-id: 7ffde76fbfb6192e3aac31ccc098f31ce89181ae
|
2024-07-01 03:51:20 +08:00 |
|
hiyouga
|
46f0189e88
|
refactor pissa, improve llamaboard
Former-commit-id: 619556e46c19718f702c97df5d570a2a4c5fb13a
|
2024-06-28 01:04:24 +08:00 |
|
hzhaoy
|
89d9dd5aa5
|
fix #4579
Former-commit-id: 0fa298ff6a4febea36ea9f11c7594277a77e6e9b
|
2024-06-27 13:49:57 +08:00 |
|
hiyouga
|
9fd7a410bb
|
tiny fix about badam
Former-commit-id: 03f49267c7406e36aee35639f86e6e0383897090
|
2024-06-25 01:54:53 +08:00 |
|
Jonery
|
fa3150548e
|
Cleaner integration.
Former-commit-id: 26d4b05d424bd71f570195dd433258caf6465d92
|
2024-06-19 12:29:40 +08:00 |
|
Jonery
|
12fcfc2b72
|
Support distributed BAdam.
Former-commit-id: bdcb986e37975911c190a74d3e60bb77aa2033bd
|
2024-06-18 12:27:47 +08:00 |
|
Jonery
|
95ae30f678
|
Merge remote-tracking branch 'upstream/main'
Former-commit-id: 37834a7e79473ccf50ad7f67745b97c274c326d9
|
2024-06-17 18:44:51 +08:00 |
|
Jonery
|
ba303fd1aa
|
adapt for badam with ds zero3
Former-commit-id: fff2a020ec8713022bd8145f4a7168168ea07ca4
|
2024-06-17 18:18:10 +08:00 |
|
hiyouga
|
32f45c9e91
|
support pissa
Former-commit-id: ef8e45f2eaf466c54e9a671512a2974575677b08
|
2024-06-16 01:08:12 +08:00 |
|
hiyouga
|
bb88536166
|
add license
Former-commit-id: 69cfc98d7c81756a5ab6bf962240e393e449fef0
|
2024-06-15 17:54:33 +08:00 |
|
hiyouga
|
a30931fe0f
|
fix #4295
Former-commit-id: 08f657868f9d605b837c5d8c2946a25cc05c8735
|
2024-06-15 04:34:55 +08:00 |
|
hiyouga
|
fcb134e144
|
rename files
Former-commit-id: e1a8431770fc36c0c9ee7fed4abbc3d7fdcc5efd
|
2024-06-07 00:09:06 +08:00 |
|
hiyouga
|
dfa686b617
|
rename package
Former-commit-id: a07ff0c083558cfe6f474d13027642d3052fee08
|
2024-05-16 18:39:08 +08:00 |
|