Commit Graph

23 Commits

Author SHA1 Message Date
hiyouga
bf5ffeeae0 simplify readme
Former-commit-id: 92dab8a90b
2024-04-02 20:07:43 +08:00
hiyouga
e7ade84bba fix plots
Former-commit-id: 5907216a1c
2024-03-31 19:43:48 +08:00
hiyouga
2f878bde11 support ORPO
Former-commit-id: 17bf8a2c3a
2024-03-31 18:29:50 +08:00
hiyouga
89c400633a update trainers
Former-commit-id: 8c77b10912
2024-03-28 18:16:27 +08:00
hoshi-hiyouga
ae9ad13f2a fix ds optimizer
Former-commit-id: 3bcd41b639
2024-03-26 23:39:56 +08:00
hiyouga
27151b8c65 release v0.6.0
Former-commit-id: 6f2b563f12
2024-03-25 22:38:56 +08:00
hiyouga
8717e98200 fix #2777 #2895
Former-commit-id: 9bec3c98a2
2024-03-20 17:59:45 +08:00
hiyouga
4a4e4b4354 support layerwise galore
Former-commit-id: 8664262cde
2024-03-10 00:24:11 +08:00
hiyouga
868444e124 allow non-packing pretraining
Former-commit-id: bdb496644c
2024-03-09 22:21:46 +08:00
hiyouga
2c010c72b8 support galore
Former-commit-id: 28f7862188
2024-03-07 22:41:36 +08:00
hiyouga
d1e6e02461 fix #2649
Former-commit-id: 4e5fae2fac
2024-03-01 13:02:41 +08:00
hiyouga
b27e91222c format style
Former-commit-id: 638234ceee
2024-01-20 20:15:56 +08:00
hiyouga
2f7684a8ee fix tests
Former-commit-id: f6d6e00337
2024-01-20 19:58:04 +08:00
hiyouga
4e3bfb799d support function calling
Former-commit-id: d9f1cae351
2024-01-18 09:54:23 +08:00
hiyouga
6378864390 fix #2161
Former-commit-id: 898ec3696a
2024-01-11 17:04:13 +08:00
hiyouga
61960189b2 fix #1789
Former-commit-id: 4571068e1e
2024-01-09 18:31:27 +08:00
hiyouga
1cb390b9b2 implement rm server #1543
Former-commit-id: 7df4f3ab20
2023-12-03 20:52:54 +08:00
hiyouga
ba6d290d0b fix #1668
Former-commit-id: 1585962eb7
2023-11-30 21:02:00 +08:00
hiyouga
ecfc7d1b50 fix #1658
Former-commit-id: 77d1b14fc2
2023-11-28 20:57:24 +08:00
hiyouga
682d81caa9 fix #1567
Former-commit-id: 99a3f06377
2023-11-20 18:46:36 +08:00
hiyouga
48d6d925f7 fix #1558
Former-commit-id: 1740131d63
2023-11-19 14:15:47 +08:00
hiyouga
f441932bd1 support full-parameter PPO
Former-commit-id: ce78303600
2023-11-16 02:08:04 +08:00
hiyouga
06a4820836 disentangle model from tuner and rename modules
Former-commit-id: 4736344eb1
2023-11-15 16:29:09 +08:00