bukejiyu
|
29ed617f0f
|
[v1 loader]qwen Offline fp8 (#4036)
* support offline fp8
* update ut
* update ut
* update ut
* fix
* update
* update
|
2025-09-15 13:44:11 +08:00 |
|
周周周
|
17b414c2df
|
MoE Default use triton's blockwise fp8 in TP Case (#3678)
|
2025-08-29 11:07:30 +08:00 |
|
bukejiyu
|
3200a80de3
|
[v1 loader]support fp8 (#3593)
* support fp8
* update ci
|
2025-08-26 02:42:46 -07:00 |
|
xiaoxiaohehe001
|
9afa236e39
|
[NewFeatures] support eplb (#3547)
* [NewFeatures] support eplb
* fix eplb
|
2025-08-26 16:19:30 +08:00 |
|
bukejiyu
|
9408e667a5
|
[bugfix]fix blockwisefp8 and all_reduce (#3243)
* fix
* update
* fix linear for prequant loader
|
2025-08-06 23:54:33 +08:00 |
|
bukejiyu
|
20839abccf
|
qwen3_moe (#3084)
|
2025-08-06 14:45:27 +08:00 |
|
Zero Rains
|
25698d56d1
|
polish code with new pre-commit rule (#2923)
|
2025-07-19 23:19:27 +08:00 |
|
Yuanle Liu
|
61b3997b85
|
refactor rl get_name_mappings_to_training (#2847)
Deploy GitHub Pages / deploy (push) Has been cancelled
* refactor rl get_name_mappings_to_training
* fix tp>1
* change variable name(ffn1->up_gate_proj/ffn2->down_proj)
* change variable name(linear_weight->weight/linear_bias->bias)
* add rl names mapping for vl
* fix ernie 0.3B error
* fix develop code
* fix
|
2025-07-15 07:31:42 -07:00 |
|
chen
|
888780ffde
|
[Feature] block_wise_fp8 support triton_moe_backend (#2767)
|
2025-07-09 19:22:47 +08:00 |
|
Jiang-Jia-Jun
|
92c2cfa2e7
|
Sync v2.0 version of code to github repo
|
2025-06-29 23:29:37 +00:00 |
|