李泳桦
cb8d87b945
[fix] fix clearing caches synchronization and add more logs ( #4212 )
...
* [fix] fix clearing caches synchronization and add more logs
* [chore] print cache_ready_signal in log
2025-09-23 19:36:38 +08:00
ltd0924
de4feff147
[Feature]CP support data clear ( #4214 )
...
CE Compile Job / ce_job_pre_check (push) Has been cancelled
CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled
CE Compile Job / FD-Clone-Linux (push) Has been cancelled
CE Compile Job / Show Code Archive Output (push) Has been cancelled
CE Compile Job / BUILD_SM8090 (push) Has been cancelled
CE Compile Job / BUILD_SM8689 (push) Has been cancelled
CE Compile Job / CE_UPLOAD (push) Has been cancelled
* Update serving_chat.py
* Update serving_completion.py
* Update serving_completion.py
* mv connection_manager init
* [BugFix] fix kv cache
* fix format
* [Feature] support clear data
---------
Co-authored-by: Yuanle Liu <yuanlehome@163.com >
Co-authored-by: RAM <gstian5555@outlook.com >
2025-09-23 16:53:39 +08:00
RAM
01f6934162
[Executor] Adjust signal sending order in RL training ( #3773 ) ( #4066 ) ( #4178 )
...
CE Compile Job / ce_job_pre_check (push) Has been cancelled
CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled
CE Compile Job / FD-Clone-Linux (push) Has been cancelled
CE Compile Job / Show Code Archive Output (push) Has been cancelled
CE Compile Job / BUILD_SM8090 (push) Has been cancelled
CE Compile Job / BUILD_SM8689 (push) Has been cancelled
CE Compile Job / CE_UPLOAD (push) Has been cancelled
* Adjust processing order
* fix bug
* fix update_parameters bug
* refine code
2025-09-22 14:31:36 +08:00
李泳桦
0fa28b1068
[fix] fix ep group all-reduce ( #4140 )
...
* [fix] fix ep group all-reduce
* [fix] fix clear/update lock not working when workers > 1
* [chore] add preemption triggered info log
* [fix] fix code style
* fix model_weights_signal (#4092 )
* fix model_weights_signal
---------
Co-authored-by: Yuanle Liu <yuanlehome@163.com >
2025-09-18 10:34:49 +08:00
李泳桦
7ccbcc5a62
[feat] support prefix cache clearing when /clear_load_weight
is called ( #4091 )
...
CE Compile Job / ce_job_pre_check (push) Has been cancelled
CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled
CE Compile Job / FD-Clone-Linux (push) Has been cancelled
CE Compile Job / Show Code Archive Output (push) Has been cancelled
CE Compile Job / BUILD_SM8090 (push) Has been cancelled
CE Compile Job / BUILD_SM8689 (push) Has been cancelled
CE Compile Job / CE_UPLOAD (push) Has been cancelled
* [feat] support clearing prefix cache (cherry-picked from release/2.1)
* [fix] fix ipc suffix, use port instead
* [fix] fix prefix caching not enabled
* [fix] fix code style
* [fix] wait for rank0 to update weight status
2025-09-16 11:11:20 +08:00
ltd0924
0f42771a84
[Feature] support model weight update in ep ( #3802 )
...
* Update config.py
* Update ep.py
* Update fused_moe_backend_base.py
* Update dynamic_weight_manager.py
* Update worker_process.py
* fix ci
2025-09-02 20:52:47 +08:00
gaoziyuan
6fdd83da10
fix some bug ( #3434 )
2025-08-18 14:39:13 +08:00
YuanRisheng
502ee92a0a
Unify server-side and model-side Config (Part3) ( #3047 )
...
* merge model config
* fix arch
* fix rl
2025-07-29 17:07:44 +08:00
Zero Rains
25698d56d1
polish code with new pre-commit rule ( #2923 )
2025-07-19 23:19:27 +08:00
Yuanle Liu
dda4a9f848
rl update ( #2861 )
2025-07-16 00:33:10 -07:00
Jiang-Jia-Jun
9fd74f75bd
Update dynamic_weight_manager.py
2025-07-03 15:55:22 +08:00
Jiang-Jia-Jun
05c670e593
[Sync] Update to latest code ( #2679 )
...
* [Sync] Update to latest code
* Add new code files
* Add new code files
* update code
* Try to fix build.sh
* Try to fix build.sh
* Update code
* Update requirements.txt
* Update code
---------
Co-authored-by: Jiang-Jia-Jun <jiangjiajun@baidu.com >
2025-07-03 15:43:53 +08:00