FastDeploy

mirror of https://github.com/PaddlePaddle/FastDeploy.git synced 2025-12-24 13:28:13 +08:00

Author	SHA1	Message	Date
Jiaxin Sui	8e0f4dfd0c	[XPU] [CI] Xpu Ci Refactor (#5252 ) * add xpu ci * add case * add case * fix ci bug * Update Docker image tag to 'latest' in CI workflow * Fix set -e usage in run_xpu_ci_pytest.sh * add pd case * add case * Configure pip to use Tsinghua mirror for dependencies Set the global pip index URL to Tsinghua mirror. * fix ci bug * fix bug * fix bug --------- Co-authored-by: suijiaxin <suijiaxin@Suis-MacBook-Pro.local> Co-authored-by: root <root@gajl-bbc-onlinec-com-1511964.gajl.baidu.com> Co-authored-by: root <root@gajl-bbc-onlinec-com-1511972.gajl.baidu.com>	2025-12-02 17:15:51 +08:00
YuBaoku	69e003abcb	[CI] Fix return_code check in test_chunked_moe.py (#5326 )	2025-12-02 15:41:26 +08:00
lizexu123	c563eca791	[Feature] support reward model (#5301 ) * Your commit message here * add test * update develop * support reward * support enable_chunk_prefill * support bingfa * support convert is reward * update test * delete print * fix enable_thinking * add document * fix place * fix test * fix * support enable_prefix_caching * add no-enable_prefix-caching test * fix * support enable_prefix_caching * delete print * fix document * fix * fix test * fix document and delete chinese * udpate * enable_thinking * fix test	2025-12-02 14:55:31 +08:00
qwes5s5	117980dd4e	[LogProbs]Enable prompt logprobs output and modify data transmission method for the online interface. (#5089 ) * add prompt logprobs * Merge prompt_logprobs_tensors and prompt_logprobs * fix param check * trigger ci * fix unitest * fix logprobs bug	2025-12-02 13:49:51 +08:00
YuanRisheng	af39819fcd	Revert "[CI] 【Hackathon 9th Sprint No.18】NO.18 功能模块单测补充 (#5064 )" (#5290 ) This reverts commit `7bac016c77`.	2025-12-02 13:43:36 +08:00
YuanRisheng	ded7765dec	Revert "[CI] 【Hackathon 9th Sprint No.41】NO.41 功能模块单测补充 (#5062 )" (#5291 ) This reverts commit `373b5c3807`.	2025-12-02 13:43:13 +08:00
YuBaoku	68533ebd95	[CI] disable test_chunked_moe.py in unit_test (#5322 )	2025-12-02 10:39:50 +08:00
xiaolei373	84e2f6aa75	[CI]add clear to run-batch ci (#5307 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details Publish Job / publish_pre_check (push) Has been cancelled Details Publish Job / print_publish_pre_check_outputs (push) Has been cancelled Details Publish Job / FD-Clone-Linux (push) Has been cancelled Details Publish Job / Show Code Archive Output (push) Has been cancelled Details Publish Job / BUILD_SM8090 (push) Has been cancelled Details Publish Job / BUILD_SM8689 (push) Has been cancelled Details Publish Job / PADDLE_PYPI_UPLOAD_8090 (push) Has been cancelled Details Publish Job / PADDLE_PYPI_UPLOAD_8689 (push) Has been cancelled Details Publish Job / Run FD Image Build (push) Has been cancelled Details Publish Job / Run FastDeploy Unit Tests and Coverage (push) Has been cancelled Details Publish Job / Run FastDeploy LogProb Tests (push) Has been cancelled Details Publish Job / Extracted partial CE model tasks to run in CI. (push) Has been cancelled Details Publish Job / Run Base Tests (push) Has been cancelled Details Publish Job / Run Accuracy Tests (push) Has been cancelled Details Publish Job / Run Stable Tests (push) Has been cancelled Details CI Images Build / FD-Clone-Linux (push) Has been cancelled Details CI Images Build / Show Code Archive Output (push) Has been cancelled Details CI Images Build / CI Images Build (push) Has been cancelled Details CI Images Build / BUILD_SM8090 (push) Has been cancelled Details CI Images Build / Run FastDeploy Unit Tests and Coverage (push) Has been cancelled Details CI Images Build / Run FastDeploy LogProb Tests (push) Has been cancelled Details CI Images Build / Extracted partial CE model tasks to run in CI. (push) Has been cancelled Details CI Images Build / Run Base Tests (push) Has been cancelled Details CI Images Build / Publish Docker Images Pre Check (push) Has been cancelled Details	2025-12-01 21:18:19 +08:00
Jiaxin Sui	b0113cb0fc	[XPU][CI] Change XPU CI Base Value (#5318 ) * Add '小度' keyword to assertion in run_w4a8.py * Add keywords to assertion in run_ep_online.py * Add keywords to assertion in run_w4a8.py * Update run_45T.py * Update run_ep_online.py * Refactor assertion for response content keywords * Update run_w4a8.py * Update run_w4a8.py	2025-12-01 21:02:09 +08:00
Juncai	0925d44f18	[PD Disaggregation] support different tp_size for prefill and decode (#5296 ) * up * up * up * fix	2025-12-01 17:50:20 +08:00
Jiaxin Sui	b467e9dadc	[XPU][CI]Change W4A8 Case Base Value (#5309 )	2025-12-01 15:25:33 +08:00
Longzhi Wang	add524d80c	[Feature] support chunked moe (#4575 ) * [Feature] support chunked moe * update * update * fix and add test * update * fix conflict and modity test * fix fused_moe * fix fused_moe * fix docstring * fix * fix typo * fix test * fix * fix * fix test * fix test	2025-12-01 15:17:18 +08:00
Jundong Liu	6f42c37359	[Deterministic] Move paddle version batch invariant pkg to Fastdeploy (#4763 ) * Move batch invariant pkg to Fastdeploy * fix problem and pre-commit * move test * Change testcase to FD style * Add testcase for log_softmax * Add testcase for mean * Add testcase for addmm * fix pre-commit * API check v0.9 * move to layers and add comment about log_softmax * Update fastdeploy/model_executor/layers/batch_invariant_ops/batch_invariant_ops.py 存在于原版代码注释中的版本控制遗留的内容，确实应该去除 Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * Update tests/batch_invariant/test_batch_invariance_op_mean.py Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * Update tests/batch_invariant/test_batch_invariance_op_logsoftmax.py Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * Update fastdeploy/model_executor/layers/batch_invariant_ops/batch_invariant_ops.py Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> * change comment after copilot fix * fix bug about addmm * avoid global effect by enable_torch_proxy --------- Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com> Co-authored-by: YuBaoku <49938469+EmmonsCurse@users.noreply.github.com>	2025-12-01 11:25:48 +08:00
Yonghua Li	a535050b11	[FDConfig] remove engine client args, use fd_config instead (#5217 ) * [refactor] remove engine client args, use fd_config instead * [chore] update * [fix] fix * [fix] fix * [chore] rename config to fd_config * [fix] fix run_batch * [ci] add ci case for engine client --------- Co-authored-by: Jiaxin Sui <95567040+plusNew001@users.noreply.github.com>	2025-11-28 01:20:54 -08:00
kevin	2d69d91ab8	add aksk check (#5273 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details	2025-11-28 14:28:24 +08:00
Juncai	1a559c973f	Revert "[CI] 【Hackathon 9th Sprint No.33】NO.33 功能模块单测补充 (#5056 )" (#5286 ) This reverts commit `a12eaf9171`.	2025-11-28 10:48:35 +08:00
ddchenhao66	fc88eebc32	[CI][XPU] add pd disaggregation (#5179 ) * [CI][XPU] add pd disaggregation * Clarify comments and install iproute2 Updated comments to clarify script purpose and added installation of iproute2. --------- Co-authored-by: ddchenhao66 <dhaochen163.com> Co-authored-by: Jiaxin Sui <95567040+plusNew001@users.noreply.github.com>	2025-11-28 10:44:27 +08:00
lizhenyun01	aba4fc657f	[Feature] support flash_mask_attention backend (#5134 ) * [Feature] suppert flash_mask_attention backend * fix unittest * clean code	2025-11-28 10:12:16 +08:00
Divano	b935101008	Create test_prompt_ids.py	2025-11-28 10:11:51 +08:00
YuBaoku	6a6bf4ea24	[CI] Fix test streaming with stop str (#5275 ) * [CI] add output for last_token in test_streaming_with_stop_str * [CI] Adapt empty last_token check	2025-11-27 20:51:39 +08:00
chen	35f85baf09	[BugFix]fix v1 loader lm head fp32 (#5270 )	2025-11-27 20:12:56 +08:00
xiaolei373	b52ec268f7	[CI]fix run batch unit test (#4628 )	2025-11-27 19:38:04 +08:00
YuBaoku	1372d6d01d	[CI] disable test_engine_client.py unit test (#5272 )	2025-11-27 17:37:54 +08:00
fl0w2o48	e63d715fc3	[BugFix][Metrics] Fix Prometheus Multiprocess Metrics Issues and Add ZMQ Communication Metrics (#5185 ) * [Feature] add metrics for ZMQ and fix multiprocess metrics * fix test_metrics.py --------- Co-authored-by: Jiaxin Sui <95567040+plusNew001@users.noreply.github.com>	2025-11-27 15:05:09 +08:00
Juncai	ce9a49f6bf	[PD Disaggregation] Add unittest for splitwise deployment with using rdma (#5189 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details * Add splitwise deployment with using rdma * clean cuda	2025-11-27 14:27:17 +08:00
xunyoyo	373b5c3807	[CI] 【Hackathon 9th Sprint No.41】NO.41 功能模块单测补充 (#5062 ) * Add tests for SplitwiseConnector functionality This commit introduces a comprehensive test suite for the SplitwiseConnector class, implementing various tests to ensure the correct functionality of task dispatching, message sending, and connection handling. The tests cover scenarios for both prefill and decode roles, including checks for task promotion, message serialization, and error handling. * Add innode splitwise test helpers * Refine Splitwise connector test stubs * Add to_tensor stub for splitwise tests * Update splitwise connector tests	2025-11-27 14:24:19 +08:00
essos	84c7fa49a5	[CI]【Hackathon 9th Sprint No.50】NO.50 功能模块 fastdeploy/entrypoints/engine_client.py 单测补充 (#5045 ) * update test utils * update test utils code * update test file name * Add engine client tests and documentation - Add CLAUDE.md documentation - Update test_engine_client.py with new test cases 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude <noreply@anthropic.com> * Fix import errors and assertion failures in test_engine_client.py for PR #5045 - Add missing mock for fastdeploy.entrypoints.engine_client module - Fix AssertionError: max_model_len parameter validation (1024 vs 2048) - Implement flexible assertions to handle parameter validation differences - Use assertIsInstance for boolean parameters instead of exact value matching - Apply SOP容错测试模式 for CI environment compatibility - All pre-commit checks pass (black, isort, flake8, ruff) 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude <noreply@anthropic.com> * fix with mock * add more test to new code --------- Co-authored-by: Claude <noreply@anthropic.com>	2025-11-27 12:43:00 +08:00
SunLei	c424e08dc5	[Speculative Decoding] split draft_tokens into standalone post-processing path (#5205 ) * refactor(mtp): split draft_tokens into standalone post-processing path for MTP + logprobs * Restore Request.__repr__ implementation * ci * add envs * fix unittest	2025-11-27 11:22:41 +08:00
xunyoyo	a12eaf9171	[CI] 【Hackathon 9th Sprint No.33】NO.33 功能模块单测补充 (#5056 ) * Add cache messager unit tests * Refactor test_cache_messager.py with new stubs Updated copyright information and modified function names for clarity. * Add missing stubs for cache messager tests --------- Co-authored-by: Tao Luo <luotao02@baidu.com>	2025-11-27 11:05:50 +08:00
Yonghua Li	cead6b26fa	[Metrics] Update time_to_first_token to include tokenization & queue time, and remove redundant metrics (#4993 ) * [update] update time_to_first_tokens to include queue time, and remove first_token_latency and infer_latency * [doc] update docs * [ci] fix test * [chore] delete redundant code --------- Co-authored-by: Jiaxin Sui <95567040+plusNew001@users.noreply.github.com>	2025-11-26 14:42:17 +08:00
kxz2002	2d787590c4	[Feature] The 45VL supports prompt_token_ids + messages input. (#5148 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details * support prompt_token_ids + messages * fix bug * refact code structure * support cache mm items * refact code structure * delete test cases * modify unit test * add unit test * add unit test * fix append * add check for messages	2025-11-25 23:11:44 +08:00
Yonghua Li	09379183e2	[BugFix] fix work metrics not returned by metrics api (#4912 ) * [BugFix] fix work metrics not returned by metrics api * [fix] fix conflict * [fix] fix ci	2025-11-25 19:12:29 +08:00
xunyoyo	edf0d09257	[CI] 【Hackathon 9th Sprint No.24】NO.24 功能模块单测补充 (#5055 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details * Add tp_utils tests * Add header and tidy tp_utils test stubs	2025-11-25 11:34:57 +08:00
xunyoyo	daf8b386eb	[CI] 【Hackathon 9th Sprint No.17】NO.17 功能模块单测补充 (#5054 ) * Refactor text processor tests to use unittest * Add helpers for text processor tests	2025-11-25 11:32:27 +08:00
Echo-Nie	a418d7b60b	[CI] Add Unittest (#5187 ) * add test * Delete tests/model_executor/test_w4afp8.py * Rename test_utils.py to test_tool_parsers_utils.py * add test * add test * fix platforms * Delete tests/cache_manager/test_platforms.py * dont change Removed copyright notice and license information.	2025-11-25 11:00:34 +08:00
kevin	8e4e3ff510	[Feature] support eplb in api_server (#4782 ) * support eplb in api_server * update code * add eplb test case * update eplb * support tp+dp eplb * update test cese * update code * update code * fix bug * update copilot review * update test case name	2025-11-24 20:22:29 +08:00
Jiaxin Sui	5ff93d4998	[XPU][CI] change VL model to 28B-VL-thinking (#5169 ) * Enhance run_ci_xpu.sh with caching and prefill options * Update model path and configuration in run_ci_xpu.sh * Add '北朝' keyword to assertion in run_45vl.py * Enhance process termination logic in run_ci_xpu.sh * Set timeout for CI_XPU job to 60 minutes * Remove extra newline in stop_processes function	2025-11-24 16:50:18 +08:00
xunyoyo	7bac016c77	[CI] 【Hackathon 9th Sprint No.18】NO.18 功能模块单测补充 (#5064 ) * Add unit tests for DeepEP buffer functionality This file contains unit tests for the DeepEP buffer helpers and runners, including various test cases for buffer allocation, cleanup, and dispatching processes. * Refactor DeepEP tests to use scoped stubs * Add licensing information to test_ep.py Added licensing information to the test file.	2025-11-24 15:52:34 +08:00
YuBaoku	98f1ab46a9	[CI] add output for last_token in test_streaming_with_stop_str (#5170 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details	2025-11-24 10:49:17 +08:00
周周周	e297406263	[Others] unitest tests/layers/test_attention_layer.py (#5174 )	2025-11-23 22:21:01 +08:00
kevin	cceaba1c8d	[Feature] remove to_numpy (#5162 ) * remove to_numpy * update code * update name * update code * update code * update code	2025-11-21 21:54:26 +08:00
kevin	c068a4f642	[Feature] dyc8 support prefixcache (#5125 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details Publish Job / publish_pre_check (push) Has been cancelled Details Publish Job / print_publish_pre_check_outputs (push) Has been cancelled Details Publish Job / FD-Clone-Linux (push) Has been cancelled Details Publish Job / Show Code Archive Output (push) Has been cancelled Details Publish Job / BUILD_SM8090 (push) Has been cancelled Details Publish Job / BUILD_SM8689 (push) Has been cancelled Details Publish Job / PADDLE_PYPI_UPLOAD_8090 (push) Has been cancelled Details Publish Job / PADDLE_PYPI_UPLOAD_8689 (push) Has been cancelled Details Publish Job / Run FD Image Build (push) Has been cancelled Details Publish Job / Run FastDeploy Unit Tests and Coverage (push) Has been cancelled Details Publish Job / Run FastDeploy LogProb Tests (push) Has been cancelled Details Publish Job / Extracted partial CE model tasks to run in CI. (push) Has been cancelled Details Publish Job / Run Base Tests (push) Has been cancelled Details Publish Job / Run Accuracy Tests (push) Has been cancelled Details Publish Job / Run Stable Tests (push) Has been cancelled Details CI Images Build / FD-Clone-Linux (push) Has been cancelled Details CI Images Build / Show Code Archive Output (push) Has been cancelled Details CI Images Build / CI Images Build (push) Has been cancelled Details CI Images Build / BUILD_SM8090 (push) Has been cancelled Details CI Images Build / Run FastDeploy Unit Tests and Coverage (push) Has been cancelled Details CI Images Build / Run FastDeploy LogProb Tests (push) Has been cancelled Details CI Images Build / Extracted partial CE model tasks to run in CI. (push) Has been cancelled Details CI Images Build / Run Base Tests (push) Has been cancelled Details CI Images Build / Publish Docker Images Pre Check (push) Has been cancelled Details * dyc8 support prefixcache * fix cache_trans test case * update code	2025-11-21 19:46:26 +08:00
chenjian	3ea1b44a58	[Optimization] Improve perf for fd response token with internal adapter (#4992 ) * [Optimize] Improve perf for fd response token with internal adapter * fix * fix bug * fix ci * fix ci * fix ci * fix ci	2025-11-21 19:02:03 +08:00
xiaoxiaohehe001	6ca2651995	[Feature] Support noaux for eplb (#5143 ) * support noaux eplb * noaux_eplb * noaux_eplb * noaux_eplb	2025-11-21 14:10:32 +08:00
essos	79f18331b6	[CI]【Hackathon 9th Sprint No.51】NO.51 功能模块 fastdeploy/scheduler/dp_scheduler.py 单测补充 (#5046 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details * update test utils * Add comprehensive unit tests for DP scheduler functionality - Add test_dp_scheduler.py with full-featured unit tests supporting both normal and standalone modes - Add test_dp_scheduler_simple.py with lightweight mock-based tests for easy execution - Add comprehensive README.md documenting test architecture and usage - Tests cover DPLocalScheduler and DPScheduler classes with focus on: - Request lifecycle management and TTL support - Response handling and routing - Resource-based scheduling and constraint handling - Multi-threading and concurrent operations - Splitwise role support (prefill vs decode) - Error handling and edge cases - Thread-safe operations with proper synchronization 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude <noreply@anthropic.com> * Remove tests/multimodal/test_utils.py This file appears to be duplicate or misplaced, removing it to clean up the test structure. 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude <noreply@anthropic.com> * update * fix * rm unused file --------- Co-authored-by: Claude <noreply@anthropic.com>	2025-11-21 10:52:33 +08:00
kevin	7454480e07	[Feature] support bos download retry (#5137 ) * support bos download retry * update code * update code	2025-11-21 10:18:32 +08:00
Yonghua Li	43097a512a	[BugFix] [PD Disaggregation] fix v1 scheduler prefill node profile run & ipc transfer protocol (#5132 ) Some checks failed CE Compile Job / ce_job_pre_check (push) Has been cancelled Details CE Compile Job / print_ce_job_pre_check_outputs (push) Has been cancelled Details CE Compile Job / FD-Clone-Linux (push) Has been cancelled Details CE Compile Job / Show Code Archive Output (push) Has been cancelled Details CE Compile Job / BUILD_SM8090 (push) Has been cancelled Details CE Compile Job / BUILD_SM8689 (push) Has been cancelled Details CE Compile Job / CE_UPLOAD (push) Has been cancelled Details Deploy GitHub Pages / deploy (push) Has been cancelled Details * [fix] fix v1 scheduler profile run for append attention in prefill node * [fix] skip send_signal if kv signal not inited for gpu and xpu * [fix] extend fix to flash_attn & mla_attn * [fix] fix v1 pd run in ipc transfer protocol * [ci] add test for v1 pd profile run using ipc transfer protocol * [style] fix code style check * [style] fix code style again * [fix] fix profile run * [update] remove --num-gpu-blocks-override in example script * [chore] rename forward_meta is_profiling to is_dummy_or_profile_run	2025-11-20 21:39:22 +08:00
周周周	385fe6dade	[Others] clean code (#5133 )	2025-11-20 18:44:08 +08:00
周周周	6fa34102e8	[Others]get_block_shape_and_split_kv_block clean code (#5123 )	2025-11-20 16:40:04 +08:00
yangjianfengo1	af715db763	[Scheduler] Support chunk prefill for video input (#5107 ) * add video chunk prefill * add vit_merge=True for test_tokenizer_client.py --------- Co-authored-by: YuBaoku <49938469+EmmonsCurse@users.noreply.github.com>	2025-11-20 16:29:13 +08:00

1 2 3 4 5 ...

472 Commits