Logo
Explore Help
Sign In
apps/FastDeploy
1
0
Fork 0
You've already forked FastDeploy
mirror of https://github.com/PaddlePaddle/FastDeploy.git synced 2025-10-05 16:48:03 +08:00
Code Issues Actions 2 Packages Projects Releases Wiki Activity
Files
0074b423a94b73fa70153e4ba6af4887dae1cc03
FastDeploy/fastdeploy/model_executor
History
bukejiyu 9408e667a5 [bugfix]fix blockwisefp8 and all_reduce (#3243)
* fix

* update

* fix linear for prequant loader
2025-08-06 23:54:33 +08:00
..
graph_optimization
update flake8 version to support pre-commit in python3.12 (#3000)
2025-07-24 01:43:31 -07:00
guided_decoding
Unify server-side and model-side Config (Part3) (#3047)
2025-07-29 17:07:44 +08:00
layers
[bugfix]fix blockwisefp8 and all_reduce (#3243)
2025-08-06 23:54:33 +08:00
model_loader
qwen3_moe (#3084)
2025-08-06 14:45:27 +08:00
models
qwen3_moe (#3084)
2025-08-06 14:45:27 +08:00
ops
moe preprocess op support 160 experts and fused_moe triton kernel name add K (#3121)
2025-08-01 10:46:20 +08:00
__init__.py
polish code with new pre-commit rule (#2923)
2025-07-19 23:19:27 +08:00
forward_meta.py
[Executor] Refactor GetBlockShapeAndSplitKVBlock Kernel (#2989)
2025-07-31 00:09:31 +08:00
load_weight_utils.py
fix load_pre_sharded_checkpoint (#3152)
2025-08-04 10:44:20 +08:00
pre_and_post_process.py
[stop sequence] support stop sequence (#3025)
2025-07-29 14:17:37 +08:00
Powered by Gitea Version: 1.24.5 Page: 90ms Template: 6ms
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API