This website requires JavaScript.
Explore
Help
Sign In
apps
/
FastDeploy
Watch
1
Star
0
Fork
0
You've already forked FastDeploy
mirror of
https://github.com/PaddlePaddle/FastDeploy.git
synced
2025-10-05 00:33:03 +08:00
Code
Issues
Actions
2
Packages
Projects
Releases
Wiki
Activity
Files
ac5f86053614085ed9103c86f2bc129fc545aa2c
FastDeploy
/
custom_ops
/
gpu_ops
/
fp8_gemm_with_cutlass
History
Jiang-Jia-Jun
92c2cfa2e7
Sync v2.0 version of code to github repo
2025-06-29 23:29:37 +00:00
..
fp8_common.h
[LLM] First commit the llm deployment code
2025-06-09 19:20:15 +08:00
fp8_fp8_fp8_dual_gemm.cu
[LLM] First commit the llm deployment code
2025-06-09 19:20:15 +08:00
fp8_fp8_half_block_gemm.cu
Sync v2.0 version of code to github repo
2025-06-29 23:29:37 +00:00
fp8_fp8_half_cuda_core_gemm.cu
[LLM] First commit the llm deployment code
2025-06-09 19:20:15 +08:00
fp8_fp8_half_cuda_core_gemm.h
[LLM] First commit the llm deployment code
2025-06-09 19:20:15 +08:00
fp8_fp8_half_gemm.cu
[LLM] First commit the llm deployment code
2025-06-09 19:20:15 +08:00
per_channel_fp8_fp8_half_gemm.cu
[LLM] First commit the llm deployment code
2025-06-09 19:20:15 +08:00