mirror of https://github.com/PaddlePaddle/FastDeploy.git synced 2025-12-24 13:28:13 +08:00

Files

charl-u 02eab973ce [Doc]Add English version of documents in docs/cn and api/vision_results (#931 )

* 第一次提交

* 补充一处漏翻译

* deleted:    docs/en/quantize.md

* Update one translation

* Update en version

* Update one translation in code

* Standardize one writing

* Standardize one writing

* Update some en version

* Fix a grammer problem

* Update en version for api/vision result

* Merge branch 'develop' of https://github.com/charl-u/FastDeploy into develop

* Checkout the link in README in vision_results/ to the en documents

* Modify a title

* Add link to serving/docs/

* Finish translation of demo.md

2022-12-22 18:15:01 +08:00

docs

[Doc]Add English version of documents in docs/cn and api/vision_results (#931 )

2022-12-22 18:15:01 +08:00

scripts

[backend][Serving]Fix paddle backend get outout tensor error (#741 )

2022-11-29 18:34:56 +08:00

src

[Serving][Backend] Backend support zero_copy_infer and Serving reduce the output memory copy (#703 )

2022-11-28 14:07:53 +08:00

CMakeLists.txt

support build cpu images (#341 )

2022-10-11 14:17:27 +08:00

Dockerfile

[Doc][Serving]modify serving doc (#718 )

2022-11-28 15:14:10 +08:00

Dockerfile_cpu

[Serving]fix cpu images version (#736 )

2022-11-29 11:10:43 +08:00

Dockerfile_ipu

[Serving]: add ipu support for serving. (#10 ) (#470 )

2022-11-02 09:50:58 +08:00

README_CN.md

fix docker serving doc (#899 )

2022-12-19 10:21:43 +08:00

README_EN.md

fix docker serving doc (#899 )

2022-12-19 10:21:43 +08:00

README.md

Fd serving add docker images correlation and docs (#311 )

2022-10-08 16:08:07 +08:00

README_EN.md

简体中文 | English

FastDeploy Serving Deployment

Introduction

FastDeploy builds an end-to-end serving deployment based on Triton Inference Server. The underlying backend uses the FastDeploy high-performance Runtime module and integrates the FastDeploy pre- and post-processing modules to achieve end-to-end serving deployment. It can achieve fast deployment with easy-to-use process and excellent performance.

Prepare the environment

Environment requirements

Linux
If using a GPU image, NVIDIA Driver >= 470 is required (for older Tesla architecture GPUs, such as T4, the NVIDIA Driver can be 418.40+, 440.33+, 450.51+, 460.27+)

Obtain Image

CPU Image

CPU images only support Paddle/ONNX models for serving deployment on CPUs, and supported inference backends include OpenVINO, Paddle Inference, and ONNX Runtime

docker pull registry.baidubce.com/paddlepaddle/fastdeploy:1.0.1-cpu-only-21.10

GPU Image

GPU images support Paddle/ONNX models for serving deployment on GPU and CPU, and supported inference backends including OpenVINO, TensorRT, Paddle Inference, and ONNX Runtime

docker pull registry.baidubce.com/paddlepaddle/fastdeploy:1.0.1-gpu-cuda11.4-trt8.4-21.10

Users can also compile the image by themselves according to their own needs, referring to the following documents:

FastDeploy Serving Deployment Image Compilation

Task	Model
Classification	PaddleClas
Detection	PaddleDetection
Detection	ultralytics/YOLOv5
NLP	PaddleNLP/ERNIE-3.0
NLP	PaddleNLP/UIE
Speech	PaddleSpeech/PP-TTS
OCR	PaddleOCR/PP-OCRv3

README_EN.md

FastDeploy Serving Deployment

Introduction

Prepare the environment

Environment requirements

Obtain Image

CPU Image

GPU Image

Other Tutorials

Serving Deployment Demo