Add CUDA acceleration (requires onnxruntime-gpu)

This commit is contained in:
Michael Hansen 2022-04-07 16:54:48 -04:00
commit e56b4579a6
14 changed files with 432 additions and 141 deletions

View file

@ -18,7 +18,7 @@ This will start a web server at `http://localhost:59125`
See `mimic3-server --debug` for more options.
## Endpoints
### Endpoints
* `/api/tts`
* `POST` text or [SSML](#ssml) and receive WAV audio back
@ -30,6 +30,13 @@ See `mimic3-server --debug` for more options.
An [OpenAPI](https://www.openapis.org/) test page is also available at `http://localhost:59125/openapi`
### CUDA Acceleration
If you have a GPU with support for CUDA, you can accelerate synthesis with the `--cuda` flag. This requires you to install the [onnxruntime-gpu](https://pypi.org/project/onnxruntime-gpu/) Python package.
Using [nvidia-docker](https://github.com/NVIDIA/nvidia-docker) is highly recommended. See the `Dockerfile.gpu` file in the parent repository for an example of how to build a compatible container.
## Running the Client
Assuming you have started `mimic3-server` and can access `http://localhost:59125`, then: