← All templates

Speaches (OpenAI-compatible speech API)

Published, not hardware-testedociVoice / audio

faster-whisper speech-to-text (and Kokoro/Piper text-to-speech) behind an OpenAI-compatible API.

Needs 3.5 GiB VRAM · CUDA on NVIDIA

No GPU big enough is free right now.

New machines appear as hosts connect them.

No speed benchmarks yet, so we match on price.

Template details
Rating
No ratings yet
Runs
0
Min VRAM
3.8 GB
Backend
CUDA · NVIDIA GPUs
Version
speaches-openai-audio-v1.0.0
Image digest
2b7951985fc3…
Health check
SSH probe on the leased endpoint

OpenAI-compatible API on the leased HTTP endpoint (per-job API key from the rental's access panel). Models download from Hugging Face on first use into /work. Speaches enforcing API_KEY on /v1 and exempting /health is UNVERIFIED. Thin AmpleRun wrapper (templates/images/workspace, variant audio-speaches) over ghcr.io/speaches-ai/speaches:0.8.3-cuda-12.6.3@sha256:c0da392c37e76a01ba479239b43124c67baf8913ae0f491071d6ac544641dad7 (digest verified 2026-09-26); NOT BUILT, so image_digest is a placeholder that can be neither published nor rented. Licence: MIT (https://github.com/speaches-ai/speaches/blob/master/LICENSE). Needs an NVIDIA driver that supports CUDA 12.6 or newer. Status: not hardware-tested.

We match GPUs on their measured runtime and dedicated VRAM. A hardware report doesn't qualify another vendor's backend.

#audio #speech-to-text #tts #openai-api

Reported on real GPUs

People report these themselves. We haven't checked them.

No reports yet.

Sign in to report how this template runs on your GPU.

Reviews

No reviews yet.

Sign in to leave a review.