HomeTemplates › Whisper Server
🎙️ Voice AI

Whisper as a server, with your own API

Audio transcription via OpenAI-compatible API

From
loading…
cheapest machine that meets this template
Deploy in one click →
Video memory
8 GB
Setup
~4 min
Access
port 8000
Billing
hourly, in BRL

Whisper transcribes audio accurately, including noisy recordings and strong accents. Running on your machine, you pay for the GPU hour and transcribe as many hours of audio as fit in it — there is no per-minute charge.

faster-whisper-server exposes Whisper large-v3 via 100% OpenAI-compatible API (/v1/audio/transcriptions). High-quality transcription with diarization support.

What it is for

How to deploy

  1. Create your account and add balance (card or Pix, no subscription).
  2. In the console, pick the Whisper Server template and a machine — the console hides the ones that do not meet the requirement.
  3. In about 4 minutes the setup finishes and the access address shows up in the panel, on port 8000.

Done? Just destroy the machine and billing stops with it. No contract, no minimum commitment.

FAQ

How many hours of audio per GPU hour?

With the large model on an 8–16 GB card, transcription runs several times faster than real time, which usually means dozens of audio hours per rented hour.

Does it identify speakers?

Plain Whisper transcribes without separating speakers. Add a diarisation step to your pipeline for that.

Is there a compatible API?

Yes, the server exposes an HTTP API you can call from n8n, a script or your own backend.

Run Whisper Server today
You only pay for the hours the machine is running.
Deploy in one click →

Related templates

F5-TTS
Clone any voice in 5 seconds of audio
Chatterbox TTS
Clone voices with 5s of audio — MIT license, commercial use allowed
OpenVoice v2
MyShell multilingual voice cloning — expressive voice in seconds