Our own product · Platform and cloud
A GPU transcription service that costs nothing when nobody is talking, and rejects unauthorised requests before a container starts
We built our own streaming and batch speech-to-text service to sit alongside a managed provider — partly for cost at volume, partly for the workloads that cannot leave our infrastructure. The design problem was idle GPU cost, and the answer was to have no idle GPUs.
Python · WhisperLiveKit · WhisperX · Modal · RunPod · Docker · NVIDIA T4