Faster Whisper Server logo
AI · self-hosted

Faster Whisper Server in one click.

Faster Whisper Server is an openai-compatible speech-to-text transcription api you can run on your own machine. Launch it in Even Stacks in one click, serve it over trusted HTTPS, and let your AI agents operate it.

What is Faster Whisper Server?

Faster Whisper Server provides an OpenAI-compatible transcription API backed by the faster-whisper inference library. It supports GPU acceleration, multiple model sizes, and concurrent transcription requests.

CategoryAI
Default port8808
Container imagefedirz/faster-whisper-server:latest-cpu

Run Faster Whisper Server in Even

Even Stacks launches Faster Whisper Server as a managed container on your own machine. No compose files, no manual setup.

  1. Open the Even Stacks panel in Even. The container engine starts on demand.
  2. Find Faster Whisper Server in the one-click services and click Launch.
  3. Even pulls the image, starts it on port 8808, and serves it over trusted HTTPS at https://faster-whisper.localhost.

Drive Faster Whisper Server with your AI agents

Even ships an MCP, so any agent you run inside Even (Claude Code, Codex, and more) can operate Faster Whisper Server directly, at both the UI level and the container level.

Through the Even MCP an agent can pull a model, call the local endpoint, and wire it into a tool, all on your own hardware. It runs commands inside the container, reads its logs, and for web apps opens and clicks through the interface in Even's own browser pane. Ask once, for example "set up Faster Whisper Server and get it ready", and the agent handles it end to end.

How agents reach it: Submit audio for transcription via POST http://localhost:8808/v1/audio/transcriptions with multipart form data (field name: file). Response is OpenAI-compatible JSON. Models are downloaded to an internal cache on first use; no persistent data volume is needed.

More AI services

Other one-click services you can launch in Even Stacks.