Alternatives
Self-hosted alternatives to ElevenLabs
Projects in the registry that can run on your own hardware and serve at least one of the same use cases as ElevenLabs.
- Source
- Registry relations over fetched repository records — the evidence sits on each repository page
- Verified
- Evidence not verified
- Confidence
- High
01How this list was built
- Compared against
- ElevenLabs
- Delivery
- SaaS products
- Shared use cases
- Voice agent, AI video workflow, Translation
- Selection rule
- open source, or self-hostable, and shares a use case
- Candidates found
- 7
02Alternatives
ComfyUI
Self-hosted projects
Node-based interface for image and video generation models, where a pipeline is an explicit graph that can be saved, versioned and re-run. Runs locally on your own GPU.
- Repository
- comfyanonymous/ComfyUI
- Stars
- —
- Licence
- unknown
- In stacks
- 0
Shares: Ai video workflow
faster-whisper
Libraries
Reimplementation of Whisper on CTranslate2 with substantially lower memory use and faster inference, including int8 execution on CPU. The usual engine behind self-hosted transcription.
- Repository
- SYSTRAN/faster-whisper
- Stars
- —
- Licence
- unknown
- In stacks
- 2
Shares: Voice agent
LiveKit Agents
Hybrid — hosted or self-hosted
Framework for real-time voice and video agents on LiveKit’s WebRTC infrastructure, with turn detection, interruption handling and pluggable speech and model providers.
- Repository
- livekit/agents
- Stars
- —
- Licence
- unknown
- In stacks
- 1
Shares: Voice agent
Open WebUI
Self-hosted projects
Self-hosted chat interface for local and hosted models, with user accounts, groups, document upload and built-in retrieval. Runs in Docker against Ollama, vLLM or any OpenAI-compatible endpoint.
- Repository
- open-webui/open-webui
- Stars
- 149,870
- Licence
- Other
- In stacks
- 4
Shares: Translation
PaddleOCR
Libraries
OCR toolkit with detection, recognition and layout models, including Chinese and other East Asian scripts, plus table and formula recognition. Runs offline on CPU or GPU.
- Repository
- PaddlePaddle/PaddleOCR
- Stars
- —
- Licence
- unknown
- In stacks
- 0
Shares: Translation
Pipecat
Frameworks
Open-source framework for real-time voice and multimodal agents, composing speech-to-text, model and speech synthesis services into a streaming pipeline with barge-in support.
- Repository
- pipecat-ai/pipecat
- Stars
- —
- Licence
- unknown
- In stacks
- 1
Shares: Voice agent
WhisperX
Libraries
Whisper-based transcription pipeline adding word-level timestamps, forced alignment and speaker diarisation. Runs locally on GPU for batch transcription of recordings.
- Repository
- m-bain/whisperX
- Stars
- —
- Licence
- unknown
- In stacks
- 1
Shares: Translation
03Repository facts
| Repository | Stars | Health |
|---|---|---|
| comfyanonymous/ComfyUI | — | Health not measuredunknown |
| SYSTRAN/faster-whisper | — | Health not measuredunknown |
| livekit/agents | — | Health not measuredunknown |
| open-webui/open-webuiUser-friendly AI Interface (Supports Ollama, OpenAI API, ...) | 150k | Maintenance 100%, Releases 80%, Contributors 50%, Issues 99%, Community 100%88 |
| PaddlePaddle/PaddleOCR | — | Health not measuredunknown |
| pipecat-ai/pipecat | — | Health not measuredunknown |
| m-bain/whisperX | — | Health not measuredunknown |
04Ask about your own case
Which alternative is right depends on how many people will use it, what the data is and what you can operate. Ask and the answer is researched for your organisation.