Run LLMs, image, and audio models locally with an OpenAI-compatible API, optional GPU acceleration, and a built-in web UI for managing and testing models.

48.7k stars4.4k forksMITlast commit Actively maintained
Cloud text-to-speech platform that converts text into realistic, multi-speaker audio. Offers voice cloning, speech styles, SSML/pronunciation controls, multi-language support, multi-voice dialogues, and a low-latency API for integration into apps, videos, podcasts, IVR and localization.
Run LLMs, image, and audio models locally with an OpenAI-compatible API, optional GPU acceleration, and a built-in web UI for managing and testing models.

48.7k stars4.4k forksMITlast commit Actively maintained
Self-hostable tool to convert non-DRM eBooks into audiobooks with chapter support, metadata, multilingual TTS engines, and optional voice cloning via a web UI or CLI.
20.1k stars1.7k forksApache-2.0last commit Actively maintained
Self-hosted, OpenAI API-compatible server for streaming transcription, translation, and speech generation using faster-whisper and TTS engines like Piper and Kokoro.

3.6k stars444 forksMITlast commit Actively maintained
Every option on this page is open source and free to run on your own hardware, so you own the data and there is no subscription to cancel. 3 of 3 shipped a commit in the last six months. Licences in this list: MIT, Apache-2.0. In exchange you take on hosting, backups and updates yourself.
Browse everything in Model Serving & Inference.