Orpheus-FastAPI is a high-performance self-hosted text-to-speech server built with FastAPI.
It provides an OpenAI-compatible /v1/audio/speech endpoint together with a modern web interface for generating speech locally. The project is designed to work with external inference servers such as llama.cpp, LM Studio, and GPUStack, and focuses on fast GPU-accelerated synthesis with support for expressive tags, long-form generation, and multilingual voice models.
This is free and open source software.
Key Features
- Provides an OpenAI-compatible text-to-speech API endpoint.
- Includes a responsive web interface with waveform visualisation.
- Supports multilingual speech synthesis with multiple voices across several languages.
- Supports emotion tags for more expressive generated speech.
- Handles long-form audio generation through batching and crossfaded stitching.
- Can connect to external inference servers such as llama.cpp, LM Studio, and GPUStack.
- Offers Docker deployment options for GPU, ROCm, and CPU-based setups.
- Optimised for fast local inference on RTX-class GPUs.
Website: github.com/Lex-au/Orpheus-FastAPI
Support:
Developer: Alexander J.
License: Apache License 2.0
Orpheus-FastAPI is written in Python. Learn Python with our recommended free books and free tutorials.
Related Software
| Speech Tools | |
|---|---|
| Piper | Fast, local neural text to speech system |
| Bark | Transformer-based text-to-audio model. |
| sherpa-onnx | Speech-to-text and text-to-speech software |
| Coqui TTS | Offers pretrained models in more than 1,100 different languages |
| Dia | 1.6B parameter text to speech model |
| Tortoise | Multi-voice text-to-speech system trained with an emphasis on quality |
| Festival | General multi-lingual speech synthesis system |
| PraatSpeechAnalyser | Software for speech analysis and synthesis |
| Chatterbox | Family of text-to-speech models |
| Speech Note | Speech to Text, Text to Speech and Machine Translation |
| Mimic 3 | Lightweight Text to Speech engine |
| OrcaScreenReader | Scriptable screen reader |
| MeloTTS | High-quality multi-lingual text-to-speech library |
| Parler-TTS | Lightweight text-to-speech (TTS) model |
| Flite | Small, fast run time text to speech synthesis engine |
| RHVoice | Gives the visually impaired a synthesis voice with their screen reader |
| eSpeak NG | Continuation of the eSpeak project |
| eSpeak | Speech synthesizer using a formant synthesis method |
| Orpheus-TTS-FastAPI | High-performance self-hosted text-to-speech server |
| Gespeaker | GTK-based frontend for eSpeak |
| VoiceGen | Simple text-to-speech application |
| Glate | Google Translator and Text To Speech Service |
Read our verdict in the software roundup.
Explore our carefully curated directory of recommended free and open source software, covering every major software category.The directory forms part of our extensive collection of articles for Linux enthusiasts. It includes hundreds of detailed reviews, together with free and open source alternatives to proprietary software from companies such as Google, Microsoft, Apple, Adobe, IBM, Cisco, Oracle, and Autodesk. LinuxLinks also covers interesting projects worth exploring, Linux-compatible hardware, free programming books and tutorials, and much more. Know a useful free and open source Linux application that we haven’t covered? Tell us about it using our submission form. |


Please read our Comment Policy before commenting.