Technology · Agentic → Voice Agents · Open Source · wiki:deep
Piper is a fast, local neural text-to-speech engine. Active development lives at OHF-Voice/piper1-gpl (Open Home Foundation; pip install piper-tts, GPL-3.0). It embeds espeak-ng for phonemization, ships voice models you run on-device, and exposes CLI, Python, HTTP, and C/C++ APIs. It is the speak substrate under voice agents — not a full STT→LLM→TTS product (that job stays with Pipecat / LiveKit Agents / Vocode).
The older rhasspy/piper README points development to the OHF repo; treat OHF as primary and the Rhasspy tree as legacy redirect.
Whisper covers listen. Without a local TTS pack, “voice round-trip” only exists as a stage name inside pipeline frameworks. Piper is the research default when spoken replies must stay on-device (privacy, offline Studio, Home Assistant-class boxes) and still bind to the same gate rules as text agents: speech is channel chrome; approve / spend / mutate stay elsewhere.
Text enters Piper; espeak-ng phonemizes; a neural voice model synthesizes audio. Operators pick a voice ONNX (or trained voice), call the CLI / Python API / HTTP server, and stream or write WAV/raw audio to a transport that Pipecat or LiveKit Agents already owns.
piper-tts (or build libpiper). Related: whisper · pipecat · livekit-agents · vocode · topics/03-channels
piper1-gpl is GPL-3.0 — embedding in proprietary closed agents has copyleft implications; confirm counsel before shipping. Claims below are backed by science sources on disk.
Product / local TTS framing
OHF-Voice/piper1-gpl README · STRONG
“A fast and local neural text-to-speech engine that embeds espeak-ng for phonemization.”
Legacy redirect (Rhasspy → OHF)
rhasspy/piper redirect · STRONG
“Development has moved: https://github.com/OHF-Voice/piper1-gpl”
9 tags · 8 out · 9 in · 3 artifacts · 1 gaps · 0 corpus docs