Text-to-Speech
Loly 3.5
Turns written intent into audible presence. Generate a finished file, return bytes directly, stream progressive SSE events or deliver PCM16 frame by frame over WebSocket.
- Batch and binary output
- SSE and WebSocket streaming
- Custom voice identity
- Language, pace and synthesis controls
At a glance
- Modality
- Text to speech
- Transports
- REST, SSE, WebSocket
- Output formats
- WAV, MP3, PCM16
- Stream frames
- PCM16 mono, 24,000 Hz
- Languages
- 646 plus auto
- Text per request
- 5,000 characters
