Skip to content

ReadSpeaker speechServer

On-premise text-to-speech server for your own network

Run ReadSpeaker's text-to-speech engine on your own Windows Server or Linux machines. Your applications request speech over TCP/IP, the command-line tool, the REST API, or the C, Java, and .NET SDKs.

  • 150+ AI voices
  • 50+ languages
  • Self-hosted on Windows Server or Linux

INSIDE speechServer

A text-to-speech server you host yourself

ReadSpeaker speechServer is a client-server, network-based text-to-speech server. You install it on your own infrastructure, and your applications request speech synthesis over the network. It suits teams that need text-to-speech inside their own environment, from virtual assistants and notification platforms to educational applications and other server-based systems that generate audio on demand.

HOW IT CONNECTS

Four ways to integrate your application

Your application talks to speechServer the way that fits your stack. Connect over a direct network protocol, the command-line tool, a web API, or a client SDK in your own language.

TCP/IP

TCP/IP protocol

Connect your application directly to the speechServer engine over a TCP/IP communication protocol.

CLI

Command-line tool

Drive speech synthesis from scripts and server processes with the included command-line tool.

REST

REST API

Call speechServer over a REST API when it runs behind a CGI-enabled web server such as Apache.

SDK

C, Java, and .NET SDKs

Integrate with client-side SDKs for C, Java, and .NET that handle communication between your application and speechServer.

# direct socket connection
connect speechserver:5555
synth "Welcome"

Code samples are illustrative. See the user documentation for the exact API syntax.

RUN IT YOUR WAY

Runs on Windows Server and Linux

Install speechServer on the platform your operations already run.

Windows Server

Install speechServer on your Windows Server machines, alongside the services your operations already run.

Linux

Install speechServer on the Linux distributions your infrastructure is built on, from bare metal to virtual machines.

AUDIO OUTPUT

Stream the audio format your system needs

speechServer streams audio back to your application in the format your pipeline expects, from telephony-grade encodings to standard media formats.

PCM

16-bit linear

PCM Wave

16-bit linear & variants

A-law

8-bit PCM

μ-law

8-bit PCM

ADPCM

4-bit Dialogic

mp3

via the LAME package

OGG

open container

VOICES

150+ AI voices in 50+ languages

speechServer uses ReadSpeaker's catalog of more than 150 AI voices across over 50 languages, with new voices and languages added regularly. Each licensed language includes a user dictionary so you can fine-tune pronunciation for domain-specific terms.

Global coverage

150+ AI voices in 50+ languages

Explore all voices

Available languages

FINE CONTROL

Control how every word is spoken

Shape the delivery of your text down to the word, with the controls your content already carries.

SSML voice and language switch

Switch to another voice, or another language, mid-text using SSML instructions in your input. Adjust speaking rate, pitch, and volume as needed.

User dictionary and IPA

Define custom pronunciations for words and patterns with per-language user dictionaries and IPA transcription input, so domain terms read correctly.

Audio clip insertion

Insert references to audio clips in your input text. speechServer places the referenced audio at the right position in the streamed output.

Mark information

Receive mark information alongside the audio for event triggers at text positions, device and interface synchronization, and text highlighting aligned with playback.

DEPLOYMENT

Your text-to-speech stays in your network

speechServer runs entirely on your own servers, so text and audio stay inside your network.

Installed, licensed and operated by you

speechServer runs entirely on your own servers, so text and audio stay inside your network. It installs with the speechServer installer, and any licensed ReadSpeaker voice can be added. Use is governed by a license file that sets your voices, concurrent ports, synthesis speed, number of servers, and term.

  • Installation and implementation support is included from the ReadSpeaker Support Team.

speechServer vs OPEN-SOURCE

A supported alternative to do-it-yourself text-to-speech

Open-source projects can run text-to-speech locally, but you maintain the engine, the voices, and the integration yourself. speechServer gives you a commercial, supported server with a maintained voice catalog.

ReadSpeaker speechServer compared with open-source do-it-yourself text-to-speech
ReadSpeaker speechServerOpen-source do-it-yourself
TypeCommercial product, supportedFree open-source, community-maintained
Voice catalog150+ AI voices, 50+ languagesBring-your-own, varies
IntegrationTCP/IP, command line, REST, C/Java/.NET SDKsDo-it-yourself wiring
SupportReadSpeaker Support Team, installation and implementationCommunity
LicensingLicense file with defined termsProject license, no SLA
Voice updatesAdded regularlyDo-it-yourself

FAQ

Questions about speechServer

  • ReadSpeaker speechServer is an on-premise, network-based text-to-speech server. You host it on your own Windows Server or Linux machines, and your applications request speech synthesis over the network. It provides the ReadSpeaker text-to-speech engine, a server API, and tools as a complete server-side runtime.

  • speechServer is self-hosted, not cloud-based: it runs on your own Windows Server or Linux machines, and your text and the audio it generates stay inside your own network rather than passing through a third-party service. If you prefer a managed cloud API instead, ReadSpeaker speechCloud API is the hosted option.

  • Client applications connect to speechServer over a TCP/IP communication protocol, a command-line tool, or a REST API when it runs behind a CGI-enabled web server such as Apache. Client-side SDKs for C, Java, and .NET are also provided.

  • speechServer provides client-side SDKs for C, Java, and .NET that handle the communication between your application and the server, so you integrate speech synthesis in your own code rather than wiring the protocol by hand. Supported development languages are C and C++, Java, and C# / .NET on Windows.

  • speechServer runs on Windows Server and on Linux. Ask our team for the current list of supported and tested versions for your environment.

  • speechServer outputs 16-bit linear PCM, PCM Wave variants, 8-bit A-law and μ-law PCM, 4-bit Dialogic ADPCM, mp3 through the LAME package, and OGG. This range covers both telephony-grade and standard media playback.

  • speechServer supports ReadSpeaker's catalog of more than 150 AI voices across over 50 languages. New voices and languages are added regularly. Each licensed language includes a user dictionary and IPA support for custom pronunciations.

  • speechServer is a commercial, supported product with a maintained catalog of 150+ AI voices, four integration paths, and installation support from the ReadSpeaker Support Team. Open-source projects can run locally, but you maintain the engine, voices, and integration yourself, without an SLA.

  • speechServer integrates directly through TCP/IP, command line, REST, or C, Java, and .NET SDKs, for general server-based applications. ReadSpeaker speechServer MRCP uses the MRCP protocol for standards-based telephony and IVR platforms.

  • speechServer use is governed by a license file. It sets the licensed voices, the number of concurrent text-to-speech ports, the permitted synthesis speed, the number of servers covered, and the license term. Concurrency is limited only by your licensed ports and available system resources.

Ready to host your own text-to-speech server?

Tell us about your environment and the voices you need. Our team will help you size speechServer, choose your integration path, and plan the deployment on your Windows Server or Linux infrastructure.

Find your ReadSpeaker solution

Opening the conversation…

Search

    Searching…

    No results for

    Try another wording, or start from one of these.

    Search is not available right now.

    You can still reach the pages below.