The Typecast MCP server is an implementation of the Model Context Protocol that connects your AI agent and assistants like Claude, Cursor, etc directly to your Typecast account. It provides structured and secure access to speech generation, voice discovery, voice details, and subscription usage, so your agent can create expressive audio, compose speech sequences, generate timestamps, recommend voices, and check account usage on your behalf.
- Text to speech generation: Have your agent turn text into complete WAV or MP3 audio using an available Typecast voice.
- Speech composition: Direct your agent to create one audio file from an ordered mix of spoken passages and pauses.
- Timed speech alignment: Instruct your agent to generate speech with word timestamps, character timestamps, or both for captions and synchronized experiences.
- Voice discovery and recommendations: Let the agent list available built-in and custom voices, filter them by traits and use cases, recommend matches from a natural-language description, and review voice details, models, and emotions.
- Plan and usage checks: Have your agent check your Typecast plan, credit usage, concurrency limit, and custom voice allowance.