Skip to content

Products

PyAI products: Omni, Hear, Speak, Agents, and Cast

Omni is one speech-to-speech model for phone agents at human conversational pace, about 390 ms median voice-to-voice in-region. Hear and Speak are the OpenAI-compatible STT and TTS APIs underneath. Start with a sandbox key that never bills.

ModelWhat it doesEndpointPriceStatus
OmniThe definitive all-in-one AI voice agent model. Hybrid speech-to-speech, fused LLM brain, emotion-aware voices, and tool calling.wss://api.pyai.com/v1/omni$0.05/minLive
AgentsLaunch and manage AI voice agents without code.Console + wss://api.pyai.com/v1/omni$0.08/minLive Beta
CastCinematic, emotional long-form TTS for podcasts, narration, and audiobooks.POST /v1/cast/render_jobs$0.02/minLive
TraceA scorecard on the call you already ran. Not a second QA vendor.GET /v1/trace/interactions$0.12/callBeta
HearHear - speech-to-text for real phone audio, with finals you can ship.POST /v1/audio/transcriptions$0.001/minLive
SpeakSpeak - text-to-speech (TTS) that starts in milliseconds.POST /v1/audio/speech$0.04/minLive
TelephonyOne agent stack across the PyAI network.POST /v1/telephony/numbers$0.01/minLive
AMDKnow who answered. Even when it's an AI screening the call.wss://api.pyai.com/v1/amd/stream$0.004/callLive
RecapCall notes the moment you hang up. No bot to invite.GET /v1/recap/calls$0.02/minAdd-on
Add-ons:Trace ($0.12/call)Recap ($0.02/min)KB Context ($0.003/min) - Full pricing

Get an API key. Make a phone agent.

Sandbox keys never bill. Fund a live key when you go to production.