Beta · turn-based

Build a voice agent in five minutes.

Describe your agent in plain words, in a chat. Pick a voice, or make one that is yours: record a paragraph or describe it and tune it with meters. When you like the preview, press Generate API. We run the model, keep the memory, and speak the answer sentence by sentence, using instant library phrases when they match.

1

Describe it

Chat with the builder. It fills an editable configuration: personality, language, voice, knowledge, FAQs, actions and memory. Five agents per account.

2

Generate API

Each publish is a numbered version at one endpoint, with a key that works only for that agent and short-lived tokens for browsers.

3

Play the reply

You get the text and one audio link per sentence. Play them in order, and report what was actually played if the caller interrupts.

One endpoint.

curl -X POST https://speakvora.com/api/v1/agents/ag_123/turn \
  -H "x-api-key: $SPEAKVORA_KEY" -H "content-type: application/json" \
  -d '{"session":"visit-42","text":"Can I move my appointment to Thursday?"}'

{ "session": "visit-42",
  "reply": "Of course. What day and time is your current appointment?",
  "audio_url": "https://speakvora.com/live/.../....mp3",
  "source": "live", "ms": 5200 }

Honest limits of the beta

  • Fast turns, not a live phone call: you get the reply and its audio when the turn completes (about 2 to 4 seconds with your own voice). WebSocket streaming is not available yet.
  • Speech in is a file upload we transcribe, or text you send.
  • Your own voice speaks English well, French and German as beta. Other languages use catalog voices.
  • No phone numbers. Short replies by design.
$10per 1,000 turns, early access
first 200 turns free