Preview These docs describe the preview API. It is not generally available yet, and endpoints and fields may change.

VoxVane docs

VoxVane is one API for text to speech, speech to text, turn detection and real-time voice agents. Text to speech is specified in the OpenAPI document. The other products are still preview placeholders.

Product guides

Guides for each product are on their way. Each link below opens a placeholder until its guide is published.

Quickstart

  1. Get an API key

    See API keys below. Export it so the examples can read it:

    export VOXVANE_API_KEY="your-key-here"
  2. Make some speech

    Send text and a voice id. The response is audio, not JSON.

    Preview API
    curl -X POST https://api.voxvane.com/v1/text-to-speech/aria \
      -H "Authorization: Bearer $VOXVANE_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{"text": "Hello from VoxVane.", "output_format": "mp3"}' \
      --output hello.mp3
  3. List voices

    GET /v1/voices returns the ids you can pass in that path. Usage for the key is GET /v1/usage.

API keys

During the private preview, API keys are issued by invitation. There is no self-serve sign-up yet; this page will link to it when there is.

  • Send the key as Authorization: Bearer … on every request.
  • Keep it on your server. Never put it in browser or mobile code.
  • Store it in the VOXVANE_API_KEY environment variable, not in source control.

SDKs

Python and JavaScript clients call this HTTP API and return audio bytes. They do not include a model or a pronunciation table. The curl example above is the same contract the clients use. Package install steps ship with the public SDK.

PythonPublic SDK
JavaScriptPublic SDK
HTTPToday

Errors and limits

Errors return JSON: { "error": { "code": "invalid_request", "message": "…" } } with a 4xx or 5xx status. A 429 response includes Retry-After. Speech is $0.012 per 1,000 characters. GET /v1/usage reports characters, audio seconds and the customer charge.