Curated, distinctive voices. Custom voice cloning available on Business+ plans.
Native fluency, not translation. Code-switching across languages mid-sentence.
Real-time TTS for voice agents. Faster than human reaction time.
Tone, pace, emotion. Direct the delivery, not just the words.
WebSocket streaming for real-time applications. PCM, MP3, Opus outputs.
Speaker diarization, word timestamps, custom vocabularies. 95%+ accuracy.
Token-efficient, with volume discounts and burn incentives.
47 voices, 31 languages
Emotion + style control
Diarization + timestamps
Sub-200ms voice agents
Pricing is illustrative. Final rates are governed by on-chain parameters and may vary based on network state.
Tier-based access. Higher stakes unlock better economics and more capacity.
| Tier | Required stake | Access |
|---|---|---|
| Builder | 50M QUBIC | Voice TTS/STT enabled |
| Startup | 150M QUBIC | Realtime API, custom voices |
| Business | 500M QUBIC | Voice cloning, telephony |
| Enterprise | Custom | On-prem, custom voices |
Drop-in compatible with the OpenAI SDK.
from aigarth import Aigarth
client = Aigarth(api_key="sk-aigarth-...")
audio = client.voice.synthesize(
text="Welcome to Aigarth. The future of compute grows on participation.",
voice="atlas",
language="en-US",
emotion="warm",
)
audio.save("welcome.mp3")Everything in the standard tier, plus the things enterprises need.