Pricing
Not listed
Official monthly price not confirmed
- API
- Homepage and About page promote a unified API and edge SDK with Get API key and playground access
Voice · Vendor site 2026-10-04
CassetteAI builds real-time generative audio models for music, sound effects, and upcoming TTS through one SDK or hosted API. It targets games, creator apps, and robotics with streaming latencies under 50 ms, 44.1 kHz stereo output, and pay-per-use hosted pricing.
Developers shipping interactive apps that need low-latency generative music, SFX, or speech inside real-time experiences.
About page lists hosted usage at $0.01 per SFX generation and $0.02 per output minute of music without published monthly subscription tiers.
CassetteAI promises music, SFX, and TTS through one API with streaming responses fast enough for games and live creator pipelines.
What it can do
Unknown is different from unavailable. Each fact carries its own evidence.
| Capability | Value | Evidence | Checked |
|---|---|---|---|
| Latency | Homepage cites first-sample latency around 23 ms and full tracks rendered in under 10 seconds for 3-minute audio. | Facts sourced | 2026-10-04 |
| Modalities | Marketing describes Music, SFX, and upcoming TTS engines under one SDK. | Facts sourced | 2026-10-04 |
| Usage pricing | About page states pay-per-use pricing at $0.01 per SFX generation and $0.02 per music output minute. | Facts sourced | 2026-10-04 |
Understand the total cost
Pricing
Not listed
Official monthly price not confirmed
Keyframe Labs builds photoreal interactive AI avatars with ultra-low latency, emotion control, and image-to-avatar creation. Developers deploy sessions through SDKs and no-code components, choosing LLM and voice providers while routing traffic across multi-region infrastructure.
Explore toolLalals is a browser-based suite of AI music and audio tools for creators. It bundles voice changing, text-to-speech, stem splitting, co-production, mastering, sound effects, voice cloning, transcription, and related utilities in one workspace marketed to musicians and producers.
Explore toolAqua Voice provides fast voice dictation that turns speech into clear text across Mac, Windows, and iPhone apps, including AI prompts and long-form writing. It ships Avalon transcription models with Free, Pro, Max, and Business tiers.
Explore toolThe practical questions
CassetteAI builds real-time generative audio models for music, sound effects, and upcoming TTS through one SDK or hosted API. It targets games, creator apps, and robotics with streaming latencies under 50 ms, 44.1 kHz stereo output, and pay-per-use hosted pricing.
Not confirmed; About page cites pay-per-use API rates rather than a free production tier. This record does not confirm an ongoing free plan.
A listed monthly price has not been confirmed. See the plan cards for entitlements, billing commitments and seat minimums.
Homepage and About page promote a unified API and edge SDK with Get API key and playground access. API access and subscription access may have different terms; consult the linked sources.
About page lists hosted usage at $0.01 per SFX generation and $0.02 per output minute of music without published monthly subscription tiers.
No retained pricing changes yet. A current price alone does not establish a historical trend.
Reviewed vendor source
Read original source ↗Facts apply to the named version and check date. Send a sourced correction if something changed.