Pricing
Not listed
Official monthly price not confirmed
- API
- Velma REST and WebSocket APIs publish per-hour rates for transcribe, detect, triage, and redact models
Voice intelligence · Vendor site 2026-10-04
Modulate builds audio-native voice intelligence with its Velma model suite for transcription, deepfake detection, emotion and language signals, redaction, and conversation triage. Developers integrate batch REST and real-time WebSocket APIs, while ToxMod targets voice safety in gaming and social communities.
Trust-and-safety, contact-center, and fraud teams that need speech understanding beyond text LLMs on live audio.
Pricing is usage-based per hour of audio processed; enterprise platform and ToxMod plans require sales conversations.
Modulate combines research-grade voice models with enterprise platform and ToxMod moderation offerings, emphasizing audio-native understanding for safety, analytics, and developer APIs.
What it can do
Unknown is different from unavailable. Each fact carries its own evidence.
| Capability | Value | Evidence | Checked |
|---|---|---|---|
| Model breadth | API catalog spans transcription, deepfake detection, emotion, accent, language, music detection, triage, and PII redaction. | Facts sourced | 2026-10-04 |
| English fast STT | English Fast Speech-to-Text lists $0.025 per hour batch and $0.05 per hour streaming. | Facts sourced | 2026-10-04 |
| Deepfake detection | Deepfake Detection is priced at $0.25 per hour for batch and streaming with sub-three-second verdicts marketed. | Facts sourced | 2026-10-04 |
| Platform positioning | Homepage highlights an Ensemble Listening Model architecture for conversational understanding in voice. | Facts sourced | 2026-10-04 |
Understand the total cost
Pricing
Not listed
Official monthly price not confirmed
TTS.ai is a multi-model text-to-speech platform with 36+ models, 314+ voices, voice cloning, speech-to-text, music, and marketplace listings. The homepage offers free generation without an account while paid plans add commercial licenses, API access, and longer downloads.
Explore toolai-coustics provides real-time audio intelligence SDKs that clean speech, remove background voices, and score audio quality for voice-AI pipelines. Voice Focus and related models target sub-30ms enhancement for production voice agents, with minute-based SaaS plans and enterprise SLAs.
Explore toolAltered provides professional AI voice changing software for real-time calls and post-production workflows. Altered Studio focuses on speech-to-speech voice morphing, cloning, and editing, while Real-Time Pro targets gamers, call centers, and voice restoration use cases.
Explore toolThe practical questions
Modulate builds audio-native voice intelligence with its Velma model suite for transcription, deepfake detection, emotion and language signals, redaction, and conversation triage. Developers integrate batch REST and real-time WebSocket APIs, while ToxMod targets voice safety in gaming and social communities.
Self-serve API signup advertised with pay-as-you-go hourly audio pricing. This record does not confirm an ongoing free plan.
A listed monthly price has not been confirmed. See the plan cards for entitlements, billing commitments and seat minimums.
Velma REST and WebSocket APIs publish per-hour rates for transcribe, detect, triage, and redact models. API access and subscription access may have different terms; consult the linked sources.
Pricing is usage-based per hour of audio processed; enterprise platform and ToxMod plans require sales conversations.
No retained pricing changes yet. A current price alone does not establish a historical trend.
Reviewed vendor source
Read original source ↗Facts apply to the named version and check date. Send a sourced correction if something changed.