ttsMP3.com is a browser-based text-to-speech tool that converts typed text into natural-sounding speech across dozens of languages and accents. Users can listen online or download MP3 files for e-learning, presentations, YouTube, and accessibility projects.
TTS WebUI is a free Gradio-based web interface for text-to-speech, audio, and music generation. It supports dozens of open models including Bark, MusicGen, and Tortoise with flexible local installation options and ongoing GitHub development.
TurboDoc extracts structured data from invoices, receipts, and bank statements using AI document processing. Users can upload PDFs or images, connect Gmail or Outlook inboxes, export CSV or XLS, and automate accounts-payable workflows on tiered monthly plans.
Twilio is a cloud communications platform providing APIs and SDKs for programmable SMS, voice, email, video, authentication, and customer engagement workflows. The homepage positions Twilio as an agentic-era engagement stack connecting channels, customer context, and conversational AI building blocks.
Uberduck provides AI vocals, text-to-speech, voice cloning, and related media tools for creators and marketers. Paid tiers add commercial licenses, API access, image generation, and monthly credit pools for synthetic voice content.
Ultimate Vocal Remover (UVR) is a free, open-source desktop application for separating vocals and other stems from songs on Windows, macOS, and Linux. The official site distributes UVR5 downloads, encourages donations, and describes the tool as the best vocal remover available at no charge.
UneeQ (digitalhumans.com) builds AI digital humans and an Immersive Training Platform for sales, customer service, leadership, and education roleplay with realistic conversational simulations. The site also covers Digital Human OS components, Synanim animation, LLM orchestration, kiosks, and XD Studios creative services for branded avatars.
Unholy.ai is a lightweight web app that analyzes song lyrics for themes it labels as unholy, such as erotica, blasphemy, adultery, or explicit content.
Unmixr AI is a multilingual voice content studio for voiceovers, dubbing, transcription, voice cloning, and turning slides or scripts into narrated video. One credit balance covers voice, video, and dubbing workflows with API and automation options for production teams.
Unreal Speech provides a low-latency text-to-speech API powered by Kokoro-82M, streaming audio in about 300ms with optional per-word timestamps. Pricing tiers scale by monthly character allotments, and a browser studio lets developers preview voices before integrating.
Uplift AI builds Urdu text-to-speech, voice APIs, and AI calling agents for Pakistan, targeting appointment booking, lead qualification, and routine customer service. The homepage highlights 100+ realistic Urdu voices, a Studio for speech, developer APIs, and Calling agents with a Start Free entry point.
Vapify is a white-label agency platform for reselling and managing Vapi.ai voice agents under your brand. Agency pricing sells sub-accounts, markup controls, and GoHighLevel integrations for voice AI agencies.
Vatis Tech transcribes audio and video to text with 95%+ accuracy across 50+ languages, offering both a team transcription platform and developer speech APIs. Platform plans bill in euros while the site also publishes async API rates.
Vedika is an astrology intelligence API for birth charts, natural-language answers, voice guidance, compatibility, PDF reports, MCP tools, and SDKs spanning Vedic, Western, and KP systems. The homepage demo shows computed ephemeris context, evidence paths, and subscription wallet plans for production integrations.
Vexo makes the Vexo Band, an always-on AI bracelet that listens to conversations, remembers context, and proactively helps with everyday tasks. The wearable connects to a phone over Bluetooth LE and includes onboard memory and health sensors.
VideoDubber provides AI video translation, dubbing, voice cloning, and text-to-speech across 150+ languages with minute-based plans. Subscriptions bundle translation and cloning minutes, watermark-free exports, and premium voice or lipsync features on higher tiers.
VideoLingo localizes videos with AI translation, bilingual subtitles, dubbing, and optional voice cloning through minute-based subscriptions. Plans scale supported input languages, output resolution, and API access for multilingual content pipelines.
Vishaya AI is a prototype platform for generating multilingual AI-powered courses with structured lessons and audio. The official site notes the project is no longer actively maintained while still demonstrating course generation flows.
Vision Agents is an open-source Python framework for low-latency voice and video AI agents. Developers can mix LLM, speech, and vision models from many providers and deploy realtime agents on Stream’s global edge network.
VlogMusic.io resolves to AISong, an online AI song generator that creates royalty-free music from prompts or lyrics with genre and voice selection. The platform also offers related video, image, stem-splitter, and remix tools on the same site.
Vocali.se is a free online service that separates vocals and instrumentals from uploaded songs to create karaoke tracks using a Demucs-powered AI engine. Processing completes in under three minutes without requiring software installs or account registration.
Vocality Health provides AI voice technology for healthcare, including clinical interpretation and intelligent voice agents for providers. The public site focuses on voice-driven patient and clinical workflows.
Not listed
API unknown
Checked Oct 4, 2026
Showing 720 of 804 tools
Plan facts carry individual sources and check dates. Unknown values are excluded from required filters.