Kaldi is an open source speech recognition toolkit maintained by the research community. It provides training and decoding recipes for automatic speech recognition experiments, with code hosted on GitHub and documentation on kaldi-asr.org.
Kalinda is an AI intake specialist for personal injury law firms. It answers and returns calls, qualifies leads, and drives retainers from first contact through signed intake so firms convert more of the demand they already generate.
Karaoke Eternal is a self-hosted karaoke party system with browser-based mobile queuing and a browser player supporting MP3+G, MP4 video, and WebGL visualizations. Guests join via QR codes while the server runs on your own hardware without ads or telemetry.
Katalog is an audio-first read-it-later app that saves article links, optimizes them for listening, and narrates them with enhanced AI voices. Paid tiers add Conversational Reading voice Q&A while metering saved articles and generated audio each month.
Kea AI provides voice and text AI agents that answer restaurant phones, take orders, and handle guest questions around the clock. The homepage includes an interactive demo, FAQ, and claims average monthly savings from recovered missed order revenue.
Keyframe Labs builds photoreal interactive AI avatars with ultra-low latency, emotion control, and image-to-avatar creation. Developers deploy sessions through SDKs and no-code components, choosing LLM and voice providers while routing traffic across multi-region infrastructure.
KIN is a voice-first shared family calendar app that keeps households aligned on schedules through voice-powered planning. Marketing emphasizes family scheduling assistance rather than generic task lists.
Kits AI provides studio-quality AI audio tools for musicians, including custom voice models, vocal conversions, cloning, and royalty-free outputs. Creators train or pick voices, convert recordings, and download stems from browser or desktop apps.
klink.cloud is an omnichannel customer conversation platform combining inbox, voice, chat, WhatsApp, LINE, and case management. Its Kai AI agents handle chat and voice using your content, with outcome-based pricing on resolved conversations.
koolio.ai is an AI audio workspace for podcasters that combines transcription, music, SFX, voice, and cover art in one credit pool. Studio-grade processing and automation help creators polish raw recordings into publish-ready episodes quickly.
Krisp delivers AI meeting assistance with noise cancellation, transcription, recordings, and summaries, alongside separate call-center and developer voice products. Meeting plans are sold per user with a trial before purchase.
KugelAudio provides European-hosted voice AI and text-to-speech APIs with natural voices in 39 languages, GDPR compliance, and EU Sovereign Cloud included on every plan. The vendor highlights streaming latency, voice cloning, and conversational and story presets.
Kveeky is an AI text-to-speech and voiceover platform with 700+ voices, voice cloning, localization, and credit-based monthly plans for video, e-learning, ads, and podcasts. Creators generate studio-style narration online without recording talent for each project.
Ladder is an AI sales platform for home services enterprises. Reps record customer conversations while Ladder coaches closers, automates follow-ups and busywork, drives reviews and referrals, and scales proven talk tracks with outcome-based pricing tied to sold projects.
LALAL.AI separates vocals, instruments, and stems from audio and video using AI, with voice changer, cleaner, and cloner tools. Paid tiers add fast-queue minutes, larger uploads, and API access on Pro.
Lalals is a browser-based suite of AI music and audio tools for creators. It bundles voice changing, text-to-speech, stem splitting, co-production, mastering, sound effects, voice cloning, transcription, and related utilities in one workspace marketed to musicians and producers.
Lami.ai is an AI music generator that turns text descriptions or lyrics into original songs in minutes. The site also offers vocal removal, stem splitting, AI covers with voice models, lyrics generation, and exports in MP3, WAV, or MP4 formats for creators who want royalty-free music workflows.
Lamucal delivers AI-enhanced chords, lyrics, tabs, and melody for songs with real-time playback sync across phone, tablet, and PC. Features include vocal removal, smart transpose, loops, speed control, MIDI tabs, and practice tools for guitar and piano players.
Lance provides collaborative AI agents for hotels to streamline operations and guest service. Offerings include Agent Builder, a 24/7 Voice Agent, AI Housekeeping, and AI Labor Scheduling, with customer stories citing faster room turns and fewer missed guest calls.
Leoline is a voice-first storytelling companion for children that turns spoken prompts into unique, kid-safe adventures. Parents control content while kids chat aloud instead of typing to hear fresh stories each session.
Leon is an open-source personal assistant you can self-host on your own server. Built with Node.js and Python, it listens for voice or text requests, runs skills through a modular core/brain architecture, and keeps data under your control instead of a third-party cloud assistant.
Letterly turns speech into polished text for messages, notes, emails, journals, and meeting recaps across iOS, Android, Mac, Windows, and web. It supports dictation inside other apps, long recordings up to 90 minutes, and transcripts with speaker splits.
LibreTime is an open-source radio broadcast and automation platform for community and internet stations. It imports music, sweepers, and full programs, schedules shows and podcasts, and plays out audio to a broadcast console, transmitter, or cloud stream.
Free
API unknown
Checked Oct 4, 2026
Showing 456 of 804 tools
Plan facts carry individual sources and check dates. Unknown values are excluded from required filters.