SnapOtter is self-hosted file-processing infrastructure for converting, compressing, OCR-ing, transcribing, and running local AI on images, video, audio, PDFs, and documents. Teams use a web UI, REST API, and pipelines on hardware they control so sensitive files never leave the network.
Sona8 provides employee listening by voice across an entire workforce. The platform interviews employees in their own language by voice and surfaces leadership-facing themes plus the underlying sentences behind those themes.
Soundful generates royalty-free music and loops with producer-aided AI templates for creators, brands, and enterprises. Users can start from genres, download tracks, and upgrade for more styles, stems, and commercial licenses.
SOUNDRAW is an in-house-trained AI music generator that produces royalty-free beats with mixer controls for length, intensity, and instrumentation. It targets creators who need copyright-safe background music and optional artist plans for distribution features.
Soundry AI builds generative music tools aimed at empowering human creativity, including Groove for remixing songs into new styles and Sample Planet for mixdown-ready samples. The company positions its products for DJs, producers, and TikTok creators.
SoundTools hosts 35+ browser-based audio and AI voice utilities—including text-to-speech, vocal removal, stem splitting, voice changing, effects, and converters—with no signup on the homepage. Processing runs in the browser so files stay on the user device.
Soundwise.ai transcribes audio and video to text in the browser using Whisper-based models, supporting many common media formats. It offers unlimited local transcription on the free tier and a Pro tier for faster cloud processing.
SpaceGen markets AI coworkers—persona-driven agents for research, writing, design, editing, scouting, web development, and QA—that can join a customer team or run outsourced workflows managed by SpaceGen with human backup.
Speak4Me is a mobile text-to-speech app that reads PDFs, websites, and scanned documents aloud with natural voices and speed controls. It also offers ChatWithMe Q&A on files and a free-for-schools education program.
SpeakNotes records or uploads meetings, lectures, podcasts, and YouTube links to produce transcripts and AI summaries in 50+ languages. It adds a meeting bot for Google Meet, Zoom, and Teams plus sharing to Notion or Slack.
Speaktor converts text into natural speech and supports audiobooks, podcasts, voice-over video, and WAV exports across 55+ languages. Subscriptions sell Lite, Pro, and Team seats with monthly minute allowances and dubbing on higher tiers.
SpeechBrain is an open-source PyTorch toolkit for speech and text technologies, covering ASR, enhancement, separation, TTS, speaker recognition, and chatbot pipelines. It ships recipes, Hugging Face checkpoints, and documentation for researchers.
SpeechCraftPro sells AI-generated speeches for events such as weddings, graduations, keynotes, and eulogies based on user prompts about occasion and talking points. Customers purchase credits, sign in, and receive drafts tailored to speech type.
Speechify is a voice AI assistant that reads books, PDFs, and web pages aloud with natural voices, plus voice typing, AI chat, summaries, and podcast-style audio creation across browser, mobile, and desktop apps. Premium unlocks higher-quality voices, faster playback, and integrations.
SpeechPulse is a desktop dictation and transcription app for Windows and macOS that uses Whisper-based speech recognition with optional offline mode, 99-language transcription, AI punctuation, and templates that can call external LLM APIs for cleanup.
SpeechReader converts pasted text, PDFs, and images into natural text-to-speech audio with thousands of voices across 60+ languages. Free Demo, Basic, and Premium plans differ by daily credits, downloads, and access to premium or studio voices.
SpeedyAudios is a WhatsApp-based service that transcribes forwarded voice notes into text in seconds. Users pin the chat, forward any audio, and receive transcripts or summaries without installing a separate app or creating an account.
Speko is a router for voice AI that benchmarks speech and language models across languages and exposes a hosted STT, LLM, and TTS data plane. Pricing lists Speko infrastructure at $0.09 per minute all-in or provider rates plus 5% when routed externally, with $10 signup credit.
Spreaker Create is Spreaker’s free online podcast editor and creation studio for recording, editing, and publishing shows. Spreaker also offers hosted podcast plans with distribution to major directories, monetization, and analytics through separate hosting tiers.
Stabbl is a French invoicing web app that generates compliant PDF invoices in seconds for freelancers and auto-entrepreneurs. It auto-fills company legal data via Pappers, validates mandatory invoice fields, and automates VAT handling.
Stable Audio is Stability AI's music and sound generation product with a web app and DAW plugin. The official pricing page advertises monthly subscriptions and licensing tiers compared by credits, generation length, uploads, and commercial rights.
StableAudio.net is a third-party AI music creation web app that turns text descriptions into songs with optional vocals or instrumentals. The homepage sells monthly Basic, Standard, and Pro subscriptions plus signup credits for new users.
Stable Audio Open Online hosts a browser demo of Stability AI's open text-to-audio model, generating stereo samples up to 47 seconds from text prompts. The page explains model capabilities and offers a free online generator interface.
Staccato is an AI music copilot with a DAW plugin for MIDI creation, extension, rewriting, and accompaniment from text prompts. Paid subscriptions unlock unlimited monthly credits, commercial use, and optional MIDI or audio uploads on Pro.
$11.99/mo · starting plan
API unknown
Checked Oct 4, 2026
Showing 648 of 804 tools
Plan facts carry individual sources and check dates. Unknown values are excluded from required filters.