Speech in both directions: text-to-speech and voice cloning, speech-to-text and transcription, real-time voice agents, dubbing, and audio intelligence. Every listing is a real product with a live website — dead tools get removed, not archived.
209 tools listed · 163 added in the last 30 days · newest: Sensory, Talon, Cerence
Cumulative bids, uncapped. Rebid pays only the difference. Season freezes from now.
Ad disclosure: this ranking is paid advertising placement, ordered solely by cumulative captured bids (ties: earlier capture, then lowest intent id) — methodology.
| # | Tool | Season total | Min to take this spot | Clicks | CTR | CPC |
|---|
| 3Play Media AI plus human captioning, transcription, and audio description | First bid €3 → #1 |
| ACE Studio AI singing voice generation with note-level editing | First bid €3 → #1 |
| AIVA AI music composition assistant | First bid €3 → #1 |
| ASAPP Generative AI for contact centers | First bid €3 → #1 |
| Abridge AI medical conversation documentation | First bid €3 → #1 |
| Acapela Group Text-to-speech voices and custom voice services | First bid €3 → #1 |
| Adobe Podcast Browser-based AI speech enhancement and recording | First bid €3 → #1 |
| Altered Professional AI voice changer | First bid €3 → #1 |
| Amazon Polly AWS text-to-speech service offering neural and generative voices via API | First bid €3 → #1 |
| Ambience Healthcare AI operating system for clinical documentation | First bid €3 → #1 |
| Amira Learning AI reading tutor that listens to students | First bid €3 → #1 |
| Amphion Open-source toolkit for audio, music and speech generation | First bid €3 → #1 |
| Aqua Voice Voice-first AI text editor | First bid €3 → #1 |
| AssemblyAI Speech AI models via API | First bid €3 → #1 |
| AudioPen Converts rambling voice notes into clear written text | First bid €3 → #1 |
| AudioShake AI sound separation for industry | First bid €3 → #1 |
| Augmedix Ambient AI medical documentation | First bid €3 → #1 |
| Auphonic Automatic audio post production | First bid €3 → #1 |
| Ava Real-time captions for deaf and hard-of-hearing people in conversations | First bid €3 → #1 |
| Avoca AI phone platform that answers, books, and coaches calls for HVAC and home services | First bid €3 → #1 |
| Balto Real-time guidance for contact centers | First bid €3 → #1 |
| BandLab Free social music studio with SongStarter AI idea generation | First bid €3 → #1 |
| Beatoven.ai Royalty-free AI music generation | First bid €3 → #1 |
| Bland Enterprise AI phone calls | First bid €3 → #1 |
| Boomy Make generative music instantly | First bid €3 → #1 |
| Camb AI AI dubbing and speech translation | First bid €3 → #1 |
| Cartesia Real-time voice intelligence | First bid €3 → #1 |
| Castmagic Turns podcast and meeting audio into show notes and written content | First bid €3 → #1 |
| Cephable Hands-free device control via voice, face, and head movements | First bid €3 → #1 |
| Cerence Voice assistant and speech technology for cars | First bid €3 → #1 |
| Cleanvoice AI audio editing for podcasts | First bid €3 → #1 |
| Cognigy Enterprise conversational and voice AI | First bid €3 → #1 |
| Commure Ambient AI scribe and healthcare automation | First bid €3 → #1 |
| ConverseNow Voice AI that takes restaurant phone and drive-thru orders | First bid €3 → #1 |
| Corti AI infrastructure for healthcare conversations | First bid €3 → #1 |
| CosyVoice Alibaba FunAudioLLM's open multilingual TTS with zero-shot and cross-lingual voice cloning | First bid €3 → #1 |
| Coval Simulation and evals for voice agents | First bid €3 → #1 |
| Cresta Generative AI for the contact center | First bid €3 → #1 |
| Cyanite AI music tagging and search | First bid €3 → #1 |
| Daily Real-time voice and video for AI | First bid €3 → #1 |
| DeepFilterNet Open-source real-time noise suppression for speech audio | First bid €3 → #1 |
| DeepScribe Ambient AI medical documentation | First bid €3 → #1 |
| Deepdub AI dubbing and localization | First bid €3 → #1 |
| Deepgram Voice AI platform for developers | First bid €3 → #1 |
| Dia (Nari Labs) Open-weights expressive dialogue TTS model | First bid €3 → #1 |
| Dreamtonics Synthesizer V AI singing voice synthesis software | First bid €3 → #1 |
| Dubformer AI dubbing with human quality control | First bid €3 → #1 |
| Dubverse AI video dubbing and subtitling platform covering Indian and global languages | First bid €3 → #1 |
| ELSA Speak Speech-recognition app that coaches English pronunciation | First bid €3 → #1 |
| ESPnet End-to-end speech processing toolkit for ASR, TTS and more | First bid €3 → #1 |
| ElevenLabs Lifelike AI voices | First bid €3 → #1 |
| F5-TTS Open flow-matching text-to-speech model offering fast zero-shot voice cloning | First bid €3 → #1 |
| Fadr Free AI stem separation, remixing, and MIDI extraction | First bid €3 → #1 |
| Fireflies.ai AI notetaker for meetings | First bid €3 → #1 |
| Fish Audio Open voice generation and cloning | First bid €3 → #1 |
| Freed AI scribe for clinicians | First bid €3 → #1 |
| FunASR Alibaba's open speech recognition toolkit with pretrained models | First bid €3 → #1 |
| GPT-SoVITS Popular open-source few-shot voice cloning and TTS web UI trained from short voice samples | First bid €3 → #1 |
| Gaudio Lab AI audio separation and spatial audio technology | First bid €3 → #1 |
| Gladia Speech-to-text API for real products | First bid €3 → #1 |
| Gliglish AI language teacher for practicing speaking with feedback and roleplays | First bid €3 → #1 |
| Good Tape Simple secure audio transcription built for journalists | First bid €3 → #1 |
| Granola AI notepad for meetings | First bid €3 → #1 |
| Hamming Automated voice-agent testing | First bid €3 → #1 |
| Happy Scribe Transcription and subtitle platform combining AI with human review | First bid €3 → #1 |
| HappyRobot AI voice agents that handle calls for logistics and freight operations | First bid €3 → #1 |
| Heidi Health AI medical scribe | First bid €3 → #1 |
| Hi Auto AI voice ordering for quick-service restaurant drive-thrus | First bid €3 → #1 |
| Hume AI Emotionally intelligent voice AI | First bid €3 → #1 |
| IndexTTS Open zero-shot voice-cloning TTS system from Bilibili | First bid €3 → #1 |
| InnoCaption Real-time captioning for phone calls using AI and stenographers | First bid €3 → #1 |
| Inworld Real-time voice AI for applications | First bid €3 → #1 |
| Jammable AI voice conversion and song covers | First bid €3 → #1 |
| Jellypod AI podcast studio with custom hosts | First bid €3 → #1 |
| Kaldi Established open-source toolkit for speech recognition research | First bid €3 → #1 |
| Kintsugi Detects signs of depression and anxiety from short voice samples | First bid €3 → #1 |
| Kits AI AI voice tools for musicians | First bid €3 → #1 |
| Kokoro TTS Lightweight open-weight text-to-speech model (~82M params) known for fast, quality synthesis on modest hardware | First bid €3 → #1 |
| Krisp AI noise cancellation and meeting AI | First bid €3 → #1 |
| Krotos AI-assisted sound design software for film and games | First bid €3 → #1 |
| Kyutai Open science lab behind Moshi voice AI | First bid €3 → #1 |
| LALAL.AI AI stem splitter for vocals and music | First bid €3 → #1 |
| LANDR AI mastering, distribution, and creative tools for musicians | First bid €3 → #1 |
| LMNT Low-latency text-to-speech API with voice cloning, built for real-time voice agents | First bid €3 → #1 |
| Langua AI conversation partner with human-like voices for language learners | First bid €3 → #1 |
| Limitless Wearable and app that records, transcribes, and recalls your conversations | First bid €3 → #1 |
| Listnr AI text to speech generator | First bid €3 → #1 |
| LiveKit Realtime infrastructure for voice agents | First bid €3 → #1 |
| Loudly AI music for creators | First bid €3 → #1 |
| Masterchannel AI mastering tuned for streaming platforms | First bid €3 → #1 |
| MeloTTS Open multilingual text-to-speech library from MyShell that runs in real time on CPU | First bid €3 → #1 |
| Millis AI Low-latency voice agent platform | First bid €3 → #1 |
| Modulate Voice moderation with machine learning | First bid €3 → #1 |
| Moises AI music practice and stem separation | First bid €3 → #1 |
| Montreal Forced Aligner Aligns speech recordings with transcripts at phone level | First bid €3 → #1 |
| Moonshine Fast on-device speech-to-text models for edge hardware | First bid €3 → #1 |
| Mubert Royalty-free AI music generation | First bid €3 → #1 |
| Mureka AI music generation platform with editable stems and vocals | First bid €3 → #1 |
| Murf AI Studio-quality AI voiceover | First bid €3 → #1 |
| Musicfy AI voice and music creation | First bid €3 → #1 |
| NVIDIA Riva NVIDIA's GPU-accelerated speech AI SDK for deploying ASR, TTS and translation services | First bid €3 → #1 |
| Nabla Ambient AI assistant for clinicians | First bid €3 → #1 |
| Nagish AI-captioned phone calls for deaf and hard-of-hearing users | First bid €3 → #1 |
| Narakeet Text to speech video maker | First bid €3 → #1 |
| NaturalReader Text-to-speech app that reads documents, web pages and PDFs aloud with AI voices | First bid €3 → #1 |
| Neuphonic Low-latency speech AI with cloning | First bid €3 → #1 |
| Neutone Real-time neural audio effect plugins for music production | First bid €3 → #1 |
| Notta Meeting transcription and summarization across languages | First bid €3 → #1 |
| Numa AI agents that answer calls and book appointments for auto dealerships | First bid €3 → #1 |
| Observe.AI Conversation intelligence for contact centers | First bid €3 → #1 |
| OpenVoice Open-source instant voice cloning that transfers a reference speaker's tone across languages | First bid €3 → #1 |
| Orpheus TTS Open-source Llama-based speech model from Canopy Labs with emotive tags and zero-shot voice cloning | First bid €3 → #1 |
| Otter.ai AI meeting notes and transcription | First bid €3 → #1 |
| Papercup AI dubbing for video | First bid €3 → #1 |
| Parler-TTS Open TTS models controlled by natural-language style prompts | First bid €3 → #1 |
| Parloa Agentic AI for customer service calls | First bid €3 → #1 |
| Phonely AI phone support agents | First bid €3 → #1 |
| Phonic Voice AI platform with built-in reliability | First bid €3 → #1 |
| Picovoice On-device voice AI platform | First bid €3 → #1 |
| Pindrop Voice security, caller authentication, and deepfake detection | First bid €3 → #1 |
| Pipecat Open-source voice agent framework | First bid €3 → #1 |
| Piper Fast local neural text-to-speech engine for low-power devices | First bid €3 → #1 |
| PlayAI (Play.ht) AI voice generation | First bid €3 → #1 |
| Podcastle AI-powered podcast recording and editing | First bid €3 → #1 |
| Podsqueeze Generates show notes, timestamps, and posts from podcast episodes | First bid €3 → #1 |
| PolyAI Enterprise voice assistants that sound human | First bid €3 → #1 |
| RVC WebUI Widely used open-source retrieval-based voice conversion tool behind many AI voice covers | First bid €3 → #1 |
| Rask AI Localize video into 130+ languages | First bid €3 → #1 |
| ReadSpeaker Text-to-speech for websites, apps and embedded products | First bid €3 → #1 |
| Regal Voice AI agent platform for contact centers | First bid €3 → #1 |
| Resemble AI Voice cloning and detection | First bid €3 → #1 |
| Resound AI podcast editing that finds and removes filler words and mistakes | First bid €3 → #1 |
| Respeecher AI voice cloning for studios | First bid €3 → #1 |
| Retell AI AI phone call agents | First bid €3 → #1 |
| Rev AI Speech-to-text APIs by Rev | First bid €3 → #1 |
| Rilla Speech analytics that coaches outside sales reps in home services trades | First bid €3 → #1 |
| Rime Text-to-speech for real-time products | First bid €3 → #1 |
| RipX DAW Stem-based DAW built around AI audio separation and editing | First bid €3 → #1 |
| Riverside AI podcast and video recording | First bid €3 → #1 |
| RoEx Automated AI mixing and mastering for multitrack audio | First bid €3 → #1 |
| Salesken Real-time AI for sales conversations | First bid €3 → #1 |
| Sameday AI phone answering that sells and books appointments for home service companies | First bid €3 → #1 |
| Sanas Real-time accent translation | First bid €3 → #1 |
| Sarvam AI Voice-first AI for Indian languages | First bid €3 → #1 |
| Scribeberry AI medical scribe and documentation | First bid €3 → #1 |
| Sensory Embedded wake word, speech and voice biometrics for devices | First bid €3 → #1 |
| Sesame Natural conversational voice companions | First bid €3 → #1 |
| Silero VAD Compact open-source voice activity detector used widely in speech and voice-agent pipelines | First bid €3 → #1 |
| Sindarin Conversational speech AI persona engine | First bid €3 → #1 |
| Siro AI that records and coaches face-to-face field sales conversations | First bid €3 → #1 |
| Slang.ai Digital phone concierge that answers calls for restaurants | First bid €3 → #1 |
| Smallest AI Ultra-low-latency text-to-speech | First bid €3 → #1 |
| Snipd Podcast player with AI transcripts, chapters, and highlight capture | First bid €3 → #1 |
| Sonauto AI song generator producing full tracks with vocals from prompts | First bid €3 → #1 |
| Soniox Accurate multilingual speech AI | First bid €3 → #1 |
| Sonix Automated transcription, translation, and subtitles for audio and video | First bid €3 → #1 |
| SoundHound AI Voice AI platform powering in-car assistants and restaurant phone/drive-thru ordering | First bid €3 → #1 |
| Soundful AI music generator for brands | First bid €3 → #1 |
| Soundraw AI music for creators | First bid €3 → #1 |
| Soundverse AI music and voice generation assistant | First bid €3 → #1 |
| SpeechBrain Open-source PyTorch toolkit for speech recognition, enhancement, diarization and other audio tasks | First bid €3 → #1 |
| Speechify Text to speech across every device | First bid €3 → #1 |
| Speechmatics Enterprise speech recognition | First bid €3 → #1 |
| Splice Sample platform with AI-assisted sound search and stack creation | First bid €3 → #1 |
| Stable Audio Text-to-audio music and sound generation from Stability AI | First bid €3 → #1 |
| StyleTTS 2 Open TTS model using style diffusion and adversarial training for natural-sounding speech | First bid €3 → #1 |
| Suki AI voice assistant for clinicians | First bid €3 → #1 |
| Sully.ai AI employees for medical practices | First bid €3 → #1 |
| Suno Make any song with AI | First bid €3 → #1 |
| Sunoh.ai AI medical scribe from ambient conversations | First bid €3 → #1 |
| Supertone Real-time AI voice conversion and speech enhancement tools | First bid €3 → #1 |
| Superwhisper AI dictation for Mac | First bid €3 → #1 |
| Synthflow No-code AI voice agents | First bid €3 → #1 |
| Tali AI AI medical scribe for clinicians | First bid €3 → #1 |
| Talon Hands-free computer control via voice and eye tracking | First bid €3 → #1 |
| Telnyx Voice AI on private network infrastructure | First bid €3 → #1 |
| Thoughtly AI phone agents for business | First bid €3 → #1 |
| Toma AI phone agents for car dealerships | First bid €3 → #1 |
| Tortoise TTS Open-source multi-voice text-to-speech system focused on expressive, realistic prosody | First bid €3 → #1 |
| Trint AI transcription and collaborative editing for media teams | First bid €3 → #1 |
| TurboScribe Whisper-powered transcription service that converts audio and video files to text | First bid €3 → #1 |
| Typecast AI voices with emotion control | First bid €3 → #1 |
| Udio AI music creation | First bid €3 → #1 |
| Ultimate Vocal Remover Open-source GUI for AI stem separation — isolates vocals and instrumentals from music | First bid €3 → #1 |
| Ultravox Speech-native model and voice agents | First bid €3 → #1 |
| Univerbal Swiss AI speaking coach for practicing conversations in 20+ languages | First bid €3 → #1 |
| Vapi Voice AI agents for developers | First bid €3 → #1 |
| Verbit AI transcription and captioning for media, legal, and education | First bid €3 → #1 |
| Vocode Open-source voice AI agents | First bid €3 → #1 |
| Vogent Voice AI for real-world calls | First bid €3 → #1 |
| VoiceInk Open-source local dictation app for macOS | First bid €3 → #1 |
| Voiceitt Speech recognition built for people with non-standard speech | First bid €3 → #1 |
| Voicemaker AI text-to-speech voice generator | First bid €3 → #1 |
| Voicemod Real-time AI voice changer | First bid €3 → #1 |
| Voicenotes AI voice note app that transcribes and answers questions about your notes | First bid €3 → #1 |
| Vozy Conversational voice AI for customer operations | First bid €3 → #1 |
| WellSaid Enterprise AI voice studio | First bid €3 → #1 |
| Whisper (OpenAI) Open-source speech recognition | First bid €3 → #1 |
| Whispp Converts whispered or impaired speech into a clear natural voice | First bid €3 → #1 |
| Willow Voice AI dictation app for Mac, Windows and iPhone | First bid €3 → #1 |
| Wispr Flow AI-powered dictation | First bid €3 → #1 |
| Wondercraft AI audio studio | First bid €3 → #1 |
| XRAI Live subtitles for real-world conversations in AR glasses and phones | First bid €3 → #1 |
| Zonos Zyphra's open-weight multilingual TTS model with voice cloning and control over emotion and speaking rate | First bid €3 → #1 |
| eMastered Instant online AI audio mastering | First bid €3 → #1 |
| iZotope AI-assisted mixing, mastering, and audio repair plugins (Ozone, RX) | First bid €3 → #1 |
| pyannoteAI Speaker intelligence and diarization | First bid €3 → #1 |
| sonible AI-powered smart EQ, compressor, and limiter plugins | First bid €3 → #1 |
| telli AI voice agents for outbound customer calls | First bid €3 → #1 |