audio Tools
251 best audio tools and apps, curated and ranked by community upvotes on ProductCool. Updated daily as new audio products launch.
Free read aloud for any article with your Cartesia voices. No subscription on top of your Cartesia plan.
Hemory comes from Hear + Memory. It's not meeting transcription: Hemory listens on the phone or Apple Watch you already own, and keeps every conversation as memory. Your day is auto-split into moments with speaker labels, then settles into a private, searchable memory. Connect it to Claude, Codex, Cursor, or any agent over MCP, and your AI finally has real context to revisit what you heard and build on it.
Small Wins is a 14-day audio movement programme delivered to WhatsApp every morning. No app, no login, no equipment, no floor work. Each day: a 30-second personal intro from your coach, a 7-minute guided session (movement, breathing, a food swap), then a quick check-in that shapes tomorrow. Designed by a physiotherapist, enhanced by AI, built for women 55+ who want to feel stronger without the gym.
Generate custom character voices and direct scene dialogue across Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.
ElevenLabs runs in a browser, so getting a line into an edit means generate, download, find the file, import, repeat. Robot Voice Bridge is a native Mac app that turns that into a voiceover session. Write and direct the script in one window, shape each line with performance tags, fix names with an IPA pronunciation editor, then press S to drop the selected take at the playhead in Pro Tools or Logic. Drag takes into Premiere, Final Cut or anywhere else. Uses your own ElevenLabs account.
Walkie turns speech into polished text in any app, gives you meeting transcription with notes and who said what, and reads anything aloud in 950+ voices — free and fully on-device when you want. Mac, Windows, Linux, and iPhone.
Hold a key, talk, and the words land wherever you're typing. Blurt is a free, open-source Mac dictation app powered by AssemblyAI's Dictation API.
Hola AI answers calls when you can't, speaks with callers, takes messages, filters spam, and sends instant call summaries. It also records and transcribes conversations, helps you understand why someone called, and gives you the context you need before calling back. Stay reachable without picking up every call.
CodaBridge lets you listen to real sperm whale recordings, shape a personal synthetic coda, compare timing patterns, and explore evidence in Context Lab. Astra can investigate the measurements using CodaBridge's evidence tools, connecting its explanations to source rows, controls, and uncertainty. The goal is to make exploration useful and traceable without treating similarity as translation.
Foleyfy turns short recordings into organized sound-effect packs. Choose events, capture a few variations, review the takes, and export named WAV files in a ZIP. Native iPhone app with on-device processing. Currently in development.
Build AI agents by simply describing what you want. Duvi brings voice, chat and actions into one agent that knows your business, connects to your tools, and works across your website, phone and WhatsApp without building separate experiences for every channel.
Free real-time voice changer app for PC & Mac. Hundreds of community AI voices, real-time effects, soundboard, text to speech and Audio Lab — all running locally. Less than 50ms delay on ultra fast voices and quality conversions that fully change your voice even if you are singing!
MosMos goes beyond voice dictation by turning both individual thoughts and group conversations into usable writing. Speak naturally in any app, get fast and accurate text in the style you choose, or ask MosMos to search the web for current information. Its personal glossary remembers specialized terms after you add them once. In multi-speaker meetings, it tracks precise timestamps, distinguishes speakers, and creates structured notes, summaries, decisions, and action items for easier follow-up.
Simulate realistic callers at scale with custom personas, scenarios, interruptions, noise, accents, and network conditions. NovaSynth runs those calls against your voice agent, scores audio and transcripts across 30+ dimensions, and surfaces the failures and fixes that matter to your team.
Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet, built for natural conversation.
Voiskey starts from what you meant, not just what you said. Speak a rough thought and it comes back shaped for where it's going and who's reading it: casual with a friend, composed with a colleague, technical with an AI. It arrives 5x faster than typing, cleaned up and ready to send, while you still sound like you. Voiskey is now available on iOS, macOS, Android, and Windows, in over 100 languages. It's free to start. Join during launch and get a free month of Pro.
VoxCPM2 is a tokenizer-free text-to-speech system that generates high-quality, multilingual speech. It solves the problem of rigid and unnatural synthetic voices by enabling creative voice design and producing true-to-life voice clones. This tool is for developers, content creators, and researchers who need versatile, authentic, and controllable speech synthesis across multiple languages.
Your thoughts shouldn’t have to slow down for a keyboard. With Loqua, you can speak naturally to turn rough ideas into ready-to-use writing, understand what’s on your screen, listen when you’d rather not read, and move work forward by voice - from rewriting and translation to scheduling and coding workflows. Less typing, less context switching, more time in flow.
Convert supported audio to editable MIDI locally in your browser, with free Raw MIDI downloads and an optional one-time cleaned MIDI file.
Dictantor captures meetings and voice notes across Mac, iPhone, and Apple Watch. On Mac, it records microphone and system audio together and prompts you before meetings. Transcription runs on device in 25 European languages, separating your voice from everyone else. Search returns you to exact spoken moments. No Dictantor account or subscription. Supported recordings can sync through your private iCloud account.
Annotate your Mac screen with voice, capture full-screen screenshots, and keep copied text ready to reuse from the notch.