386+ AI Audio & Voice tools and assistants (updated 2026)

This page lists 386 AI tools for AI Audio & Voice, each with a direct link to its official site.

386 matching tools
Top ratedNewestName A–Z
AlphyAlphy converts up to 10‑hour audio files in 40+ languages into accurate transcripts, offers quick summaries and key takeaways, enables timestamped Q&A, and transforms content into formats like Twitter threads, blogs, newsletters, and quizzes.alphy.appAnyMusicAnyMusic converts text or lyrics into up to five-minute songs with multi‑voice vocals, layered instrument arrangements (including niche/ethnic instruments), distinct song sections, tempo/style/mood controls, and stems/vocal removal for editing and demos.anymusic.aiAnyToSpeechAnyToSpeech converts text, PDFs, DOCX, URLs, and images into natural‑sounding audio across 16 languages, offering 100+ voices and voice‑cloning from a 30‑second clip. It transcribes and cleans audio, supports translation, and is available via web and Android.anytospeech.comApplioApplio is an open-source AI voice cloning tool featuring over 26,000 models, multi-language support, and cross-platform compatibility. Its user-friendly interface and modular codebase cater to both novice and experienced users interested in advanced audio technology.applio.orgFreearticle2audioarticle2audio turns web articles into spoken audio with natural pauses and contextual voice‑over for images. It summarizes tables, explains code, provides two American English voices, and runs as a web app addable to mobile homescreens, offering a Listen page.article2audio.appaskInputaskInput gathers spontaneous voice notes from team members, transcribes them, and AI‑drives a written summary of key themes and actions. It supports multiple languages and eliminates the need for synchronous meetings, freeing calendar space.askinput.comFreeAudify AIUser-friendly platform for voice synthesis with customizable options and instructions, making it versatile for both developers and creatives.audify-ai.ahmedtokyo.comAudio DiaryAudioDiary records spoken journal entries, automatically transcribes them, and uses AI to produce summaries and personalized goals. Users can attach photos, edit transcripts, tag entries, and export audio, text, images, or PDF. End‑to‑end encryption and cross‑platform availability support secure journaling.audiodiary.aiFreeAudio PenAn app that converts your voice notes into concisely summarized text.audiopen.aiAudioBotAudioBot converts written text to natural‑sounding MP3 audio using over 500 AI voices in multiple languages, including diverse Spanish accents. Users can tweak pitch, speed, and tone, making it useful for video, podcasts, and accessibility.audio-bot.comAudiogenAudiogen is an AI audio creation tool that generates high-quality, royalty-free sounds with endless variations. It supports content creators and audio professionals through features like sound refinement, inpainting, and an upcoming extensive sound library.audiogen.coFreeAudioGenius.aiAudioGenius.ai clones a speaker’s voice accurately for videos, podcasts, and dubbing, and offers real‑time multilingual translation for global meetings and support. Unlimited audio minutes and API integration enable scalable, brand‑consistent voice content.audiogenius.aiAudioreadAudioread transforms articles, PDFs, emails, URLs, and RSS feeds into natural‑sounding audio in 80+ languages, with adjustable speed, MP3 downloads, and private podcast feeds for cross‑device streaming. It offers AI summaries, privacy mode, Slack integration, and an API for developers.audioread.comAudiowaveAIAudiowaveAI turns articles, blogs, PDFs, ePubs, and other text into natural‑sounding audio in 100+ languages, offering up to ten distinct voices. Browser‑based playback, shareable files, and flexible pay‑per‑word credits suit creators and learners.audiowaveai.comAVbeamAVbeam lets users drag‑and‑drop multiple source and target audio files (MP3, WAV, OGG, FLAC) to detect partial matches. It tolerates noise, filtering, amplification, and displays aligned segments with timestamps, similarity %, and playback.avbeam.comFreeBangin' Audio RecorderBangin’ is an audio recording tool that captures high-quality sound, features speech timestamping, intuitive editing, and advanced transcription. It syncs across Apple devices, making it ideal for musicians and anyone needing to document auditory ideas.banginaudiorecorder.comBehnevisBehnevis Persian is a web editor that converts English-based Finglish or Pinglish into Persian script, allows word‑level correction, includes speech‑to‑text, and supports exporting to email, documents, or Word via an add‑on.behnevis.comBlahgetTalkieMoney is a voice-based expense tracker that enables users to log income and expenses using natural language processing. It features smart categorization, enhanced speech recognition, and flexible data management through both voice and typing modes.apps.apple.comFreeBlueprintsBlueprints by Mozilla.ai is a developer-focused platform for creating and integrating AI workflows. It offers tools for document parsing, speech-to-text transcription, synthetic audio detection, and customizable agents, ensuring efficient data management and privacy-conscious model training.blueprints.mozilla.aiFreeBRAIVBraiv Player is an online video hosting platform that supports multi-language content through automatic captions, translations, and AI dubbing. It enables brand customization and provides analytics for measuring audience engagement and language preferences.braiv.coFreeBusyScribeBusyscribe transcribes voice messages from platforms like WhatsApp into readable text, supporting over 65 languages. Its AI-powered service provides accurate, instant transcriptions, ensuring clear communication while maintaining data security and privacy.busyscri.beClipboard TTSClipboard TTS automatically reads clipboard changes with 49 languages and 100+ voices, supports translation, OCR, dictionary, text editing, history, experimental AI summaries, and user‑defined order for mutations on Windows and Linux.clipboardtts.comCloneDubCloneDub automatically translates and dubs videos into 20+ languages while preserving original audio and music. It supports SRT subtitles, offers custom voice cloning, and delivers downloadable audio, video, or combined files within minutes, helping creators reach global audiences efficiently.clonedub.comConfidentierConfidentier is an AI-powered speech analysis tool that evaluates audio and video presentations, offering feedback on delivery, common mistakes, and audience engagement strategies. It helps users refine communication skills and deliver impactful presentations.confidentier.comFree

Nearby categories

View all
Snapshot mode · This page is served from a database snapshot exported on 2026-09-16, not live data. This notice disappears once the live API is connected.