386+ AI Audio & Voice tools and assistants (updated 2026)

This page lists 386 AI tools for AI Audio & Voice, each with a direct link to its official site.

386 matching tools
Top ratedNewestName A–Z
F5-TTSF5‑TTS converts text into natural‑sounding, multi‑language audio with emotion control. It supports zero‑shot voice cloning from a reference file, real‑time processing, and speed adjustment, ideal for audiobooks, e‑learning, and accessibility.f5tts.orgFree5.0GIF with SoundGif Sound enhances GIFs by adding customized sound effects, transforming them into dynamic MP4 videos. The AI analyzes visuals to generate complementary audio, facilitating easy sharing on social media while maintaining high-quality output.gifwithsound.netFree5.0Good TapeGood Tape offers secure and automated transcription services for interviews or other recordings in various languages.mygoodtape.com5.0GoodListenGoodListen is an AI podcast studio that generates highlights, chapters, and clips from long episodes, enhancing the listening experience. It processes extensive audio content and integrates with platforms like Spotify and YouTube for easy access to valuable snippets.goodlisten.co5.0InstadeskInstadesk consolidates VoiceBot, ChatBot, Call Center, Live Chat, and Ticket System across 20+ channels, using multilingual speech recognition and NLP to automate ticket creation. It offers agent assistance, compliance monitoring, and training modules to improve response speed and shorten onboarding.instadesk.comFree5.0LazybirdLazybird turns text into realistic spoken audio using over 200 voices across 100+ languages. Users control accent, tone, speed, pauses, pitch, and pronunciation. Download files for videos, podcasts, audiobooks, or educational content with commercial rights.lazybird.app5.0ListnrListnr AI is a text-to-speech tool offering over 1,000 voices in 142 languages. It features voice cloning, emotion fine-tuning, and supports multiple formats for seamless integration in content creation, enhancing accessibility and engagement.listnr.tech5.0Meloflow AIMeloflow is an AI music generator that enables users to create original tracks and enhance existing audio. Key features include vocal removal, audio layering, and an AI cover generator, catering to musicians and content creators across various sectors.meloflow.ai5.0MS Text-to-Speech DownloaderMicrosoft TTS Downloader converts written text into high‑quality, natural‑sounding speech using Azure’s Text‑to‑Speech service. With a single click, users can play back or download audio, batch‑process multiple files, and bypass Azure credential setup.microsoft-tts-downloader.comFree5.0Mumble NoteMumble Note is an iOS and Mac AI voice‑note app that records, transcribes, and summarizes audio into searchable text, tasks, and outlines. It supports 40+ languages, syncs with productivity tools, and encrypts data end‑to‑end.mumblenote.com5.0MyGPT LinkMyGPT lets users build custom ChatGPT‑style bots inside Telegram, choosing from GPT‑4o, GPT‑3.5‑turbo, or Claude 3‑5‑sonnet. It adds DALL·E 3 image generation, Whisper transcription, GPT‑4 Vision image understanding, and text‑to‑speech, with quick setup and open‑source scripts.mygpt.link5.0Narrator: Audiobook MakerNarrator converts ePub, PDF, DOCX, TXT, and RTF files into natural‑sounding speech in over 25 languages. Playback speed ranges from 0.5× to 3×, and audio can be exported as a single .m4a file. Works offline after voice download.narratorapp.coFree5.0notevibes.comNotevibes transforms text, PDFs, URLs, images, and audio into studio‑quality voiceovers, podcasts, and audiobooks using 550+ voices across 57 languages. It auto‑summarizes content, supports multi‑speaker dialogues, and delivers MP3/WAV downloads for commercial use.notevibes.com5.0Novels AINovels AI generates personalized audiobooks in various genres, allowing users to customize characters and influence narratives. Advanced voice synthesis creates immersive audio experiences, expanding the possibilities of AI-driven storytelling for a tailored listening journey.novels-ai.comFree5.0PodnotesPodnotes transcribes podcasts and audio into text, auto‑generating chapters, summaries, timestamps, and speaker‑segregated transcripts. It converts them into social media posts, blog articles, and other formats across 19+ languages with a one‑click generator.podnotes.app5.0PodsqueezePodsqueeze automates podcast transcription with speaker tags, timestamps, and subtitle export. It produces show notes, summaries, short clips, and audiograms, trims audio, edits subtitles, and offers AI voice tuning and topic suggestion for streamlined production.podsqueeze.com5.0Prankify AICelebrity Voice AI creates realistic voiceovers in the style of over 100 personalities. Neural‑network synthesis lets users adjust pitch, pacing, and emphasis. It’s ideal for narration, promos, or education, offering instant custom text playback and an updated voice library.prankify.lol5.0ReplicastudiosReplica Studios provides realistic AI voice cloning and text‑to‑speech with multilingual libraries and adjustable parameters. It integrates with game engines, animation, and audio editors, supporting batch and real‑time API processing for scalable character voice‑overs, narration, and dialogue.replicastudios.com5.0Soca AISoca AI is a versatile generative AI tool for voice character creation. It provides studios like AI Creator, Advanced Gen AI, Creative, and Dubbing for generating content, voices, videos, quizzes, and cloning voices. Personalized unique voices cater to businesses, marketing, talent development, creative agencies, media, education, and more.soca.ai5.0SoundifySoundify generates royalty‑free audio clips from text prompts in real time, letting users set duration, volume, and speed. It offers preset sound libraries and outputs files ready for use in videos, podcasts, games, or visual projects.soundifytext.io5.0SpeechlabSpeechlab automates speech‑to‑speech translation, enabling bulk video/audio dubbing across 20+ languages. It offers real‑time interpretation with sub‑3‑second latency, API integration, role‑based collaboration, fine‑tuned voice synthesis, and seamless workflow.speechlab.aiFree5.0Splitter.aiSplitter.ai automatically separates audio into 5‑stem (vocals, drums, bass, piano, other) or 2‑stem (vocal, instrumental) tracks, removes reverb, and processes YouTube and cloud uploads. It offers an API for developers and supports producers, DJs, forensic, and karaoke use.splitter.aiFree5.0StockmusicGPTAI tool that creates royalty‑free music, sound effects, and covers from text or image prompts, offering remixing, upscaling, style replication, stem‑splitting, vocal removal, mastering, and audio enhancement across diverse genres.stockmusicgpt.com5.0SwiftinkSwiftink turns spoken audio into written text with high‑speed, hardware‑accelerated speech‑to‑text models. It supports multi‑hour recordings, 95+ languages, domain‑aware terminology, and offers API and plugin integrations, enabling swift turnaround for large media files.swiftink.io5.0

Nearby categories

View all
Snapshot mode · This page is served from a database snapshot exported on 2026-09-16, not live data. This notice disappears once the live API is connected.