386+ AI Audio & Voice tools and assistants (updated 2026)

This page lists 386 AI tools for AI Audio & Voice, each with a direct link to its official site.

386 matching tools
Top ratedNewestName A–Z
VoiSparkVoiSpark is an AI voice generator for text-to-speech and voice cloning, offering 500+ natural voices in 30+ languages. It enables custom emotions, styles, and unique vocal identities, with seamless integration for voiceovers in videos, podcasts, and apps.voispark.com2.5Cleanvoice AICleanvoice AI automates podcast post‑production by removing background noise, filler words, pauses, mouth sounds, and breath artifacts in 20+ languages. It offers transcription, summaries, show notes, chapter markers, multi‑track editing, a drag‑and‑drop interface, and an API for batch processing.cleanvoice.ai3.5Hume AIHume AI offers emotion‑intelligent text‑to‑speech, real‑time speech‑to‑speech, and expressive voice cloning across 100+ languages. Developers use TypeScript, Python, .NET, or Swift SDKs to build voice‑design, stage‑direction, and emotion‑analysis features for content creation.hume.ai3.4Typecast AITypecast: AI voice generator for content creation - Emotional TTS, Voice cloning & extensive character library for efficient VSTB, Product marketing & Training videos.typecast.ai3.4AudioXAudioX is an AI audio generation tool that converts text, images, and videos into high-quality music and sound effects. It offers customizable audio parameters, multi-track editing, and supports 30+ music styles for versatile creations.audiox.app2.8KrispKrisp delivers real‑time noise cancellation, accent conversion, and multilingual voice translation for meetings and call centers. It records calls, transcribes, and summarizes, syncing to CRMs. Developers can embed its voice SDK into custom applications.krisp.ai3.3Genve AIGenve AI is a video localization platform that automates transcription, translation, and AI dubbing with voice cloning and lip-sync. It scales multilingual content creation for 30+ languages, preserving brand voice and streamlining distribution for social media, ads, and training.genve.ai2.5AudimeeAudimee is an AI‑driven audio platform that transforms vocal recordings into studio‑quality covers or new takes. It offers pre‑trained voice personas, custom model training, vocal isolation, stem splitting, and seamless DAW integration for streamlined production.audimee.com3.4Kits AIKits AI offers studio‑quality audio tools for musicians and voice artists, including AI voice cloning, vocal isolation, stem splitting, and an instrument library. Accessible via web or API, it supports rapid iteration and collaborative remote demos.kits.ai3.3Elser AIElser AI is an all-in-one anime creation studio that converts text or images into anime art, comics, and short films, offering image/video generation, character creation, voice cloning, lip sync, sound design, storyboarding, and upscaling.elser.aiFree2.0kikivoice.aiKikiVoice is an AI voice cloning tool designed for creators, enabling rapid generation of realistic voice clones from short audio samples. It offers versatile models for various applications, including voiceovers and multilingual content creation.kikivoice.ai2.0DupDubDupDub converts ideas into polished text, offers AI text‑to‑speech with 700+ voices across 90 languages, creates animated speaking avatars, automates video editing with subtitles and effects, and provides voice cloning and API integration for streamlined media production.dupdub.comFree3.3LingvanexLingvanex delivers on‑premise machine translation and speech‑to‑text for over 100 languages, with APIs, SDKs, desktop and mobile apps, enabling secure, offline multilingual content processing, summarization, and data anonymization for business intelligence and compliance.lingvanex.comFree3.2SUNO AI音楽生成器-SongMakerSongMaker AI converts text prompts into complete songs with generated melodies, instrument arrangements, vocal synthesis and versioning. Users select styles, create variations, and export ready tracks for prototyping, production handoff, or custom media use.songmakerai.netFree1.3VozardVozard is an AI‑powered voice changer that delivers low‑latency, real‑time voice transformations for gaming, streaming, and online calls. It offers 200+ presets, adjustable parameters, background effects, export options, and integrates with Discord, Zoom, OBS, Twitch, and Fortnite.imobie.com1.3WriteVoiceWriteVoice is a voice-to-text application that converts speech to punctuated text at 4x typing speed with 97%+ accuracy. It handles accents and technical terms, integrates with popular productivity tools, and is privacy-focused with no data storage.writevoice.ioFree1.3TrancyTrancy delivers bilingual subtitles for YouTube, Netflix, and educational platforms, featuring a reading mode, AI‑powered word lookup, grammar analysis, and part‑of‑speech tagging. It offers customizable translation engines, TTS voices, adjustable display options, and offline learning decks.trancy.orgFree2.8Sound Effect GeneratorA web-based platform that produces seamless looping sound clips up to 20 seconds long, featuring an intuitive interface for generation and preview, a history panel, and instant browser playback or downloads—ideal for editors, game developers, podcasters, and creators.soundeffectgenerator.comFree2.1Dubbing AIDubbing AI is a free, real-time voice changer tailored for gamers and social media users. It enables transforming your voice to match game characters or anime personas, supporting 40 languages across popular platforms for immersive social experiences.dubbingai.ioFree3.0AssemblyAIAssemblyAI offers real‑time and batch speech‑to‑text transcription across 99+ languages, featuring speaker diarization, sentiment analysis, and language identification. It supports medical terminology, PII redaction, and custom prompts for precise conversational insights.assemblyai.com2.2MiniMaxMiniMax is an AI platform providing text, speech, video and music models for developers and creators — supporting agentic text workflows, real-time speech synthesis and voice cloning, emotion-aware video rendering, and precise vocal/instrument music generation via APIs and SDKs.minimaxi.comFree2.9Vbee AI VoiceVbee Aivoice is an AI text-to-speech platform that converts text into natural-sounding audio across multiple languages. It offers various voices, supports voice cloning, and provides MP3/WAV output, ideal for podcasts, e-learning, and audiobooks.vbee.vnFree2.7AudioPod AIAudiopod AI is a platform for voice and audio processing, offering speaker separation, AI dubbing, high-quality stem separation, and noise reduction, making it suitable for content creators, podcasters, and educators to enhance audio quality.audiopod.ai2.2AudioNotesAudionotes AI tool for effortless voice-to-text conversion, organization, summarization, and content generation.audionotes.appFree

Nearby categories

View all
Snapshot mode · This page is served from a database snapshot exported on 2026-09-16, not live data. This notice disappears once the live API is connected.