Best AI Audio & Music Tools
AI music generators, voice cloning, text-to-speech, and audio tools
The most natural AI voices available — text-to-speech, voice cloning and dubbing in 29 languages, free tier included.
Build a audio & music tool? This is the list AI assistants read.
These are the audio & music tools ChatGPT names when someone asks for a recommendation, and ElevenLabs is the one it names first. If you built one that isn't here, adding it is free — the listing publishes after review. Want it live in minutes with a Verified badge instead? That option is on the form, one-time, no subscription.
All Audio & Music Tools (87)
ElevenLabs
Ultra-realistic AI voice generation and cloning
Suno
Create complete AI songs with vocals and instruments
Udio
Professional AI music generation with vocals
Podcast.ai
Generate full AI podcast episodes with hosts
Resemble AI
Enterprise AI voice cloning and synthesis platform
Boomy
Create and release AI songs to streaming platforms
Mubert
Infinite AI music streams for content
Beatoven.ai
Emotion-based AI music for videos and podcasts
Loudly
AI music with full instrument and genre control
Splash Pro
Create AI songs with vocals from text prompts
Deepgram
Enterprise speech-to-text API with audio intelligence
WellSaid Labs
Enterprise AI voices for professional content
📬 Get the best new AI tools delivered weekly
One concise email with fresh launches, trending picks, and featured standouts.
LOVO AI
AI voice generator with video editor
Listnr
AI voice for podcasts with hosting and distribution
Coqui
Discontinued — open-source TTS, community forks available
Adobe Podcast
AI podcast enhancement with one-click studio sound
Cleanvoice AI
Auto-remove filler words and cleanup audio
Auphonic
Automatic audio post-production and distribution
Riverside.fm
Remote podcast studio with AI editing
Hume AI
AI toolkit for emotionally intelligent voice creation and emotion detection
Cartesia
Ultra-low-latency TTS API — 90ms for voice agents and real-time apps
Descript
Edit audio and video like a document — AI-powered podcast and video editor
Murf AI
AI voice generator — studio-quality voiceovers in 120+ voices
Speechify
Turn any text into audio — the leading text-to-speech reading app
Play.ht
900+ ultra-realistic AI voices for voiceovers, podcasts, and voice apps
Podcastle
AI podcast platform with voice cloning, noise removal, and auto-editing
Soundraw
AI music generator — create custom royalty-free music for any content
AIVA
AI composer for cinematic, orchestral, and game soundtracks
AssemblyAI
Speech AI API — transcription, diarization, and audio intelligence for developers
NaturalReader
Text-to-speech app for reading PDFs and documents aloud — 200+ AI voices
Castmagic
Turn podcast episodes into show notes, blog posts, social content, and newsletters automatically — AI content repurposing for audio
Alitu
All-in-one podcast maker — automatic audio cleanup, silence removal, and one-click publishing for non-technical creators
Lalal.ai
AI audio stem separator — split any song into isolated vocals, instrumentals, drums, bass, guitar, and more with no artifacts
Autopod
AI podcast editor plugin for Adobe Premiere Pro — auto multi-camera switching, silence removal, and social clip generation
Podsqueeze
AI podcast content generator — show notes, transcripts, blog posts, and social content from your RSS feed automatically
Voicemod
Real-time AI voice changer and soundboard for gaming and streaming
Snipd
AI podcast player that captures highlights and syncs to Notion/Obsidian — turn podcasts into knowledge
Podwise
Extracts mind maps, takeaways, and quotes from podcasts without listening — AI podcast research tool
Speechlab
AI video dubbing in 30+ languages while preserving original speaker voices — enterprise localization
Artlist
Royalty-free music licensing for video creators with AI matching.
Moises
AI music separation — isolate vocals, stems, and instruments
Voxtral TTS
Mistral's TTS model — 9 languages, emotional expression, beats ElevenLabs Flash v2.5
Voxtral Transcribe 2
Mistral's STT family — sub-200ms realtime + best-in-class batch at $0.003/min
Spotify AI DJ
Premium-only AI feature that curates and narrates a personalized Spotify music stream
Whisper
OpenAI's open-source speech-to-text model powering countless transcription tools
Retell AI
Developer platform for low-latency AI voice agents that handle real phone calls
Sesame AI
Viral full-duplex conversational voice AI behind the Maya/Miles demo and open CSM model
Orate
On-device macOS text-to-speech that turns selected text from any app into a menu-bar listening queue
The Daily FM
Daily summary podcasts generated from your own sources — articles, X handles, Hacker News, podcasts, and state legislatures
FluidVoice
Free open-source macOS dictation app with fully on-device speech-to-text and a local model that polishes the output
Willow
Voice dictation and AI keyboard for Mac, Windows, and iPhone that types into any app you already use
Unreal Speech
Low-cost text-to-speech API on Kokoro-82M with 300ms streaming, 10-hour requests, and per-word timestamps
Voicetypr
On-device voice dictation for Mac and Windows with a one-time licence instead of a subscription
talat
On-device meeting recording, dictation and file transcription with speaker ID and local AI notes
SpeakNotes
Meeting-bot transcription that produces structured notes with owned action items, not raw transcripts
Whisper Memos
iPhone and Apple Watch voice recorder that emails back formatted AI transcripts and custom-prompt summaries.
Audioread
Turns articles, PDFs and emails into life-like speech delivered as a private podcast feed you play in any podcast app.
SpeechText.AI
Pay-as-you-go transcription with domain-specific speech models, speaker identification, 50+ languages and an API.
TalkToType
Push-to-talk dictation into any app on Windows or Mac, plus meeting recording that captures both sides of a call without joining as a bot, with transcripts and AI summaries.
VoiceScriber
On-device iPhone transcription in 100+ languages that works in airplane mode — no internet, no uploads, no cloud, with editing, export and a home-screen widget.
Despeech
A Whisper transcription API that takes an audio URL instead of an upload, handling download, queueing, diarization and storage in one call across 90+ languages.
HoldSpeak
Hold-to-talk macOS dictation that transcribes 100% on-device in 100+ languages, works offline, has unlimited usage, and is sold once from $19 with no subscription.
HyperWhisper
Open-source macOS and Windows dictation that runs local models offline or your own keys across Groq, OpenAI, Deepgram, AssemblyAI, ElevenLabs, Mistral, Grok and Gemini.
Pepys
Transcription with no length cap and no subscription: speaker-labelled transcripts in 99+ languages, AI summaries and chat, plus API/MCP access, where one credit purchase unlocks every pro feature for good.
OBSIDIAN Neural
An open-source VST3/AU plugin that generates audio samples on your own CPU in about 10 seconds, fully offline with no account — a live performance instrument rather than a song generator.
Transgate
Transcription and translation for calls, meetings and video in 50+ languages with AI summaries, smart highlights and chat-with-transcript, offered pay-as-you-go by the hour as well as monthly.
Miso Labs
Low-latency text-to-speech foundation models for voice agents, with 110ms latency and one-shot cloning
EKHOS AI
Unlimited on-device audio and video transcription with speaker ID, 98 languages and no cloud upload
Ecrett Music
Scene, mood and genre-based royalty-free music generation with unlimited downloads from $4.99/month
Pinch
Pay-as-you-go dubbing and real-time translation API that preserves the original speaker's voice
EdgeWhisper
Fully on-device macOS voice dictation on Mistral Voxtral — sub-500ms, offline, 13 languages
ZenMic
Text, PDF or URL to multi-speaker AI podcast with a full script editor and its own RSS feed
MuseGen
AI music studio with vocals, stems, MIDI export and music video generation on one credit pool
Voice-Swap
Rights-first AI voice conversion for musicians, with licensed artist voices and a DAW plugin
GetTranscribe
Transcribes and structurally analyses social video — hooks, scenes, patterns
BizCrush
In-person-first meeting recorder with 60-language real-time interpretation
Typecast
Expressive AI text-to-speech with emotion control, voice cloning and an API
MusicGPT
AI music generation with remix, stems, sound FX, TTS and commercial use
PlainScribe
Pay-as-you-go AI transcription in 47 languages with 7-day auto-delete and no subscription
Staccato
Text-to-MIDI AI plugin that creates, accompanies, rewrites and extends music inside your DAW
DeVoice
AI transcription for audio, video and YouTube in 100+ languages with speaker labels and SRT export
Transkriptor
Audio and video transcription in 100+ languages with AI summaries, transcript chat and an hours-based bulk ladder
Kinjari
Unlimited music distribution to 100+ platforms for $3/mo with a 90% royalty split and free YouTube Content ID
PodText
Podcast transcription from an Apple, Spotify, iHeart or RSS link, with AI summary, chapters and SRT/VTT export
Murf AI
Studio-quality AI voiceovers in 120+ voices
MP3 to MIDI Converter
Free AI tool that converts MP3 audio into editable MIDI notes in seconds
MusicGPT Pro
AI music generation tool for creators and musicians
Want your tool featured here?
Featured tools appear first on this page and get surfaced to AI search engines like ChatGPT and Perplexity. Every plan includes a permanent dofollow backlink to your site.