WaveSpeed
WaveSpeed gives creators and developers one web workspace and API for generating images, video, audio, and 3D assets across a large model catalogue.
Audio tools in AI Tools Directory.
WaveSpeed gives creators and developers one web workspace and API for generating images, video, audio, and 3D assets across a large model catalogue.
LANDR combines AI mastering, stem separation, instrument generation and audio cleanup with music distribution, samples and plugins through its LANDR Studio subscription.
Read Aloud Reader converts text, PDFs, documents, ebooks, and articles into highlighted neural speech, with speed controls, browser reading, and downloadable MP3 audio.
Fish Audio generates expressive multilingual speech, clones permitted voices from short recordings, handles multi-speaker dialogue, transcribes audio, and provides real-time APIs for voice applications.
Krisp combines AI noise cancellation, bot free meeting notes, transcription, recordings, accent conversion, and workflow integrations for clearer calls and structured follow up.
Moises separates songs into instrumental stems, detects chords and tempo, creates practice tracks, changes pitch and speed, and includes AI music and voice tools.
MicMonster turns scripts into AI voiceovers with multiple speakers, pronunciation controls, reusable settings, and audio exports for videos, lessons, presentations, and audiobooks.
Cleanvoice edits podcasts and spoken recordings with AI noise removal, filler cleanup, synchronized multitrack edits, transcripts, show notes, and timeline exports for further editing.
LALAL.AI separates vocals and instruments, cleans spoken audio, removes echo, and provides voice changing across web, desktop, mobile, VST, and API tools.
Voice.ai is a free AI voice changer that transforms your voice in real time for Discord, Zoom, and games, with thousands of community voices, fast voice cloning, and text-to-speech.
All-in-one AI audio platform with 1,000+ voices, voice cloning, stem splitting, music generation, and text-to-speech. Trusted by over 3 million musicians and content creators.
AI audio and media production platform covering text-to-speech, voice agents, dubbing, music, transcription, sound effects, and a full creator studio. Freemium, credits-based.
Palabra.ai translates live speech across 60+ languages in under a second, with voice cloning and integrations for meetings, events, streams, and enterprise API use.
AirMusic is a freemium AI music platform for generating songs, lyrics, vocals, instrumentals, covers, stems, MIDI files, and music videos.
Riverside's free drag-and-drop transcription tool uses advanced AI from OpenAI to transcribe audio or video files in over 100 languages, with a user-friendly interface capable of processing hour-long interviews in less than 2 minutes
AI voice generator covering 646 languages with zero-shot voice cloning, text-to-speech, and voice design from text descriptions. One-time credit packs, no subscription required.
RecCloud is a browser-based AI platform for transcription, subtitles, dubbing, vocal removal, and video generation with no software install required.
AI voice companion that answers calls from loved ones with dementia in the caregiver's cloned voice, with 24/7 coverage, real-time monitoring, and emergency escalation.
Superscribe transcribes your voice in real time and automatically logs billable hours as you dictate. One tool for dictation and time tracking, with PDF invoice export built in.
AI Song Maker is a text-based music generator that produces full songs, lyrics, and vocals. It includes tools for vocal removal, track extension, and commercial licensing for creators.
A versatile AI music production suite that combines text-to-song generation with practical editing tools like stem splitting, voice cloning, and MIDI support.
Descript's transforms complex audio and video editing into a text-editing task. It can rapidly label speakers, clone voices realistically with Overdub (it removes filler words as well), produce speedy transcripts, remove gaps in recordings without affecting meaning and provide cohesive output by splicing together clips from different sources
One subscription for 200+ AI models covering chat, image, video, voice and music. Switch models mid-thread.
Generates full-length AI songs up to 8 minutes from text prompts, with commercial licensing, a lyrics generator, and audio production tools included.
Media.io combines AI video, image and music generation with editing tools for product ads, visual stories, talking avatars, photo enhancement and audio cleanup.
Mubert is a generative AI music platform that builds royalty-free tracks in real time, matched to your video, podcast, or app by mood, genre, and duration.
Audo Studio rapidly enhances audio by eliminating background noise, reducing echoes, and adjusting volume, providing great results for a wide range of users.