Stable Audio
Text-to-audio & SFX generation
Generate music tracks and sound effects up to 3 minutes from text with timing control — includes open models.
28 tools for video creators. Every entry is labelled with exactly how far we have verified it — hands-on tested, pricing checked, or listed only. We never claim a test we did not run.
Showing 1–24 of 28
Text-to-audio & SFX generation
Generate music tracks and sound effects up to 3 minutes from text with timing control — includes open models.
Automatic audio post-production
Automatic audio post-production web service for podcasts, broadcasters, and video creators. Leveler, noise reduction, and loudness normalization.
Artistic video generation studio
Superstudio canvas for music videos, audio-reactive visuals, and stylized animation used by musicians and artists.
AI audio enhancement
Free AI speech enhancement that removes background noise and echoes from any microphone recording instantly.
Open voice cloning platform
Top-ranked open-source TTS with instant voice cloning in seconds and a huge community voice library.
Podcast to video audiograms
Turn podcast episodes into shareable audiogram videos with waveforms, captions, and full transcripts.
Most realistic AI voice cloning
The industry standard for AI voiceovers, voice cloning, and dubbing in 29+ languages. Perfect for faceless YouTube channels and documentaries.
AI music video animator
Frame-by-frame AI animation synced to audio stems — the go-to tool for trippy, audio-reactive music videos.
Google DeepMind video generation
State-of-the-art text-to-video with native audio generation, strong prompt adherence, and 4K-quality cinematic output inside Gemini and Flow.
AI audio cleaning
Automatically removes filler words, mouth sounds, stuttering, and dead air from podcast and video recordings.
Expressive character video
Character-3 model generates talking, singing, emoting characters from a single image and audio track.
AI research & podcast notebook
Upload sources and get grounded summaries, study guides, and viral Audio Overview podcasts from your own documents.
Clone your voice with AI
Create a text-to-speech model of your own voice. Edit audio by editing text. Fix mistakes in recordings without re-recording.
AI podcast recording studio
Record studio-quality remote interviews, enhance audio with Magic Dust, clone your voice, and edit text-based.
Studio-quality remote recording
4K remote recording with local tracks, AI transcription in 100+ languages, Magic Clips, and text-based editing.
AI voiceovers for videos
Versatile AI voice generator offering over 120+ realistic voices in 20 languages. Includes pitch control, emphasis editing, and background music syncing.
Real-time voice cloning
Real-time voice cloning API with emotion control, localization, and deepfake detection. Enterprise-grade security and compliance.
AI noise cancellation
AI-powered noise cancellation for calls and recordings. Removes background voices, barking dogs, and keyboard clicks in real-time.
Text to speech everywhere
Listen to any text — docs, PDFs, web pages — with celebrity AI voices across mobile, desktop, and browser.
Enterprise AI voiceover
Studio-quality synthetic voices with commercial licensing, pronunciation controls, and team libraries for L&D and ads.
AI stem splitter
Extract vocals, drums, bass, and instruments from any track with the Phoenix neural network — used by DJs and editors.
Real-time AI voice changer
Live voice changing and soundboard for streamers and gamers with AI voices and custom voice creation.
Ultra-low-latency realistic TTS
Sonic model delivers 40ms-latency lifelike speech for real-time agents, gaming, and interactive video.
AI voices for musicians
Train and share AI voice models for music, convert vocals artist-to-artist, and generate royalty-free vocals.