Stable Audio
Text-to-audio & SFX generation
Generate music tracks and sound effects up to 3 minutes from text with timing control — includes open models.
35 tools for video creators. Every entry is labelled with exactly how far we have verified it — hands-on tested, pricing checked, or listed only. We never claim a test we did not run.
Showing 1–24 of 35
Text-to-audio & SFX generation
Generate music tracks and sound effects up to 3 minutes from text with timing control — includes open models.
Automatic audio post-production
Automatic audio post-production web service for podcasts, broadcasters, and video creators. Leveler, noise reduction, and loudness normalization.
AI music video animator
Frame-by-frame AI animation synced to audio stems — the go-to tool for trippy, audio-reactive music videos.
AI audio enhancement
Free AI speech enhancement that removes background noise and echoes from any microphone recording instantly.
Open voice cloning platform
Top-ranked open-source TTS with instant voice cloning in seconds and a huge community voice library.
Podcast to video audiograms
Turn podcast episodes into shareable audiogram videos with waveforms, captions, and full transcripts.
AI video translator & lip sync
Translate, redub, and lip-sync videos with multi-speaker detection and script editing.
Most realistic AI voice cloning
The industry standard for AI voiceovers, voice cloning, and dubbing in 29+ languages. Perfect for faceless YouTube channels and documentaries.
AI avatars and video translation
Create AI avatar videos with lip-sync translation in 40+ languages. Clone your voice and appearance for scalable personalized video content.
Google DeepMind video generation
State-of-the-art text-to-video with native audio generation, strong prompt adherence, and 4K-quality cinematic output inside Gemini and Flow.
Video translation & dubbing
Localize videos into 130+ languages with voice cloning, lip-sync, and multi-speaker detection — used by top educational channels.
AI audio cleaning
Automatically removes filler words, mouth sounds, stuttering, and dead air from podcast and video recordings.
Realistic physics video generation
Kuaishou video model known for realistic human motion, lip-sync, and 1080p clips up to 2 minutes.
Artistic video generation studio
Superstudio canvas for music videos, audio-reactive visuals, and stylized animation used by musicians and artists.
Expressive character video
Character-3 model generates talking, singing, emoting characters from a single image and audio track.
AI research & podcast notebook
Upload sources and get grounded summaries, study guides, and viral Audio Overview podcasts from your own documents.
AI voiceovers for videos
Versatile AI voice generator offering over 120+ realistic voices in 20 languages. Includes pitch control, emphasis editing, and background music syncing.
Clone your voice with AI
Create a text-to-speech model of your own voice. Edit audio by editing text. Fix mistakes in recordings without re-recording.
AI podcast recording studio
Record studio-quality remote interviews, enhance audio with Magic Dust, clone your voice, and edit text-based.
Studio-quality remote recording
4K remote recording with local tracks, AI transcription in 100+ languages, Magic Clips, and text-based editing.
Royalty-free AI music for videos
Generate unlimited royalty-free tracks by mood, genre, and length — customize energy per section for perfect edits.
Real-time voice cloning
Real-time voice cloning API with emotion control, localization, and deepfake detection. Enterprise-grade security and compliance.
AI noise cancellation
AI-powered noise cancellation for calls and recordings. Removes background voices, barking dogs, and keyboard clicks in real-time.
Text to speech everywhere
Listen to any text — docs, PDFs, web pages — with celebrity AI voices across mobile, desktop, and browser.