Stable Audio
Text-to-audio & SFX generation
Generate music tracks and sound effects up to 3 minutes from text with timing control — includes open models.
40 tools for video creators. Every entry is labelled with exactly how far we have verified it — hands-on tested, pricing checked, or listed only. We never claim a test we did not run.
Showing 1–24 of 40
Text-to-audio & SFX generation
Generate music tracks and sound effects up to 3 minutes from text with timing control — includes open models.
Automatic audio post-production
Automatic audio post-production web service for podcasts, broadcasters, and video creators. Leveler, noise reduction, and loudness normalization.
Expressive character video
Character-3 model generates talking, singing, emoting characters from a single image and audio track.
AI audio enhancement
Free AI speech enhancement that removes background noise and echoes from any microphone recording instantly.
Open voice cloning platform
Top-ranked open-source TTS with instant voice cloning in seconds and a huge community voice library.
Best-in-class AI image generation
The gold standard for artistic AI imagery — V7 delivers photorealistic characters and consistent styles for thumbnails.
Open-source image-to-video
Stability AI's open video model for image-to-video generation — free to run locally and fine-tune for research.
Podcast to video audiograms
Turn podcast episodes into shareable audiogram videos with waveforms, captions, and full transcripts.
Text & image to 3D models
Generate textured 3D models from text or images in under a minute — the most popular AI 3D pipeline.
Fast text & image to 3D
Generate textured 3D models from a single image or prompt in seconds, with rigging and export options for games and animation.
Most realistic AI voice cloning
The industry standard for AI voiceovers, voice cloning, and dubbing in 29+ languages. Perfect for faceless YouTube channels and documentaries.
High-quality AI video from text & images
Fast, high-fidelity video generation with smooth motion, keyframes, and camera control from Luma Labs.
Google DeepMind video generation
State-of-the-art text-to-video with native audio generation, strong prompt adherence, and 4K-quality cinematic output inside Gemini and Flow.
AI audio cleaning
Automatically removes filler words, mouth sounds, stuttering, and dead air from podcast and video recordings.
AI content creation
AI writer with SEO optimization, paraphrasing, and image generation. Create blog posts, landing pages, and video scripts that rank.
Artistic video generation studio
Superstudio canvas for music videos, audio-reactive visuals, and stylized animation used by musicians and artists.
AI music video animator
Frame-by-frame AI animation synced to audio stems — the go-to tool for trippy, audio-reactive music videos.
xAI's real-time assistant
Frontier reasoning model with live X data, image generation, and unfiltered research for trend-aware creators.
AI research & podcast notebook
Upload sources and get grounded summaries, study guides, and viral Audio Overview podcasts from your own documents.
Clone your voice with AI
Create a text-to-speech model of your own voice. Edit audio by editing text. Fix mistakes in recordings without re-recording.
AI podcast recording studio
Record studio-quality remote interviews, enhance audio with Magic Dust, clone your voice, and edit text-based.
AI images with perfect text
The best AI image generator for readable text rendering — ideal for thumbnails, posters, and logos.
State-of-the-art open image model
Black Forest Labs' FLUX models lead open-source image generation with exceptional prompt adherence and realism.
Free open-source image upscaler
Free, private, offline AI image upscaling for Linux, macOS, and Windows — no cloud needed.