Genmo
Open-source Mochi video model
Creators of Mochi 1, an open state-of-the-art video generation model with strong motion quality and prompt adherence.
22 tools for video creators. Every entry is labelled with exactly how far we have verified it — hands-on tested, pricing checked, or listed only. We never claim a test we did not run.
Showing 1–22 of 22
Open-source Mochi video model
Creators of Mochi 1, an open state-of-the-art video generation model with strong motion quality and prompt adherence.
Tencent open-source video model
13B-parameter open video foundation model with cinematic quality rivaling closed models — free for developers.
State-of-the-art open image model
Black Forest Labs' FLUX models lead open-source image generation with exceptional prompt adherence and realism.
Text-to-audio & SFX generation
Generate music tracks and sound effects up to 3 minutes from text with timing control — includes open models.
Open-source speech recognition
OpenAI's free, open-source ASR model supporting 90+ languages — the backbone of most captioning tools.
OpenAI text-to-video model
Generate cinematic, physics-aware video clips up to 20 seconds from plain text prompts, with storyboard and remix tools built in.
Open-source image-to-video
Stability AI's open video model for image-to-video generation — free to run locally and fine-tune for research.
Alibaba open video generation
Alibaba's Wan 2.x open-source model family — strong text rendering and motion, runnable on consumer GPUs.
Open-source workflow automation
Self-hostable open-source automation with 400+ integrations. Full control over your data. Perfect for privacy-conscious creators.
Open-source frontier reasoning
Free frontier-level reasoning and writing model — R1 rivals paid models for script logic and research.
Open voice cloning platform
Top-ranked open-source TTS with instant voice cloning in seconds and a huge community voice library.
Free open-source image upscaler
Free, private, offline AI image upscaling for Linux, macOS, and Windows — no cloud needed.
AI art & production suite
Fine-tuned models, real-time canvas, and consistent characters — a full creative pipeline used by game artists and YouTubers.
All-in-one AI video hub
Access Kling, Runway, Veo, Hailuo, and more models from one interface with templates and effects.
AI voices for musicians
Train and share AI voice models for music, convert vocals artist-to-artist, and generate royalty-free vocals.
ByteDance pro video model
Top-benchmarked text and image-to-video model with multi-shot storytelling and precise camera control.
Text & image to 3D models
Generate textured 3D models from text or images in under a minute — the most popular AI 3D pipeline.
Licensed-data cinematic video model
Marey model trained exclusively on licensed footage — commercially safe HD video generation for professional filmmakers.
MiniMax cinematic video model
Strong prompt-following text and image-to-video generation with expressive character animation and camera direction.
Speech AI models via API
Universal-2 speech-to-text with sentiment, chapters, and PII redaction for building media features.
General-purpose AI assistant
OpenAI's conversational AI for brainstorming, script writing, research, and code generation. The swiss army knife for creators.
Google's multimodal AI assistant
Research, script, and brainstorm with Google's frontier model — Deep Research and 1M-token context for long video projects.