Audio8-TTS-Preview-0.6b
Generates multilingual speech and clones voices from audio samples using a compact open-source model.
Software for editing video, generating voiceovers, or cloning audio.
19 tools
Generates multilingual speech and clones voices from audio samples using a compact open-source model.
Generate, edit, and customize music from text or audio with an open-source model that runs on your local machine.
Automates short video creation from a single text prompt by generating scripts, AI visuals, voiceovers, and background music.
Generate cinematic product videos by instructing an AI agent with a library of shot recipes, styles, and templates.
Generates hand-drawn whiteboard videos from images or scripts by connecting the Codex platform to a local rendering engine.
Orchestrate end-to-end video production using an AI coding assistant to manage research, scripting, asset generation, and editing.
Generate terminal-style ASCII art animations from Python scripts and render them as MP4 videos or self-contained HTML files.
Clones a voice from a short audio sample and generates natural-sounding speech in 14 different languages without requiring fine-tuning.
Programmatically create vector animations and synchronize them to audio using a TypeScript library and a real-time preview editor.
Create MP4 videos programmatically using React components, CSS, and other web technologies.
Generates short-form videos from text prompts with automated voiceovers, captions, stock footage, and background music.
Edit videos on a professional, multi-track timeline with this open-source, cross-platform desktop application.
Generate unfiltered AI images, videos, and lip-sync animations with 200+ models via a self-hosted desktop app or web UI.
Generate long-form video from text, images, and audio with an open-source foundational model.
Render deterministic MP4 videos from HTML, CSS, and JavaScript animations using a CLI and agent-friendly framework.
Transcribe, translate, and dub audio or video content with voice cloning using a free, self-hosted web application.
Generates consistent, multi-shot videos from ideas, scripts, or novels using an autonomous multi-agent framework to handle the entire production pipeline.
Clone voices, generate multi-lingual speech, and dictate into any app with an open-source, local-first AI voice studio that runs on your machine.
MOSS-TTS-Nano is a 0.1B parameter, open-source multilingual text-to-speech model for realtime, CPU-friendly speech generation and voice cloning, supporting 20 languages with a simple local setup.