Local AI Video Generators and Editors

Turn prompts and photos into short clips with open video models, then upscale old footage, sync lips to new audio or add subtitles in another language.

Subcategories

59 tools
A local speech-to-text browser app with Whisper backends, subtitle translation and speaker labeling. Open source under Apache 2.0, with Docker support.

2.9KUpdated 9 months agoApache-2.0

Windows · Docker · Web#Hugging Face integration#Multilingual#Speaker diarization

Local AI face restoration software built on PyTorch and CUDA. Restore faces in photos and videos, with control over image quality and fidelity.

18.2KUpdated 10 months ago

#Image-to-image#Inpainting

Local AI lip-sync model for Windows and Linux that matches faces to supplied audio, with NVIDIA GPU support and an MIT-licensed codebase.

6.6KUpdated 1 year ago

Windows · Linux · Web#Batch processing#Inpainting#Multilingual

An open-source video translation and AI dubbing tool for Windows, macOS and Linux, with local offline models or cloud APIs. Licensed under GPL-3.0.

19.2KUpdated 2 days agoGPL-3.0

macOS · Windows · Linux · Docker · Web#Batch processing#Human approval#Multilingual

Favicon of WhisperX

WhisperX

1 video
Open source speech-to-text software that runs locally, aligns transcripts word by word, and can label speakers.

24.3KUpdated 4 days agoBSD-2-Clause

macOS · Windows · Linux#Batch processing#Hugging Face integration#Multilingual

Open source desktop app that turns images and motion prompts into video on Windows or Linux with an NVIDIA RTX GPU.

17.3KUpdated 11 months agoApache-2.0

Windows · Linux · Web#Hugging Face integration#Multimodal input

Offline transcription and translation software runs Whisper on Windows, Linux and macOS, with microphone capture and subtitle export.

21.8KUpdated 1 week agoMIT

macOS · Windows · Linux#Hugging Face integration#Multilingual#Speaker diarization

Favicon of Draw Things

Draw Things

4 videos
An AI image generator that runs offline on iPhone, iPad and Mac, with on-device LoRA training and optional self-hosted or managed cloud compute.

575Updated 1 day agoGPL-3.0

macOS · Linux · iOS · Docker#Image-to-image#Inpainting#LoRA

Favicon of LatentSync

LatentSync

1 video
Open-source local AI lip-sync tool with a Gradio interface and Apache 2.0 code license. GPU inference requires 8 GB or 18 GB VRAM, depending on the model.

6.1KUpdated 1 year agoApache-2.0

Web#Batch processing#Hugging Face integration#Multimodal input

A local AI portrait animation tool that transfers facial motion from video to images, with eye and lip controls. Runs on Linux, Windows and Apple Silicon.

19.1KUpdated 4 months ago

macOS · Windows · Linux · Web#Hugging Face integration

An open-source Stable Diffusion app for macOS that generates and edits images offline, with support for SDXL, ControlNet and local model training.

13.6KUpdated 3 years agoAGPL-3.0

macOS#ControlNet#Image-to-image#Inpainting

A self-hosted subtitle translator that runs in Docker and automates media library translation using local Ollama models or cloud services such as DeepL.

881Updated 3 days ago

Docker#Multilingual#Ollama integration

Self-hosted subtitle generator runs Whisper locally on CPU or NVIDIA GPU and connects to Bazarr, Plex, Jellyfin, Emby and Tautulli. MIT licensed.

1.5KUpdated 2 months agoMIT

Docker#Batch processing#Multilingual#OpenAI-compatible API

Favicon of Vibe

Vibe

1 video
Local transcription app for macOS, Windows and Linux. Uses Whisper, Nemotron and Parakeet, with Ollama analysis and optional Claude API summaries.

7.7KUpdated 4 weeks agoMIT

macOS · Windows · Linux#Batch processing#Multilingual#Ollama integration

Open-source AI image and video upscaler with local Python and portable Windows, macOS and Linux versions, licensed under BSD 3-Clause.

36.9KUpdated 2 years agoBSD-3-Clause

macOS · Windows · Linux#Batch processing

An open-source AI frame interpolation tool that processes video locally, builds on RIFE and SAFA, and supports Apple Silicon acceleration under an MIT license.

1KUpdated 1 month agoMIT

macOS

Self-hosted speech-to-text API that runs in Docker on CPU or CUDA GPUs, with Whisper, Faster Whisper and WhisperX. Open source under MIT.

3.3KUpdated 2 months agoMIT

Docker · Web#Multilingual#Speaker diarization#Voice activity detection

A local AI image processing editor for Windows, macOS, and Linux. Build visual workflows with community upscaling models. Licensed under GPL-3.0.

6KUpdated 2 months agoGPL-3.0

macOS · Windows · Linux#Batch processing#ONNX#Visual workflows

An offline speech-to-text and text-to-speech app for Linux and Sailfish OS, with local translation, voice typing and MPL-2.0 open-source licensing.

1.7KUpdated 1 week agoMPL-2.0

Linux#Multilingual#Works offline

Favicon of Video2X

Video2X

2 videos
Open-source AI video upscaler for Windows and Linux, with Anime4K, Real-ESRGAN and RIFE support. Runs locally on a Vulkan-capable GPU.

21.9KUpdated 7 months agoAGPL-3.0

macOS · Windows · Linux · Docker

An open-source subtitle editor for Windows, macOS and Linux, with Whisper transcription and translation through Ollama or LM Studio.

14.4KUpdated 1 day agoMIT

macOS · Windows · Linux#Batch processing#LM Studio integration#Multilingual

AI video generation model you can run on your own GPUs, with text-to-video and image-to-video support, ComfyUI integration and downloadable weights.

12.6KUpdated 3 months ago

Web#Multimodal input#Quantization

A self-hosted AI video generation model with ComfyUI and Diffusers support, image animation, video editing, and fine-tuning tools.

11KUpdated 9 months agoApache-2.0

#Hugging Face integration#LoRA#Multimodal input