Open Text-to-Video and Image-to-Video Models

Open video models like Wan2.2 and HunyuanVideo that turn a prompt or a still image into a short clip.

16 tools
Favicon of Wan2.2

Wan2.2

3 videos
Open source video generation models that run on your GPU, with text, image, speech and character animation options under Apache 2.0.

17.7KUpdated 1 week agoApache-2.0

#Hugging Face integration

Open-source AI video generator that runs locally with miniFLUX or SD3 models, supports Apple Silicon, and includes a browser interface.

3.2KUpdated 2 years agoMIT

macOS · Web#Hugging Face integration#Multimodal input

An open-source AI animation module for Stable Diffusion that runs locally, works with personalized models, and supports image or sketch guidance.

12.3KUpdated 2 years agoApache-2.0

Web#Hugging Face integration#LoRA#Multimodal input

An image-to-video model for animating still pictures locally, with Python and browser demos. Model weights use Stability AI Community terms.

27.3KUpdated 9 months agoMIT

Web#Hugging Face integration

Open source AI video generator you can run on your own GPUs, with image and text inputs, model training tools, and an Apache 2.0 license.

29.9KUpdated 6 months agoApache-2.0

#Hugging Face integration#Multimodal input

Favicon of SkyReels

SkyReels

1 video
AI video generator with a hosted web app for synchronized sound and reference control, plus downloadable models for local GPU inference.

7.6KUpdated 8 months ago

Web#Hugging Face integration#Multimodal input

An open-source text-to-video model you can run locally through ComfyUI or Python, with Apache 2.0 licensing and LoRA fine-tuning.

3.7KUpdated 11 months agoApache-2.0

Web#Hugging Face integration#LoRA

Local AI video generator with downloadable models, NVIDIA GPU support and LoRA fine-tuning. Code and the 2B model use Apache 2.0.

13KUpdated 11 months agoApache-2.0

Windows · Web#Hugging Face integration#LoRA#Multimodal input

An open-source portrait animation tool that runs on Ubuntu with an NVIDIA GPU, turning a still image and English speech into a talking video.

8.7KUpdated 2 years agoMIT

Linux#Hugging Face integration#Multimodal input#ONNX

An open-source image and video generation framework with 4K text-to-image models, laptop GPU support, and ComfyUI and Diffusers integrations.

9.2KUpdated 2 weeks agoApache-2.0

#ControlNet#LoRA#Multimodal input

Local AI portrait animation software turns images and audio into talking-head videos, with editable facial landmarks and Apache 2.0 source code.

4.3KUpdated 6 months agoApache-2.0

Linux · Web#Hugging Face integration#Multimodal input

Local AI lip-sync model for Windows and Linux that matches faces to supplied audio, with NVIDIA GPU support and an MIT-licensed codebase.

6.6KUpdated 1 year ago

Windows · Linux · Web#Batch processing#Inpainting#Multilingual

Open source desktop app that turns images and motion prompts into video on Windows or Linux with an NVIDIA RTX GPU.

17.3KUpdated 11 months agoApache-2.0

Windows · Linux · Web#Hugging Face integration#Multimodal input

Favicon of LatentSync

LatentSync

1 video
Open-source local AI lip-sync tool with a Gradio interface and Apache 2.0 code license. GPU inference requires 8 GB or 18 GB VRAM, depending on the model.

6.1KUpdated 1 year agoApache-2.0

Web#Batch processing#Hugging Face integration#Multimodal input

AI video generation model you can run on your own GPUs, with text-to-video and image-to-video support, ComfyUI integration and downloadable weights.

12.6KUpdated 3 months ago

Web#Multimodal input#Quantization

A self-hosted AI video generation model with ComfyUI and Diffusers support, image animation, video editing, and fine-tuning tools.

11KUpdated 9 months agoApache-2.0

#Hugging Face integration#LoRA#Multimodal input

More in Open Models