Bernini-Diffusers-v2

NEW

Key Features

Unified video generation and editing.
MLLM semantic planner.
DiT-based video renderer.
Qwen2.5-VL planning assets.
Wan2.2 diffusion components.
Complex instruction following.
Reference-guided video editing.
Public Diffusers package and Gradio demo.

The self-contained Diffusers package includes a Qwen2.5-VL planner, Bernini planning weights, and Wan2.2 diffusion components. Compared with renderer-only Bernini releases, v2 adds stronger instruction following and multi-step semantic planning, with a connector warmup and co-training recipe that improves reference-guided editing and image-to-video performance.


The model can be downloaded from Hugging Face and run with the public Bernini repository, including single-GPU inference and a Gradio demo. It is useful for text-to-video, image-to-video, video-to-video, reference-guided editing, and research into planner-renderer video systems.

Get more likes & reach the top of search results by adding this button on your site!

Embed button preview - Light theme
Embed button preview - Dark theme
TurboType Banner

Subscribe to the AI Search Newsletter

Get top updates in AI to your inbox every weekend. It's free!