MAGI-2 Preview

NEW

Key Features

Unified audio-video generation research.
100B-scale model architecture exploration.
Single-stream modeling approach.
40 Transformer layers.
Multi-head mixture-of-experts sparse core.
Top-6 routed experts per head.
Co-designed training and inference systems.
Research focus on consistency and synchronization.

The model uses a single-stream design with a 40-layer backbone, a sparse middle made of multi-head mixture-of-experts layers, and a smaller activated expert set. The training system focuses on communication, kernels, routing stability, optimization, and memory efficiency.


MAGI-2 is intended for researchers investigating large-scale generative video, physical plausibility, long-term consistency, and audio-video synchronization. The preview is explicitly a research milestone, so its results should be treated as evidence about scaling direction rather than a finished commercial product.

Get more likes & reach the top of search results by adding this button on your site!

Embed button preview - Light theme
Embed button preview - Dark theme
TurboType Banner

Subscribe to the AI Search Newsletter

Get top updates in AI to your inbox every weekend. It's free!