From the source
Runway developed Gen-2 as a text-to-video model building on latent diffusion, first solving temporal consistency with Gen-1 using input video conditioning, then removing that requirement for Gen-2.
The company's north star is enabling generation of a two-hour film, and it follows a staged release approach for safety and organic use case emergence.






