Nvidia's AI text-to-video tech could revolutionize the GIF game
Matthew David
News
Nvidia's Toronto AI Lab has unveiled its AI-generated video creation tools, "High-Resolution Video Synthesis with Latent Diffusion Models," which generates videos using Latent Diffusion Models, a type of AI that can produce videos without requiring massive computing power. The AI can make still images move realistically, upscale them using super-resolution techniques, and produce short videos at 4.7 seconds long with a resolution of 1280x2048 or longer videos with a lower resolution of 512x1024 for driving videos. The early demos for text-to-GIF productions are impressive, and while full text-to-video generation remains somewhat nebulous, improvements will undoubtedly become more commonplace.