🚀NEW LABGetting Started with Claude AgentsStart lab
Multimodal

Veo

Paper preview
Veo
Paper summary

Google Deepmind’s most capable video generation model generates high-quality, 1080p resolution videos beyond 1 minute; it supports masked editing on videos and can also generate videos with an input image along with text; the model can extend video clips to 60 seconds and more while keeping consistency with its latent diffusion transformer.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack