Google now offers a photo-to-video feature for Veo 3 through the Gemini app
Google AI Pro and Ultra subscribers can now transform photos into eight-second videos with sound through a new photo-to-video feature available on the Gemini app. According to Google, users have generated over 40M videos with Veo 3 since the model launched in May.
Google has rolled out a new photo-to-video feature in Gemini, allowing users to transform static images into dynamic eight-second video clips with sound using its Veo 3 model.
The feature is now available to Google AI Pro and Ultra subscribers in select countries. Once users upload a photo, describe their desired scene, and provide audio instructions, Gemini will bring their still images to life. The tool opens creative possibilities from animating everyday objects to adding movement to nature scenes and bringing drawings to life.
Since Veo 3's launch in May, users have generated over 40 million videos across Gemini and Flow, Google's AI filmmaking tool. The creativity spans from reimagined fairy tales to ASMR content exploring unique audio experiences.
Google emphasizes safety with extensive red-teaming, policy enforcement against unsafe content, and comprehensive evaluations to prevent misuse. All generated videos include visible watermarks indicating AI creation, plus invisible SynthID digital watermarks for authenticity verification.
Ellie Ramirez-Camara is the News Editor at Data Phoenix, where she writes the daily AI newsdesk — covering model releases, research, funding rounds, and policy across the AI and machine-learning industry. She tracks announcements from labs and startups alike and distills them into clear, source-linked reporting for practitioners.
More news

AWS releases six open-source Hugging Face deployment skills for SageMaker

Google Research releases MilleMiglia logistics benchmark generator

AWS launches AgentCore Runtime V2 with elastic memory and snapshot starts
