Veo 3.1 Image to Video
Veo 3.1 is the latest state-of-the art video generation model from Google DeepMind.
Model Catalog
Search image, video, audio, and avatar models available in Plotspace.
Veo 3.1 is the latest state-of-the art video generation model from Google DeepMind.
Generate videos from a first and last framed using Google's Veo 3.1
Extend Veo-Created Videos up to 30 seconds.
Veo 3.1 by Google, the most advanced AI video generation model in the world. With sound on!
Text-to-image generation with FLUX.2 [flex] from Black Forest Labs. Features adjustable inference steps and guidance scale for fine-tuned control. Enhanced typography and text rendering capabilities.
Image editing with FLUX.2 [flex] from Black Forest Labs. Supports multi-reference editing with customizable inference steps and enhanced text rendering.
Image-to-video endpoint for Sora 2, OpenAI's state-of-the-art video model capable of creating richly detailed, dynamic clips with audio from natural language or images.
Text-to-video endpoint for Sora 2, OpenAI's state-of-the-art video model capable of creating richly detailed, dynamic clips with audio from natural language or images.
Video-to-video remix endpoint for Sora 2, OpenAI’s advanced model that transforms existing videos based on new text or image prompts allowing rich edits, style changes, and creative reinterpretations while preserving motion.
Generate character ids to use with Sora 2 generations.
Lyria 2 is Google's latest music generation model; you can generate any type of music with this model.
CassetteAI’s model generates a 30-second sample in under 2 seconds and a full 3-minute track in under 10 seconds. At 44.1 kHz stereo audio, expect a level of professional consistency with no breaks, no squeaks, and no random artifacts.