Flexible creative formats
Generate in 16:9, 4:3, 1:1, 3:4, 9:16, or 21:9 so the framing starts with the needs of the final channel.
Seedance 2.0 multi-format video
Create text-to-video and image-to-video shots with Seedance 2.0 across six aspect ratios. Choose 5, 10, or 15 seconds, work from 480P drafts through 1080P output, and enable audio when the scene needs it.
Seedance 2.0 sample
Generated example
Generate in 16:9, 4:3, 1:1, 3:4, 9:16, or 21:9 so the framing starts with the needs of the final channel.
Use 480P for low-cost ideation, 720P for balanced iterations, and 1080P after the composition and motion are ready for more detail.
Keep the output visual-only for editing flexibility or enable audio when dialogue, ambience, effects, or music belongs inside the generated scene.
Best-fit use cases
Start from the destination format, then choose the input, duration, quality, and sound treatment that make that particular asset useful.
Develop related widescreen, square, portrait, and cinematic frames around one campaign idea without forcing every channel into the same crop.
Animate materials, lighting, camera movement, and controlled subject actions from a prompt or a carefully composed starting image.
Explore a single beat, transition, location, or character action before building a storyboard or committing to production.
Generate a complete audiovisual hook when sound matters, or keep the clip silent when it will be paired with an existing track or voiceover.

Sample breakdown
A centered rear-following shot, rain motion, reflections, and a portrait frame make this a strong format-first example. The same idea would need a different composition—not just a crop—for a wide campaign asset.
Direction to try
Follow slowly behind a person under a transparent umbrella in neon rain; preserve the centered rear view as wet pavement reflections shimmer in a moody teal-and-amber palette.
Seedance 2.0 offers a broad combination of duration, resolution, aspect ratio, input, and audio choices. That flexibility is most useful when you make those decisions in the right order. Start with the publishing context: a cinematic 21:9 concept, a widescreen presentation, a square product card, and a vertical social hook all need different composition. Once the frame is chosen, write one action that reads clearly inside it instead of expecting the same prompt to work unchanged everywhere.
Text to Video gives you freedom to define the subject, setting, and visual language. Image to Video is better when you already have a first frame that establishes brand design, a product, a character, or a location. In either mode, prompts should describe progression rather than a still image. Explain what changes at the start, how the camera or subject moves, and what the viewer should see at the end. For longer clips, use a simple beginning-to-end motion instead of stacking several unrelated scene changes.
The resolution options support an efficient iteration process. A 480P generation is useful for testing whether an idea, frame, and movement work together. Move to 720P when comparing stronger alternatives, then use 1080P once the shot direction is stable. Optional audio lets you decide whether sound should be generated with the scene or designed later. If audio is enabled, keep cues tied to visible actions and leave space between dialogue, ambience, effects, and music.
Step 1
Start with the final placement. Use lower resolution for discovery and select a duration that gives one action enough time without adding filler.
Step 2
Describe how the shot starts, the primary movement, and the intended final state. Add one camera instruction and only the supporting details it needs.
Step 3
Enable audio when it is native to the moment. Leave it off when the clip will use a separate soundtrack, narration, or post-production sound design.
Example prompt: Vertical fashion portrait in a quiet train carriage at blue hour. The subject looks up as bands of city light travel across the window and coat, the camera makes a slow controlled push-in, subtle carriage ambience and rail rhythm, no music.
Yes. Upload a first-frame image and describe how the subject, camera, lighting, and environment should change over time.
The Seedance 2.0 workflow currently offers 5, 10, and 15 second generations.
Yes. Audio is optional, so you can enable it for an audiovisual scene or leave it off when you plan to add sound during editing.
Choose 480P, 720P, or 1080P. Lower resolutions are useful for iteration, while 1080P provides more detail for a refined result.
Seedance 2.0 supports 16:9, 4:3, 1:1, 3:4, 9:16, and 21:9 on Froging AI.
Describe one coherent progression with a clear beginning, primary action, camera behavior, and final state instead of requesting several disconnected scenes.
Model guides and generators
Different shots benefit from different controls. Explore other Froging AI model pages, then return to the main generator when you want to switch workflows.
Kling 3.0
Create text-to-video or image-to-video clips with optional audio, flexible durations, and output up to 4K.
Veo 3.1
Generate landscape or vertical videos from text and images with native audio included in every result.
MiniMax H3
Direct audiovisual scenes from a prompt or first-frame image with 768P or 2K output and broad format support.
Wan 2.7
Build focused visual shots in 720P or 1080P with text-to-video and first-frame image animation workflows.
Use the main Froging AI video generator to compare models in one workspace and choose the best fit for each shot.
Open the AI video generator