Step 1
Start with the delivery format
Choose the aspect ratio for the final channel first. A vertical composition and a wide establishing shot need different subject placement and movement.
Wan 2.7 visual video generation
Generate focused visual shots with Wan 2.7 from a text prompt or opening image. Choose a five- or ten-second duration, work across landscape, portrait, square, and classic formats, and export in 720P or 1080P.
Wan 2.7 sample
Generated example
Wan 2.7 is a practical choice when the visual direction matters more than generated audio. The workflow keeps attention on subject motion, camera movement, framing, lighting, and style. Because the result is visual-only, you can design the clip around an existing voiceover, soundtrack, edit rhythm, or sound design plan without asking the generation prompt to solve every layer at once. This is especially useful for creators who already assemble final sequences in an editor.
Text to Video is best for inventing a scene and exploring several compositions quickly. State the subject and action first, identify the location, then select one camera move. Image to Video is better when a product render, campaign still, character design, or photograph already contains the composition you need. The uploaded image becomes the first frame, so the prompt should concentrate on change: how the subject moves, how the environment reacts, where the camera travels, and which visual details should stay consistent.
Choose a five-second duration for a single gesture, reveal, loop, or transition. Choose ten seconds when the shot needs a short progression, such as a camera approaching a subject before an action occurs. Use 720P for exploratory drafts and 1080P for a more detailed result after the movement works. Wan 2.7 supports five common text-to-video aspect ratios, making it useful for preparing related visual ideas for widescreen, mobile, square, and editorial placements.

Sample breakdown
This Wan 2.7 sample is presented as a visual building block. Judge composition, motion, and rhythm on their own, then decide how narration, music, or sound design should support the exported clip.
Direction to try
Keep the subject readable, describe the visible change with concrete verbs, and use one camera movement that can complete naturally inside five or ten seconds.
Step 1
Choose the aspect ratio for the final channel first. A vertical composition and a wide establishing shot need different subject placement and movement.
Step 2
Write one main action and one camera move. Add lighting, weather, material, and style details only when they affect what changes during the shot.
Step 3
Review the visual rhythm, then pair the exported clip with narration, music, ambience, or effects in your editing workflow.
Use compact five- or ten-second clips to explore motion, composition, camera behavior, and visual style without building an oversized scene.
Upload an image to establish the subject and composition, then describe the exact movement and camera change that should follow.
Create 16:9, 9:16, 1:1, 4:3, or 3:4 text-to-video shots and choose 720P or 1080P output for the intended channel.
Example prompt: A translucent blue fabric installation floating above a concrete gallery floor, the fabric slowly unfolds as sunlight moves through it, visitors remain still in the distance, gentle lateral camera slide, quiet contemporary editorial style, soft natural shadows.
Best-fit use cases
Choose Wan when the generated shot will live inside a larger edit and you want to retain control over voiceover, music, and final sound design.
Bring a product still, portrait, illustration, or landscape into motion while keeping the uploaded composition as the first-frame reference.
Create short visual sequences that can sit beneath narration, titles, music, or a live presentation without competing generated audio.
Explore how the same creative idea should be reframed for widescreen, portrait, square, 4:3, or 3:4 placements.
Compare an orbit, push-in, pan, overhead drift, locked shot, or handheld follow before selecting a direction for a larger project.
Yes. Use Text to Video to invent a shot or upload a first-frame image when you want to animate an existing composition.
Wan 2.7 currently supports 5 and 10 second generations on Froging AI.
The current Froging AI Wan 2.7 workflow produces visual-only video. Add music, narration, ambience, or effects during editing.
The available output resolutions are 720P and 1080P.
Text-to-video supports 16:9, 9:16, 1:1, 4:3, and 3:4. Image-to-video follows the composition of the uploaded first frame.
Use an image when subject identity, product details, composition, or art direction already exists and should anchor the generated motion.
Model guides and generators
Different shots benefit from different controls. Explore other Froging AI model pages, then return to the main generator when you want to switch workflows.
Kling 3.0
Create text-to-video or image-to-video clips with optional audio, flexible durations, and output up to 4K.
Veo 3.1
Generate landscape or vertical videos from text and images with native audio included in every result.
MiniMax H3
Direct audiovisual scenes from a prompt or first-frame image with 768P or 2K output and broad format support.
Seedance 2.0
Create multi-format videos from text or images with optional audio, 480P–1080P output, and 5–15 second durations.
Use the main Froging AI video generator to compare models in one workspace and choose the best fit for each shot.
Open the AI video generator