Video generation
Generate AI video from text or from an image
The catalogue holds video models from several publishers — Kling, Seedance, Wan, MiniMax H3, LTX, Grok. Some start from text alone, some from a start frame, some from reference clips. They sit in the same canvas, on the same credit balance, and can be compared on the same brief.
What the models can do
Text to video, with no start frame
Several models — Kling, Seedance, Wan, MiniMax H3, LTX — generate a shot from the description alone. Useful when no visual of the subject exists yet.
From a still to a moving shot
This is the most common production route: you approve an image, then animate it. On the canvas the image generated at the previous step wires straight into the video node — no download, no re-upload.
References to hold a character or a set
Some models accept several reference images and reference clips, which lets you pin a character, a set and a camera move at once rather than hoping for them.
Sound along with the picture
Several models produce a soundtrack with the shot. For short social video that removes a round trip to a separate tool.
The formats platforms actually take
16:9, 9:16 and 1:1 depending on the model, with durations from 2 to 30 seconds and resolutions up to native 4K. The format is chosen before generating, not by cropping after.
Comparing costs less than being wrong
Video models do not render the same motion or the same material. Running one brief through two of them and choosing afterwards beats discovering the problem after the edit.
The whole chain
From the first frame to the delivered shot, in one space.
- 1 · Settle the start frameGenerate and approve an image with one of the catalogue image models. It sets the framing, the light and the material of the shot.
- 2 · Pick a video modelThe picker shows what each one takes as input, its durations, its ratios and whether it can generate sound. Some require an input image, some do not.
- 3 · AnimateThe image wires into the video node and the shot generates. The render comes back onto the canvas, next to the image it came from.
- 4 · DeriveThe same workflow reruns to produce the other formats or the other subjects in the series, without redoing the settings one by one.
Frequently asked questions
- Can I generate AI video for free?
- The free plan gives access to the whole catalogue, video models included, with a starting credit allowance. A video generation spends considerably more credits than an image, and the cost is shown before you launch it.
- What is the maximum length of a shot?
- It depends on the model: durations run from 2 to 30 seconds depending which one you pick. The picker shows the exact range for each. For a longer sequence you chain several generations inside a workflow.
- Do I always need a start image?
- No, but it depends on the model. Several accept text alone; others require an input image. The picker says which before you launch.
- Do the models generate sound?
- Some do, some do not. It is listed in each model’s capabilities, and it is an option on the generation rather than a separate step.
- Can I produce vertical video for social?
- Yes. 9:16 is available on most catalogue video models, alongside 16:9 and 1:1. The format is chosen before generating, which avoids a destructive crop afterwards.
Generate a video
Start from text, or from an image you have just generated on the same surface.
Keep reading
- A canvas, not a chat boxAn infinite AI canvas to generate, compare and chain: every image and video model on one surface, collaboratively, shippable as an app.
- The AI image generator that does not lock you to one modelAn AI image generator with the whole catalogue in one picker: compare models on the same brief, edit, upscale, on a single credit balance.
- Nano Banana on Beemm VisionUse Google's Nano Banana 2 and Nano Banana Pro image models on one canvas: legible in-image text, conversational editing, output up to 4K.
- Kling on Beemm VisionGenerate video with Kling V3 and Kling O3 4K: text, image or reference to video, up to 15 seconds, multi-shot sequences and generated audio.
- The full catalogueEvery image and video model available in the canvas.