InstantVidBETA
All video models

Kling 3.0 AI Video Generator

Kling 3.0 turns a text prompt or image into a short video. Start with a single product action or a simple scene, then choose the output size and duration that fit your destination.

Create with this model

Studio options

Inputs
Text to video · Image to video
Duration
3–15 s
Resolution
720P · 1080P
Aspect ratio
16:9 · 9:16 · 1:1

Plan your shot

Use text mode to describe a scene from scratch. In image mode, upload a first frame and optionally an ending frame. Describe how the subject moves and how the camera follows, rather than adding several unrelated actions to one clip.

A useful starting point

For a product reveal, define the opening composition, one reveal movement and the final view. Keep any dialogue short and clearly separate spoken words from the description of the scene.

Example prompt

A ceramic coffee cup on a wooden counter. A hand gently rotates the cup as steam rises. The camera slowly pushes in, warm window light, a quiet cafe ambience and the soft sound of ceramic on wood.

Credit cost examples

  • 5s · 720P · 378 credits
  • 5s · 1080P · 504 credits

Examples use text mode and the current site tariff. References and setting combinations can change the cost. Review the studio estimate before submitting; one-time packs and subscriptions share the same model rates.

Pricing

Before you generate

Kling 3.0 offers text and image modes here. For reference images or video guidance, compare the separate Kling 3.0 Omni workflow. Different generations may vary in fine visual details.

Questions about this model

Can Kling 3.0 animate a product photo?

Yes. Upload the photo as the first frame in image mode and describe the motion you want. Add an ending frame if the last composition matters.

What is the difference between Kling 3.0 and Omni?

Kling 3.0 offers text and image workflows here. Kling 3.0 Omni also supports reference images and video, including a video editing workflow.

Which aspect ratios are offered?

The studio offers 16:9, 9:16 and 1:1 presets. With image inputs, check the framing of your uploaded image before generating.

Model documentation

Compare another model

MiniMax H3

Use MiniMax H3 to turn a product image, a written scene or reference media into a short video. Plan the subject, movement and sound together, then compare 768P and 2K credit costs.

View model guide

Seedance 2.5

Seedance 2.5 supports a scene described in text, a starting image or a combination of reference media. Use it when you want to plan a short social clip around a specific action, character or rhythm.

View model guide

Kling 3.0 Omni

Kling 3.0 Omni adds reference-based creation to text and image generation. Use reference images to guide appearance, or bring a video when you want to guide camera movement or edit an existing scene.

View model guide

Veo 3.1 Fast

Google Veo 3.1 Fast supports short scenes with native sound, text or image input and reference images. Try a focused character moment or a product scene with a clearly described soundscape.

View model guide

Wan 3.0

Alibaba Wan 3.0 provides text-to-video and first-image animation workflows. Use a short landscape or camera movement to explore a visual direction, then decide whether to increase resolution or duration.

View model guide

Vidu Q3 Pro

Vidu Q3 Pro offers text and keyframe image workflows for short clips. Start with one subject and action, then use an ending frame if the final composition needs to match a particular image.

View model guide

Runway Gen-4.5

Runway Gen-4.5 is available here as a text-to-video workflow. Use a clear written shot plan to explore a visual idea before building a longer sequence from separate clips.

View model guide