Creative Studio
API
Resources
Features
About Us
Download
Prompts · In-app tool

The AI video prompt generator built into Kling

You don’t need a third-party prompt tool — Kling ships one. It’s DeepSeek-R1, and it lives inside the video generator itself. Below: exactly where to find it, what it hands back, the one thing it consistently leaves out, and a measured before/after showing why that matters.

The Kling video composer with the DeepSeek-R1 prompt panel open — the built-in AI video prompt generator this page walks through.
On this page
  1. Where the generator is
  2. What it returns
  3. The one thing it leaves out
  4. Turning a suggestion into a prompt
  5. FAQ
The short answer

Open Kling’s video generator, click the DeepSeek whale icon at the bottom-right of the prompt box, and a panel headed DeepSeek-R1 opens. Type a rough idea in plain language and it hands back three finished prompts, each with its own Generate button. It writes the scene for you. It does not write the motion or the camera — that part is still yours, and it is the part that decides whether the clip works.

Where the generator is

1

Open the video generator

Any Kling account can reach it. The prompt box sits in the left-hand panel, under the model selector.

2

Click the DeepSeek whale icon at the bottom-right of the prompt box

It sits in the small row of icons just under the box, to the right of the Multi-Shot toggle. There is no text label — the panel that opens is headed DeepSeek-R1.

3

Type a rough idea, in plain language

Full sentences are not required. Six words was enough for the run below. An image can be attached instead, using the Upload button in the same panel.

4

Press the send arrow, and read all three

You get three separate suggestions rather than one answer. They differ in framing and emphasis, so it is worth reading past the first.

The Kling video-generation composer with the DeepSeek-R1 panel open beside it, showing a polished prompt suggestion with its own Generate button, the Upload control and the Deep Thinking toggle.
The generator in place: the DeepSeek‑R1 panel opens beside the video composer. This is the live interface — suggestion, its own Generate button, Upload and Deep Thinking, exactly as described below. (Interface as of August 2026.)

What it returns

A real run. The whole input was six words — a cat on a windowsill at sunset — and these are the three suggestions that came back, quoted exactly:

“A cat sits on a windowsill at sunset, silhouetted against the warm orange and pink sky, gazing outward.”

“A fluffy cat lounges on a wooden windowsill during sunset, the golden light casting long shadows across its fur.”

“A cat perched on a narrow windowsill watches the sunset, the deep hues of dusk painting the sky behind it.”

Each one arrives with its own Generate button, so a suggestion can go straight to render without ever touching the prompt box. Underneath the set there’s Edit and Regenerate, plus a thumbs up/down.

Two variations are worth knowing about:

  • Deep Thinking — a toggle beside the input. Switched on, the panel reports “Deep Thinking Completed” and returns three prompts that carry more framing and lighting detail (one of ours opened “From inside a room, a cat is seen on a windowsill…”).
  • Upload — attach a reference image instead of typing. Fed a photo of a potted plant on a windowsill, it returned three prompts describing that photo, down to the woven basket and the ocean beyond the glass.
The DeepSeek-R1 panel after an image upload run: the uploaded photo, the Deep Thinking Completed label, a returned prompt describing the potted plant and the ocean view, and the input area with Upload and Deep Thinking.
The image‑upload run from the bullet above, in the real panel: the uploaded windowsill photo at the top, “Deep Thinking Completed”, and a returned prompt that describes the woven basket and the ocean beyond the glass — scene detail, no motion. The disclaimer sits at the bottom.
Worth reading twice

The panel carries its own disclaimer: “This content was generated by AI. Please review and evaluate it carefully.” Treat every suggestion as a first draft, not a verified answer.

The one thing it leaves out

Across three runs — plain text, Deep Thinking, and image upload — all nine suggestions described a scene. Not one of them named a movement, and not one of them said anything about the camera. For a still image that is fine. For video it is the whole game, because whatever you leave unspecified, the model decides for you.

Here is that gap, measured. Same six-word idea, same settings — VIDEO 3.0, 720p, 3s, single shot, no audio, text-to-video — and the only difference is the prompt.

Suggestion used as-is. “A cat sits on a windowsill at sunset, silhouetted against the warm orange and pink sky, gazing outward.” The cat stays more or less still and the camera wanders: comparing the first and last frame, the framing has slid sideways by roughly a tenth of the picture width. Nobody asked for that move — it filled a gap in the prompt. Literal output, unedited.
Same suggestion, plus one motion and one camera instruction. “…gazing outward. The cat’s tail flicks once and its ears turn slowly toward the window; camera locked off.” The tail and ears now carry the shot, and first-to-last-frame drift measures zero. Literal output, unedited.

The generator writes the noun. You still have to write the verb.

Turning a suggestion into a prompt

1

Pick the suggestion closest to your shot, then press Edit

Edit drops the text into an editable field. Going straight to that suggestion's own Generate button skips this step — convenient, but it also skips the two additions below.

2

Add one movement, named specifically

Not “the cat moves” but “the cat's tail flicks once and its ears turn toward the window.” One clear physical action beats three vague ones, and it is the single biggest factor in whether a first attempt looks right.

3

Say what the camera does — even if the answer is nothing

“Camera locked off” is prompt wording, not an interface switch. Leave it out and the model picks its own move, which is exactly what happened in the first clip above.

4

Swap in details only you know

The suggestion is generic by design. Your actual colours, setting and pose will always outperform the stock version — and if you are starting from a photo, see the prompt pack for motion wording that transfers well.

Frequently asked questions

How do you write prompts for AI video generation?
Name one physical motion, then say what the camera does. Those two clauses do most of the work; subject and setting matter less than people expect, because the model can already see the image or infer the scene. Everything on this page is about getting those two right.
How do you write effective prompts for AI video generation?
The difference between a working prompt and a vague one is a verb and a camera instruction. In the measured test above, the same scene description drifted sideways on its own; adding “the cat's tail flicks once… camera locked off” held the frame still and put the movement where it was asked for.
Where is Kling's AI video prompt generator?
Inside the video generator itself. Open Kling's video tool and click the DeepSeek whale icon in the small row beneath the prompt box; a panel headed DeepSeek-R1 opens on the right. There is no separate page or sign-up for it.
Can I use a suggestion without editing it?
Yes — each suggestion has its own Generate button that renders it directly. It is not what we would recommend, though: in our runs the suggestions described the scene but never named a movement or a camera instruction, and whatever you leave unspecified the model decides for you.
Can it write a prompt from an image?
Yes. The Upload button in the same panel accepts a reference image, and the suggestions then describe that image. Note the same limitation: it describes what the picture already shows, and image-to-video needs you to say what should start moving.
Kling AI Team
Kling AI

Written and tested by the Kling AI team, who build the image-to-video workflows and prompt tooling behind Kling.

From prompt to real clip

Open the generator and add the verb

Ask DeepSeek-R1 for a starting scene, then name one movement and say what the camera does. That is the whole difference between the two clips above.