HiAPI
  • Models
  • Pricing
Search

Search HiAPI models, tools, and resources.

  • Models
  • Pricing
HiAPI

One API, All AI Models

Generate images, video, and audio with leading models through one production-ready API.

Get a free API key

AI Image API

  • All image models
  • GPT Image 2
  • Nano Banana 2
  • Seedream 5.0 Pro
  • Qwen Image 2.0 Pro
  • FLUX 1.1 Pro

AI Video API

  • All video models
  • Seedance 2.5
  • FLUX.3 Video
  • Seedance 2.0
  • Veo 3.1
  • Kling 3.0

AI Audio API

  • All audio models
  • MiniMax Music 2.6
  • MiniMax Music 1.5
  • ElevenLabs v3
  • Text to music
  • Text to speech

Product

  • Model marketplace
  • Playground
  • Pricing
  • Image API Cost Calculator
  • Free GPT Image 2 Generator
  • Free Nano Banana Image Generator
  • Outfit Preview
  • Product Photo Lab

Developers

  • Documentation
  • API Reference
  • Agent Skills
  • LLM integration index
  • Blog

Company

  • About
  • Contact support
  • Terms of Service
  • Privacy Policy

© 2026 hiapi. All rights reserved.

Open source on GitHubPython SDK on PyPI
  • TL;DR
  • How These Recipes Were Made
  • Recipe 1 — Camera Move: Slow Crane-Up Reveal (16:9)
  • Recipe 2 — Fast Action: Low Tracking Camera (16:9)
  • Recipe 3 — Visual Style: Chinese Ink-Wash Animation (16:9)
  • Recipe 4 — Vertical 9:16: Emotion Close-Up for Shorts (9:16)
  • Recipe 5 — Long Take: 12-Second One-Shot Sequence (16:9)
  • The Input Schema, Verified
  • Run These Yourself (API)
  • FAQ
GuideJul 11, 2026

Kling 3.0 Turbo Text-to-Video Prompt Recipes: Copy-Paste Prompts With Real Outputs

Five field-tested text-to-video prompts — camera moves, action tracking, ink-wash style, vertical 9:16, and a 12-second one-shot — each shown with the actual clip it generated on the live API.

hiapiKlingVideo GenerationPromptsText-to-Video

Latest models

Explore models

Contents
  • TL;DR
  • How These Recipes Were Made
  • Recipe 1 — Camera Move: Slow Crane-Up Reveal (16:9)
  • Recipe 2 — Fast Action: Low Tracking Camera (16:9)
  • Recipe 3 — Visual Style: Chinese Ink-Wash Animation (16:9)
  • Recipe 4 — Vertical 9:16: Emotion Close-Up for Shorts (9:16)
  • Recipe 5 — Long Take: 12-Second One-Shot Sequence (16:9)
  • The Input Schema, Verified
  • Run These Yourself (API)
  • FAQ

Generate it with HiAPI

Choose a model, enter your prompt, and see the result.

HiAPI Blog

Related articles

HiAPI

Generate it with HiAPI

TL;DR

  • What this is: five copy-paste text-to-video prompt recipes for kling-3.0-turbo, each paired with the real clip it produced on the live API — every video on this page was generated with the prompt printed right above it, raw first output, no retries hidden.
  • What's covered: a continuous crane reveal, a fast tracking shot, ink-wash style locking, a vertical 9:16 emotion close-up, and a 12-second single-take sequence that uses the model's flexible duration.
  • Model facts (verified against the live task API): duration is an integer from 3 to 15 seconds, resolution is 720p or 1080p, aspect_ratio accepts 16:9, 9:16, or 1:1. The schema is strict — unknown fields like seed or negative_prompt are rejected with a 400, so typos fail loudly.
  • Audio: every clip below came back with a native ambient audio track — play them with sound on.
  • Pricing (per second of output, from the pricing page): $0.13/s at 720p and $0.16/s at 1080p — the 5s clips here cost $0.65 each, and a 3s 720p probe is $0.39.
  • Speed: all five clips, submitted in parallel, were back in under ten minutes — turbo is the speed tier of the Kling 3.0 family.

How These Recipes Were Made

Each recipe has three parts: a one-line goal, the exact prompt (copy it as-is), and the clip it generated. I wrote every prompt myself and ran it through kling-3.0-turbo on the live hiapi task API; the videos are hosted here permanently, so nothing expires or gets swapped for a cherry-picked demo. Settings are in each caption so you can reproduce the exact call.

Working habits that paid off with this model:

  • Describe one continuous camera move and say "no cuts". kling-3.0-turbo will hold a crane, track, or steadicam glide for the entire clip if the move is the first thing you describe; leave the camera vague and it starts editing.
  • Declare style before content for non-photoreal looks, and name physical materials ("ink gradients on rice-paper texture") rather than just a genre.
  • Write the emotional beat as a change over time ("frown slowly relaxes into a smile") — the model treats the clip as an arc, not a pose.
  • Iterate at 3s / 720p ($0.39 a try), deliver longer. Duration is a free integer from 3 to 15, so you can probe cheap and then re-run the locked prompt at the length the cut actually needs.

Recipe 1 — Camera Move: Slow Crane-Up Reveal (16:9)

Goal: one continuous rising move that ends on a wide reveal, no cuts.

Low angle on rain-slicked cobblestones in an old harbor town at dusk, then one slow continuous crane up past mooring ropes and lantern-lit fishing boats, revealing the full marina glowing under a purple and orange sky. Single unbroken rising move, no cuts, stable cinematic camera, photoreal, reflections shimmering on the wet stone.

Crane-up from wet cobblestones to a lantern-lit marina — kling-3.0-turbo/text-to-video, 5s @ 720p, 16:9. Real output, raw first take, generated for this guide.

Why it works: the shot is written as a path — start low on the cobblestones, rise past the ropes and boats, end on the marina — so the model has a beginning, middle and destination instead of a static scene description. "Single unbroken rising move, no cuts" is the insurance clause; drop it and reveals like this tend to get chopped into two shots.

Recipe 2 — Fast Action: Low Tracking Camera (16:9)

Goal: fast subject motion with the camera matched to it, kept coherent.

Low tracking camera racing alongside a mountain biker tearing down a narrow forest singletrack, roots and rocks blurring past the lens, dry leaves kicked up behind the rear wheel, sunlight strobing through the canopy. Camera stays locked on the rider the whole way down, fast and kinetic, photoreal.

Tracking a mountain biker down a forest singletrack — kling-3.0-turbo/text-to-video, 5s @ 720p, 16:9. Real output, raw first take, generated for this guide.

Why it works: the rider's motion and the camera's motion are described as one system ("low tracking camera racing alongside ... locked on the rider"), so the subject stays framed instead of drifting out of shot. The secondary detail — leaves kicking up, sunlight strobing through the canopy — gives the speed visible evidence without adding another subject to manage.

Recipe 3 — Visual Style: Chinese Ink-Wash Animation (16:9)

Goal: lock a non-photoreal style and keep it stable for the whole clip.

Traditional Chinese ink-wash animation style. A lone red-crowned crane lifts off from a misty mountain lake at dawn, slow wingbeats rippling the water, pine silhouettes dissolving into layered mist behind it. Monochrome ink gradients on visible rice-paper texture, a single red seal-stamp accent, serene unhurried motion.

Ink-wash crane lifting off a misty mountain lake — kling-3.0-turbo/text-to-video, 5s @ 720p, 16:9. Real output, raw first take, generated for this guide.

Why it works: "Traditional Chinese ink-wash animation style" is declared before any scene content, and the prompt gives the style a physical medium — ink gradients, visible rice-paper texture, a red seal-stamp accent — so the model renders a material, not a vague aesthetic. Kling handles this genre remarkably well; the output even placed the seal stamp like a real scroll painting.

Recipe 4 — Vertical 9:16: Emotion Close-Up for Shorts (9:16)

Goal: a phone-native portrait clip where the performance — not the action — carries the shot.

Vertical close-up portrait of an elderly watchmaker examining a tiny brass gear through a jeweler's loupe, lit by one warm workbench lamp in a dark room. His focused frown slowly relaxes into a quiet satisfied smile as the mechanism starts ticking. Shallow depth of field, photoreal skin texture, subtle handheld feel.

Watchmaker's frown easing into a smile — vertical — kling-3.0-turbo/text-to-video, 5s @ 720p, 9:16. Real output, raw first take, generated for this guide.

Why it works: the emotional beat is written as a change over time ("focused frown slowly relaxes into a quiet satisfied smile as the mechanism starts ticking"), which gives the clip an arc to perform instead of a single expression to hold. Set aspect_ratio to 9:16 and the framing composes for the vertical canvas natively — no cropping a widescreen shot after the fact.

Recipe 5 — Long Take: 12-Second One-Shot Sequence (16:9)

Goal: use the flexible 3–15s duration for a mini sequence — three story beats in one unbroken shot.

One continuous shot: the camera glides through the swinging door of a busy late-night diner kitchen, follows a waitress balancing three plates as she weaves between line cooks and drifting steam, then slides past her shoulder as the plates land on the neon-lit counter in front of a customer. Single unbroken steadicam take, no cuts, warm practical lighting, photoreal.

One-shot steadicam ride through a diner kitchen to the counter — kling-3.0-turbo/text-to-video, 12s @ 720p, 16:9. Real output, raw first take, generated for this guide.

Why it works: most text-to-video models box you into fixed 5s or 10s clips; kling-3.0-turbo takes any integer from 3 to 15, and twelve seconds is enough room for an actual sequence — through the door, follow the waitress, land on the counter. Structure the prompt as ordered beats ("glides through ... follows ... then slides past") and the model paces them across the duration you paid for.

The Input Schema, Verified

Everything below was confirmed against the live task API before generating a single clip (invalid values return a 400 that spells out the accepted ones):

FieldTypeNotes
promptstringrequired
durationinteger3–15 seconds; a string like "5" is rejected
resolutionstring720p or 1080p
aspect_ratiostring16:9, 9:16, 1:1

The schema is strict: unknown fields (seed, negative_prompt, ...) are rejected with additional properties ... not allowed, so typos fail loudly instead of being silently ignored.

Run These Yourself (API)

All five recipes use the same async flow — create a task, poll until it's terminal, download the mp4:

import requests, time

API = "https://api.hiapi.ai/v1/tasks"
HEADERS = {"Authorization": "Bearer YOUR_HIAPI_KEY", "Content-Type": "application/json"}

def run(prompt, duration=5, resolution="720p", aspect_ratio="16:9"):
    body = {"model": "kling-3.0-turbo/text-to-video",
            "input": {"prompt": prompt, "duration": duration,
                      "resolution": resolution, "aspect_ratio": aspect_ratio}}
    task_id = requests.post(API, json=body, headers=HEADERS, timeout=60).json()["data"]["taskId"]
    while True:
        t = requests.get(f"{API}/{task_id}", headers=HEADERS, timeout=30).json()["data"]
        if t["status"] == "success":
            return t["output"][0]["url"]   # signed link, expires — download immediately
        if t["status"] == "fail":
            raise RuntimeError(t.get("error"))
        time.sleep(5)

print(run("Low angle on rain-slicked cobblestones in an old harbor town at dusk ..."))

The returned output[0].url is temporary — persist the bytes right away rather than hot-linking. Full parameter reference lives in the docs, and the kling-3.0-turbo API walkthrough covers task submission, polling, and error handling step by step.

FAQ

How long can clips be, and at what resolution?
Any integer duration from 3 to 15 seconds, at 720p or 1080p. The 12-second diner one-shot in Recipe 5 is a single generation, not a stitch.

Does it do vertical video?
Yes — aspect_ratio accepts 16:9, 9:16, and 1:1. Recipe 4 is a native 9:16 generation composed for the vertical frame.

Do the clips have sound?
Every clip in this guide came back with a native ambient audio track (verified on the actual files). If you need fuller generated sound design, the Kling 3.0 omni tier lists dedicated audio options.

What does a clip cost?
Billing is per second of output: $0.13/s at 720p and $0.16/s at 1080p. The four 5s clips here were $0.65 each, the 12s long take $1.56, and a 3s 720p probe is $0.39. Live numbers on the pricing page.

Turbo or Omni?
kling-3.0-turbo is the speed tier — flat $0.13/s (720p) or $0.16/s (1080p), built for volume, and every clip in this guide came back in minutes. The omni tier adds 4K output and explicit audio pricing tiers, plus image-to-video and reference-to-video variants — see the kling-3.0-omni image-to-video guide if you need to animate an existing frame. (Turbo has an image-to-video variant too: kling-3.0-turbo/image-to-video, same pricing.)

Do these prompts transfer to other models?
The structure does — continuous camera language, style-first declarations, beats over time. We keep a matching seedance-2.0-mini recipe collection if you want to compare how another model line responds to the same patterns.


All five prompts are yours to copy. Grab an API key, point the snippet above at kling-3.0-turbo, and a 3-second test clip costs $0.39 — cheaper than most image models' 4K tier.

Latest models

View all models
  • GPT Image 2From $0.007/image
  • Nano Banana 2From $0.051/image
  • Seedream 5.0 ProFrom $0.050/image
  • Seedance 2.5From $0.121/s

Explore models

TextImageVideoAudio
Back to blog
GPT Image 2From $0.007/image
Nano Banana 2From $0.051/image
Seedream 5.0 ProFrom $0.050/image
Seedance 2.5From $0.121/s
View all models
TextChat and reasoning
ImageGenerate and edit
VideoText and image to video
AudioSpeech and music
Start generating
View model pricing
View all articles
Seedance 2.5 Text-to-Video: Build Short-Form Clips with the hiapi API

Seedance 2.5 Text-to-Video: Build Short-Form Clips with the hiapi API

Seedance 2.5 Reference-to-Video for Short-Form TikTok and Reels Clips

Seedance 2.5 Reference-to-Video for Short-Form TikTok and Reels Clips

Grok Imagine Image 2.0 Image-to-Image Prompts: 4 Recipes With Real Outputs

Grok Imagine Image 2.0 Image-to-Image Prompts: 4 Recipes With Real Outputs

Grok Imagine 2.0 Text-to-Image Prompt Recipes: Copy-Paste Prompts With Real Outputs

Grok Imagine 2.0 Text-to-Image Prompt Recipes: Copy-Paste Prompts With Real Outputs

Using flux-2-klein-9b/text-to-image for E-Commerce Product Images via the hiapi API

Using flux-2-klein-9b/text-to-image for E-Commerce Product Images via the hiapi API

Flux-2-Klein-9b Image-to-Image for E-Commerce Product Photos

Flux-2-Klein-9b Image-to-Image for E-Commerce Product Photos

Start generating