HiAPI
OverviewModel MarketplaceAPI KeysUsage StatisticsCall LogsBillingReferralPlaygroundStorageChangelogContact UsSettings
Display unit
N
Powered by hiapi
Settings

Welcome

Contact Us

Create 3-15 second videos at 720p, 1080p, or 4K with square, landscape, vertical, and optional sound output.

Provider: Kuaishou

Category: video generation

Endpoint: /v1/tasks

Status: Available

Cost: See live page pricing

Back to Models

Kling 3.0 Omni Text to Video

Native Audio
by KuaishouVideo

Kuaishou Kling 3.0 Omni text-to-video: cinematic motion, multi-shot storytelling, native synced audio, up to 4K.

Pricing

Standard Usage

258 Credits/ s

Input

0 / 5000
5s
3s1-second steps15s

Choose 3-15 seconds in 1-second steps

More parameters

Generation preferences

Generate synchronized audio (sound effects / ambience); costs more when enabled

Estimated cost1,290 Credits258 Credits / s × 5s × 1080p

Output

Ready for generation

Configure your parameters and click "Run" to see the output here.

--

Public model preview

From one prompt to a complete generated video

Use this model preview to inspect the output format, subject motion, camera movement, and scene continuity.

Text-to-video
Text-to-video5s1080p16:9

Kling 3.0 Omni Text to Video public preview

Create 3-15 second videos at 720p, 1080p, or 4K with square, landscape, vertical, and optional sound output.

Prompt shown with this public preview

A red fox sprints across a snowy ridge at golden sunset and leaps over an icy stream, cinematic side light, detailed fur

Public model preview5s · 1080p · 16:9

Related video models

Choose by input method, output specifications, control surface, and budget.

Camera-focused generation

Kling 3.0 Turbo Text to Video

A controlled 3-15 second route with landscape, square, and vertical output.

Long-form video

Seedance 2.5 Text to Video

Generate up to 30 seconds with seven aspect ratios and synchronized audio.

Premium video and audio

Veo 3.1 Text to Video

A premium route with 4K options, negative prompts, seeds, and generated audio.

Compare AI video APIs

About Kling 3.0 Omni Text to Video

Create 3-15 second videos at 720p, 1080p, or 4K with square, landscape, vertical, and optional sound output.

The route accepts a prompt without requiring an input image. Use a separate image-to-video route when the opening composition must be fixed. kling-3.0-omni/text-to-video

Provider
Kuaishou
Task
Pure text-to-video
Input
Text prompt
Starting price
200 Credits / second

Model specifications

These specifications reflect the current HiAPI Kling 3.0 Omni Text to Video request contract.

Input
Text prompt
Duration
3–15s
Resolution
720p1080p4K
Aspect ratio
16:99:161:1
Audio and references
Optional generated sound switch
API protocol
POST /v1/tasksAsync
Starting tier: 200 Credits / secondBilled by output duration and the live model tier

How to use Kling 3.0 Omni Text to Video

Validate a short shot first, then increase quality or duration after motion and subject consistency are stable.

Subject
Action
Camera
01

Define subject, action, and camera

Kling 3.0 Omni Text to Video API quickstart

This is the minimum request for an asynchronous text-to-video task. Use the API docs for task retrieval, callbacks, errors, and the complete parameter contract.

curl -X POST "https://api.hiapi.ai/v1/tasks" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "kling-3.0-omni/text-to-video",
  "input": {
    "prompt": "A red fox sprints across a snowy ridge at golden sunset and leaps over an icy stream, cinematic side light, detailed fur",
    "duration": 5,
    "resolution": "1080p",
    "aspect_ratio": "16:9"
  }
}'
The create request returns a taskId; retrieve the result asynchronously.Open full API docs
Need to compare models and billing tiers?View live pricing for all models

Frequently asked questions

What type of video model is Kling 3.0 Omni Text to Video?

It is a pure text-to-video route and does not require an opening image.

Which output settings does Kling 3.0 Omni Text to Video support?

3-15 seconds, 720p, 1080p, 4K, and 16:9, 9:16, 1:1.

Can it generate sound or use audio references?

Optional generated sound switch

How is Kling 3.0 Omni Text to Video billed?

Billing follows output duration and the live tier. Resolution, mode, or audio differences are reflected by the page price and task record.

How do I call the Kling 3.0 Omni Text to Video API?

POST an asynchronous task to /v1/tasks with model set to kling-3.0-omni/text-to-video and place parameters inside input.

Lead with subject and setting, then describe action, camera direction, and speed in time order. State no cuts for a continuous take.

Draft

5s · 1080p

Final

15s · 4K

02

Choose duration, quality, and frame

Validate subject motion and camera continuity with 5s · 1080p, then move to 15s · 4K.

Playground
POST /v1/tasks
03

Validate in Playground, then integrate

Check motion stability and subject consistency online, then create an asynchronous task with the same model ID: kling-3.0-omni/text-to-video.