Novita

Kuaishou: Kling v3.0 Standard T2V

Kling/kling-v3.0-std-t2v
Context

10K

Input

text

Output

video

Tool calling

Not listed

About this model

Kling v3.0 Standard T2V is Kuaishou's text-to-video generation model designed to create cinematic video directly from natural-language prompts. It supports smooth motion, strong prompt adherence, multi-shot composition, negative prompting, configurable aspect ratios, flexible video durations, and optional native audio generation for synchronized audiovisual content.

Input modalities

text

Accepted as model input

Providers

Available routing options for this model through OneInfer.

Available

Video generation

$0.0840 / second

Video generation

$0.1260 / second

Routing

OneInfer optimized

Pricing

Current OneInfer pricing for this model.

UsagePrice
Video generation720p · 3–15s · No audio$0.0840 / second
Video generation720p · 3–15s · With audio$0.1260 / second

Benchmarks

Published evaluation results associated with this model.

No benchmark data is listed for this model.

API Guide

Refer Documentation
  1. 1

    Generate your access token

    Use your API key to generate the JWT access token required by the Models API.

  2. 2

    Add the JWT token

    Copy the generated JWT token and replace JWT_TOKEN in the example below.

  3. 3

    Check your credits

    If your balance is too low, before calling the model.

  4. 4

    Run the API example

    Choose your preferred language, copy the example, and send your first model request.

    curl -X POST https://api.oneinfer.ai/v1/ula/generate-video \
      -H "Authorization: Bearer YOUR_JWT_TOKEN" \
      -H "Content-Type: application/json" \
      -d '{
      "provider": "novita",
      "model": "Kling/kling-v3.0-std-t2v",
      "prompt": "A red fox running through a snowy forest",
      "resolution": "720p",
      "aspect_ratio": "16:9",
      "duration": 5,
      "generate_audio": false,
      "camera_fixed": false,
      "service_tier": "default"
    }'