Novita

Alibaba: Wan 2.6 T2V

alibaba/wan2.6-t2v
Context

1K

Input

text

Output

video

Tool calling

Not listed

About this model

Wan 2.6 T2V is Alibaba's text-to-video generation model designed to create high-quality cinematic video directly from natural-language prompts. It supports complex scene descriptions, realistic motion and physical dynamics, single-shot and multi-shot generation, prompt enhancement, negative prompting, multiple aspect ratios, high-resolution video generation, and optional native audio for synchronized audiovisual output.

Input modalities

text

Accepted as model input

Providers

Available routing options for this model through OneInfer.

Available

Video generation

$0.1000 / second

Video generation

$0.1500 / second

Routing

OneInfer optimized

Pricing

Current OneInfer pricing for this model.

UsagePrice
Video generation720p · 5–15s · No audio$0.1000 / second
Video generation1080p · 5–15s · No audio$0.1500 / second

Benchmarks

Published evaluation results associated with this model.

No benchmark data is listed for this model.

API Guide

Refer Documentation
  1. 1

    Generate your access token

    Use your API key to generate the JWT access token required by the Models API.

  2. 2

    Add the JWT token

    Copy the generated JWT token and replace JWT_TOKEN in the example below.

  3. 3

    Check your credits

    If your balance is too low, before calling the model.

  4. 4

    Run the API example

    Choose your preferred language, copy the example, and send your first model request.

    curl -X POST https://api.oneinfer.ai/v1/ula/generate-video \
      -H "Authorization: Bearer YOUR_JWT_TOKEN" \
      -H "Content-Type: application/json" \
      -d '{
      "provider": "novita",
      "model": "alibaba/wan2.6-t2v",
      "prompt": "A red fox running through a snowy forest",
      "resolution": "720p",
      "aspect_ratio": "16:9",
      "duration": 5,
      "generate_audio": false,
      "camera_fixed": false,
      "service_tier": "default"
    }'