AI MODEL DISCOVERY

TypeSafe Jev 1.13 model preview
TypeSafe

TypeSafe Jev 1.13

Structured decisions from text or JSON, with yes/no probabilities, choices and scores.

Model release timeline

  1. MiniMax H3 Max 16-bit Pixel model preview

    MiniMax H3 Max 16-bit Pixel

    MiniMax

    Generate 5–15 second 16-bit pixel-art videos with audio at 768P from text or a first-frame image.

  2. Wan 3.0 Prime model preview

    Wan 3.0 Prime

    Alibaba

    Alibaba Wan 3.0 Prime video generation with text, image, video, audio, document, and webpage references.

  3. Wan 3.0 model preview

    Wan 3.0

    Alibaba

    Alibaba Wan 3.0 video generation with text, image, video, audio, document, and webpage references.

  4. MiniMax H3 Max Turbo model preview

    MiniMax H3 Max Turbo

    MiniMax

    Generate 5–15 second videos from text or images at 480P, 768P, or 1080P.

  5. Seedance 2.5 model preview

    Seedance 2.5

    ByteDance

    ByteDance's next-generation video model for long-form storytelling with up to 50 multimodal references

  6. MiniMax H3 model preview

    MiniMax H3

    MiniMax

    MiniMax H3 (Hailuo 03) multimodal native-audio video generation with output up to 2K.

  7. Seedance 2.0 Mini model preview

    Seedance 2.0 Mini

    ByteDance

    ByteDance's cost-efficient Seedance 2.0 video model for faster generation and lower inference costs

  8. Seedance 2.0 Fast model preview

    Seedance 2.0 Fast

    ByteDance

    Fast & affordable multi-modal video with text, image, video & audio inputs up to 720p

  9. HappyHorse 1.1 model preview

    HappyHorse 1.1

    Alibaba

    HappyHorse 1.1 supports text-to-video, image-to-video, and reference-to-video generation with expanded aspect ratio options.

  10. Wan 2.7 model preview

    Wan 2.7

    Alibaba

    Alibaba's Wan 2.7 Video API supports full-modality inputs (text, image, video, audio) for four modes (T2V, I2V, Reference2V, Edit), delivering 720P–1080P outputs.

Turn a Written Scene into a Model Brief

Choose a Text to Video API around the shot you need, not just the model name. SeeAPI brings published text to video models into one catalog so you can explore visual styles, inspect supported tasks and continue to the relevant documentation. Start with a clear scene, then compare the controls that matter to your application.

Direct the Action, Not Just the Appearance

Direct the Action, Not Just the Appearance

A useful brief describes a subject doing something in a setting. Add camera direction only where it serves the shot. For a product reveal, for example, ask for one slow rotation with a stable camera before attempting multiple cuts. Evaluate the full sequence for object shape, readable movement and unwanted changes.

Compare Models Against the Same Deliverable

Compare Models Against the Same Deliverable

Keep the scene objective consistent while adapting requests to each documented endpoint. Compare duration, aspect ratio and resolution where supported, then review how each result handles motion and composition. Check the selected configuration on API Pricing; a different duration or output setting can change the price.

Decide Whether the Scene Needs Sound

Decide Whether the Scene Needs Sound

Some video entries also support audio generation. Open the Video with Audio category when dialogue, ambience or sound effects are part of the deliverable. Confirm the selected endpoint's audio settings rather than assuming every text-generated clip includes a soundtrack.

From Scene Brief to API Request

Use the catalog to shortlist capabilities before building a larger video workflow.

1

Write One Complete Shot

Specify the subject, action, setting and intended framing. Separate requirements from optional styling so you can judge a result consistently.

2

Read the Selected Endpoint

Open a model card and check the published request fields, supported output settings and task retrieval instructions. Use its example as the starting point.

3

Review Before Scaling

Run an authorized test in your own integration, save the request settings and inspect the complete clip. Expand only after the output meets your application's acceptance criteria.

Text to Video API: Questions Before You Integrate

Practical answers for choosing a model and preparing your first request.

Its core input is a written prompt describing a moving scene. The selected endpoint determines additional fields, defaults and limits. This category focuses on generating from text; use Image to Video when an approved starting image is the main visual reference.