Helios-Distilled — Video Generation AI Model

Helios Distilled turns a text prompt — optionally seeded with an image — into as much as 60 seconds of 24fps video. Runs 2-3× faster than Helios Base for a small quality trade. Fixed 640×384, one to five minutes.

Details

  • CategoryVideo Generation
  • Year2026
  • LicenseApache-2.0

Compliance & Provenance

  • ProviderOpen-source (BestWishYsh) · Specialized
  • EU AI Act RiskLimited Risk
  • Art. 50 TransparencyRequired — AI-generated outputs are marked

Inputs & Outputs

  • TextInput · string

    Text prompt for video generation

  • ImageInput · image · optional

    Optional input image for image-to-video generation

  • VideoOutput · video

    Generated video

Tags

  • video-generation
  • text-to-video
  • image-to-video
  • diffusion
  • distilled
  • fast

Alternatives in Video Generation

  • Cosmos3 Nano

    NVIDIA world model for text- and image-to-video, with optional synced audio.

  • Helios Base

    On-device text-to-video, up to 60s @ 24fps. 3-10 min.

  • MiniMax-H3 Ref2VA

    MiniMax-H3 Ref2VA (omni-reference) generates a 5-15s 24fps video with native stereo audio from a prompt plus a reference image and a reference clip, that clip's own soundtrack included. Video and its soundtrack come out of one denoising loop, so lip movement and sound land in sync. ~144GB bf16 weights — offloaded component-by-component, minutes-scale per clip.

  • Wan2.2 TI2V 5B

    Wan2.2 dense 5B unified text+image-to-video. Text only → 5s video from prompt; text + image → image as the first frame of the video.

Resources