Helios-Distilled — Video Generation AI Model
Helios Distilled turns a text prompt — optionally seeded with an image — into as much as 60 seconds of 24fps video. Runs 2-3× faster than Helios Base for a small quality trade. Fixed 640×384, one to five minutes.
Details
- CategoryVideo Generation
- Year2026
- LicenseApache-2.0
Compliance & Provenance
- ProviderOpen-source (BestWishYsh) · Specialized
- EU AI Act RiskLimited Risk
- Art. 50 TransparencyRequired — AI-generated outputs are marked
Inputs & Outputs
- TextInput · string
Text prompt for video generation
- ImageInput · image · optional
Optional input image for image-to-video generation
- VideoOutput · video
Generated video
Tags
Alternatives in Video Generation
- Cosmos3 Nano
NVIDIA world model for text- and image-to-video, with optional synced audio.
- Helios Base
On-device text-to-video, up to 60s @ 24fps. 3-10 min.
- MiniMax-H3 Ref2VA
MiniMax-H3 Ref2VA (omni-reference) generates a 5-15s 24fps video with native stereo audio from a prompt plus a reference image and a reference clip, that clip's own soundtrack included. Video and its soundtrack come out of one denoising loop, so lip movement and sound land in sync. ~144GB bf16 weights — offloaded component-by-component, minutes-scale per clip.
- Wan2.2 TI2V 5B
Wan2.2 dense 5B unified text+image-to-video. Text only → 5s video from prompt; text + image → image as the first frame of the video.