
None
None
Long Story Video Skill
Turn your idea into 10s–10min video—AI writes scripts, prompts, and generates footage automatically.

Ads Video Skill
Generate professional ads and sales videos—AI auto-generates scripts, prompts, and footage.
3D Science Explainer Video Skill
Convert scientific concepts into stunning 3D explain animations
Feedback
freeTrialImage.bannerPity
freeTrialImage.upgradeUnlock
- ✓freeTrialImage.benefitHd
- ✓freeTrialImage.benefitWatermark
- ✓freeTrialImage.benefitUnlimited
Grok Imagine Video 1.5 Lite
Feed it a written scene or a photograph and Grok Imagine Video 1.5 Lite returns 1080p footage of up to 15 seconds, sound included, billed per render on Venice.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Veo3.1
Create Stunning Videos with Veo3.1
FLUX 3 Video Generator

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator
Grok Imagine Video 1.5 Lite — Value Without Compromise
This is the entry-priced member of xAI's Grok Imagine Video 1.5 lineup: affordable renders with audio baked in, executed privately on Venice.
- One Member of the Grok Imagine Video 1.5 LineupLive on Venice since September 30, 2026, it covers 480p through 1080p across 1–15 second spans and lays sound down during the same render.
- A Variant for Every Starting PointBegin from typed words or from a still picture; Venice exposes both modes, and neither needs a permission gate — use the web app or call the REST API.
- Runs Private, Keeps NothingYour prompts and source photos stay out of storage, out of profiling and out of any training set, and no render log is tied to your identity — you pay by the clip instead of holding a SuperGrok plan.
How Grok Imagine Video 1.5 Lite Works
From a typed idea or one photograph to a completed Venice render in three moves.
Capabilities Inside Grok Imagine Video 1.5 Lite
Sound rendered in the same pass, resolutions from 480p to 1080p, second-by-second length control and nothing retained server-side — the practical shape of this tier.
Audio Rendered in the Same Pass
Room tone, effects and spoken dialogue arrive locked to the picture and land on the beat, so there is no separate mixing stage and no manual sync work afterward.
Every Resolution, One-Second Precision
Any length between one and fifteen seconds can be paired with 480p, 720p or 1080p output, and you set it in single-second increments — rare control for the money.
The Lowest-Priced Way Into the 1.5 Line
Four cents buys a first render on Venice, which keeps rough drafts and short social cuts cheap while the total still tracks resolution and length.
Steadier Motion, More Believable Weight
xAI's own release notes report fewer visual warps and more convincing mass and momentum through a shot when compared with the earlier version.
Prompts That Understand Camera Language
Directions such as crane, handheld, pan or push-in are interpreted rather than ignored, and descriptions of up to 4,096 characters are accepted when subject and action come first.
Nothing Kept, Nothing Trained On
Text and uploaded frames are never archived, profiled or fed into model training, and no render history attaches to you — unlike xAI's own apps, which build a library.
Common Questions About Grok Imagine Video 1.5 Lite
Answers on cost, sound, animating a photo and what stays private on Venice.
How much do renders cost across resolutions and lengths?
Each render is priced on its own and rises with quality and length: four cents for one second at 480p, five cents at 720p, eighteen cents at 1080p, capped at fifteen seconds. No subscription, and new Venice accounts get 500 credits plus a daily free allowance.
Is audio generated, and can I direct it?
It is. Sound is produced inside the same pass instead of being layered on afterward, so ambience, effects and dialogue land on the beat, and speech is cleaner and better timed than the earlier version managed. Simply describe the noises you want in your text.
Will it bring an existing photo to life?
That is exactly what the photo-driven mode does: it turns a still into moving footage at 480p, 720p or 1080p for one to fifteen seconds, sound included. A companion mode instead builds everything from written description alone.
Where does this edition sit relative to the flagship?
Both execute privately on Venice and both reach 1080p, fifteen seconds and built-in sound. This edition costs less and suits volume work; the flagship layer adds multi-reference control (up to seven images) and voice references that keep a character consistent across scenes.
Can I self-host it, audit it or fine-tune it?
The weights are xAI's alone and have never been published, which rules out hosting it yourself, tuning it or inspecting it. Welcome credits on Venice let you test it before any charge, and Wan 2.7 Enhanced stands as the nearest openly available option that also produces sound.
What happens to the prompts and images I upload?
Venice classifies everything you submit as private-tier material — nothing sits on the servers, nothing is analysed for profiling and nothing enters a training pipeline. No record of your renders points back at you, while xAI's own apps instead file each output into an account library.
Generate Privately With Grok Imagine Video 1.5 Lite
Nothing you type is logged and nothing you upload trains a model — open Venice, collect 500 welcome credits and produce a first audio-complete render before any charge.
