
None
None
Long Story Video Skill
Turn your idea into 10s–10min video—AI writes scripts, prompts, and generates footage automatically.

Ads Video Skill
Generate professional ads and sales videos—AI auto-generates scripts, prompts, and footage.
3D Science Explainer Video Skill
Convert scientific concepts into stunning 3D explain animations
Feedback
freeTrialImage.bannerPity
freeTrialImage.upgradeUnlock
- ✓freeTrialImage.benefitHd
- ✓freeTrialImage.benefitWatermark
- ✓freeTrialImage.benefitUnlimited
Grok Imagine Video 1.5 Lite
Describe a scene in words or animate a photo — Grok Imagine Video 1.5 Lite outputs 1080p video with synchronized sound for roughly $0.04 per clip.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Veo3.1
Create Stunning Videos with Veo3.1
FLUX 3 Video Generator

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator
Why Grok Imagine Video 1.5 Lite Keeps Costs Predictable
The entry point to xAI's 1.5 video line — inexpensive clips that carry their own soundtrack and run on Venice's privacy-first servers.
- The Starting Rung of the 1.5 LinexAI's most economical option in the 1.5 line has served Venice users since September 30, 2026, covering 480p through 1080p and one- to fifteen-second clips, with sound produced alongside the picture.
- Words or Stills — Both Routes CoveredTwo variants ship together: one writes motion from a description, the other animates a still. On Venice either can be reached without special permission, whether you work in the app or call the REST API.
- Nothing Stored, Nothing Trained OnWhatever you type or upload stays in the private tier: no archiving, no profiling, no training use, and no history bound to your account. You settle per clip instead of carrying a recurring SuperGrok plan.
Grok Imagine Video 1.5 Lite in Three Simple Steps
From a written idea or one still photo to a finished render on Venice — three moves cover the whole path.
What Grok Imagine Video 1.5 Lite Brings to the Table
Sound generated with the picture, three resolutions, second-by-second timing and a no-log policy — the practical toolkit inside this lightweight tier.
Sound Produced With the Picture
Ambience, effects and dialogue arrive already matched to the visuals, hitting their marks without a separate mixing step or manual sync work.
Three Resolutions, One-Second Precision
480p, 720p or 1080p can be paired with any length from one to fifteen seconds, adjustable second by second — a level of control rarely offered this cheaply.
The Lowest-Priced Way Into the 1.5 Line
Four cents buys a starting render on Venice, keeping test runs and short social cuts affordable; the total rises only as you add pixels or seconds.
Cleaner Motion, More Believable Weight
xAI's release notes report less warping and more convincing momentum and mass than the earlier generation managed across a clip.
Prompts That Read Camera Language
Directions such as push-in, pan, handheld or crane are understood, and prompts may run to 4,096 characters as long as the subject and action come first.
A Private Tier With No Retention
Your words and images are neither stored, profiled nor fed into training, and nothing links past renders to you — unlike xAI's own apps, which build a personal library.
Common Questions About Grok Imagine Video 1.5 Lite
Pricing, sound, animating photos and how privacy works on Venice — answered below.
How much does a single clip cost, and what changes the price?
Billing is per clip and climbs with both pixels and length: a one-second 480p render costs $0.04, the same length at 720p runs $0.05, and 1080p comes to $0.18, with a fifteen-second ceiling. There is no subscription to sign up for, and a new Venice account starts with a daily free allowance and 500 credits.
Is sound included, and how do I steer it?
Sound is part of the render itself, not a layer added afterwards, which keeps effects, atmosphere and dialogue locked to the action. Compared with the prior generation, speech is clearer and better timed. Just describe the noises you want inside your prompt.
Can I bring an existing photo to life?
That is exactly what the image-to-video mode is for: hand it a photo and it produces one to fifteen seconds of motion at your chosen resolution, soundtrack included. If you would rather start from nothing, the text-to-video mode works from a written description only.
What separates this tier from the flagship Grok Imagine Video 1.5?
Both are private, both reach 1080p and fifteen seconds, and both carry audio. The difference is price and control: this tier keeps costs down for drafts and high-volume jobs, while the flagship accepts up to seven image references plus voice references so a character's face and voice stay consistent between scenes.
Are the weights open, and could I self-host or fine-tune?
The model belongs to xAI and no weights have been released, which rules out hosting it yourself, training on it or inspecting it. Venice's welcome credits let you test it before spending anything, and Wan 2.7 Enhanced remains the nearest openly available option that also produces sound.
What happens to my prompts and uploaded photos?
Everything you submit falls under Venice's private tier. Text and images are not retained on the servers, not analysed for profiles and not sent into any training pipeline. Because no history is tied to your account, past renders cannot be traced back to you — the opposite of xAI's own apps, which catalogue your work.
Make Your First Private Clip With Grok Imagine Video 1.5 Lite
Your words stay unlogged and your uploads never reach a training set. Open a Venice account for 500 welcome credits and produce a first audio-included render before any payment is due.
