Grok Imagine Video 1.5 Lite
Draft a scene in words or animate a photo — output up to 1080p, 15 seconds, with sound baked in.
freeTrialImage.bannerPity
None

None

None

Long Story Video Skill

Turn your idea into 10s–10min video—AI writes scripts, prompts, and generates footage automatically.

Ads Video Skill

Ads Video Skill

Generate professional ads and sales videos—AI auto-generates scripts, prompts, and footage.

3D Science Explainer Video Skill

Convert scientific concepts into stunning 3D explain animations

AI Video Prompt Generator

Feedback

freeTrialImage.bannerPity

freeTrialImage.upgradeUnlock

  • ✓freeTrialImage.benefitHd
  • ✓freeTrialImage.benefitWatermark
  • ✓freeTrialImage.benefitUnlimited

Grok Imagine Video 1.5 Lite

Describe a scene or upload a photo and this lightweight Grok Imagine Video 1.5 Lite tier renders 1080p video with synchronized sound, from $0.04 a render.

All Tools

Discover our comprehensive AI-powered animation toolkit

Inside the Grok Imagine Video 1.5 Lite Release

xAI's most affordable video option on Venice pairs spoken sound with picture in a single pass, at a fraction of flagship pricing.

  • The Entry Point to xAI's 1.5 Video Line
    Launched on Venice on 30 September 2026, this lite release covers 480p through 1080p and runs of 1 to 15 seconds, with sound produced alongside the picture.
  • Pick Your Input: Words or a Picture
    One build starts from a written description, the other animates a photo you supply — and on Venice both are reachable without special access through the app or the REST API.
  • Nothing Kept, Nothing Trained On
    Venice places it in the private tier, so prompts and uploaded stills are never stored, profiled or fed into training, and no history is tied to your identity — you simply pay for each render instead of holding a SuperGrok plan.

Running Grok Imagine Video 1.5 Lite in Three Steps

Go from a written idea or one photo to a completed, sound-ready clip without ever leaving the browser.

Capabilities of the Grok Imagine Video 1.5 Lite Tier

Sound and picture in one pass, flexible clip lengths, budget pricing and privacy by default — the practical strengths this lite video release actually delivers.

Sound Generated With the Picture

Effects, room ambience and spoken dialogue arrive together with the visuals and land on the beat, so there is no second audio pass and no manual sync work.

Every Resolution, One-Second Precision

Pair 480p, 720p or 1080p output with any length between 1 and 15 seconds in single-second steps — a level of control rarely seen at this price point.

The Lowest-Cost Way Into the 1.5 Line

Renders begin at $0.04 on Venice, so drafts and short social cuts stay inexpensive while spending grows only with resolution and length.

Steadier Motion, Believable Weight

xAI's release notes report fewer warping artefacts and more convincing momentum and mass across a clip than the earlier generation produced.

Camera Language Understood

Describe push-ins, pans, handheld moves or crane shots and the model follows; prompts up to 4,096 characters are accepted, with subject and action placed first.

Runs Without Leaving a Trace

Your text and source photos are not stored, profiled or reused for training, and no generation record is bound to you — unlike xAI's own apps, which build an account library.

FAQ

Grok Imagine Video 1.5 Lite: Questions Answered

Pricing per resolution, audio handling, photo animation, privacy and open-source limits explained.

1

How much does a single clip cost?

Pricing is per clip and rises with resolution and length: $0.04 for one second at 480p, $0.05 at 720p and $0.18 at 1080p, scaling up to 15 seconds. There is no subscription, and new Venice accounts receive a daily free allowance plus 500 welcome credits.

2

Is audio included, and how do I steer it?

Yes — sound is produced during the render rather than layered on afterwards, so ambience, effects and speech land on the beat, with clearer and better-timed voices than the prior generation. Simply write the sounds you want into your prompt.

3

Can I animate a photo I already have?

Yes — the image variant turns a still into motion at 480p, 720p or 1080p for 1–15 seconds with sound. A second variant builds a clip from written text alone.

4

How does this lite tier compare with the flagship release?

Both are private on Venice, both reach 1080p and 15 seconds with native audio. This tier is the lower-cost choice for volume work and drafts, while the flagship adds multi-reference control (up to seven image references) and voice references that keep a character's face and voice consistent between scenes.

5

Can I self-host, fine-tune or inspect the model?

No — the weights are proprietary to xAI and have never been published, so self-hosting, fine-tuning and auditing are off the table. You can test it on Venice using welcome credits before paying per clip; Wan 2.7 Enhanced is the nearest open model with built-in audio.

6

What happens to my prompts and uploaded images?

Venice classifies every prompt and uploaded still as private-tier material: nothing is retained on its servers, nothing is profiled and nothing enters a training pipeline. No generation history links back to you, whereas xAI's own apps file your results into an account-based library.

Start Creating With Grok Imagine Video 1.5 Lite

Your words stay unlogged and your uploads never train a model. Claim 500 welcome credits on Venice and produce a first sound-complete clip before you pay anything.