On 3 June 2026, xAI announced Grok Imagine Video 1.5 and Video 1.5 Fast. The model animated a starting image from natural-language direction and later added a faster checkpoint and multi-reference control.
This release brief was checked against first-party material on 10 August 2026. The date above is the public announcement date, not the date a repository was created or a third-party provider added the model. Where access or weights arrived later, that distinction is recorded below.
Release record
| Field | Verified detail |
|---|---|
| Announcement | 3 June 2026 |
| Availability or weight release | 3 June API preview; 16 June general availability; reference-image update on 31 July |
| Release type | proprietary image-to-video generation model family |
| Access | xAI Imagine API and consumer Grok surfaces; no open weights |
| Architecture | image-conditioned video generation with jointly generated motion, physics, ambience, dialogue and sound effects |
| Maximum stated context | up to 10-second clips in the disclosed API example; no token context applies |
What changed
The preview established a new named xAI video model, and the June general release materially improved motion, physics, audio and throughput. The July references update belongs to the same model story rather than a duplicate article. Grouping the three dates preserves the earliest announcement and the later access changes.
The practical comparison is therefore not simply whether Grok Imagine Video 1.5 and Video 1.5 Fast has the largest headline score. Teams need to ask whether its architecture, access terms, latency, tool behaviour and evaluation setup match the workload they actually intend to run. A model can lead one harness while losing on cost, refusal behaviour, multilingual quality or repeatability in another.
Benchmarks worth retaining
| Evaluation | Reported result | How to read it |
|---|---|---|
| Fast generation | about 25 seconds for a six-second 720p clip | Down from more than 40 seconds for the previous model |
| Maximum disclosed resolution | 720p | API capability at preview and launch |
| Reference inputs | reference-image workflow added 31 July | Grouped capability update to the same model family |
These are release-time results, not independently reproduced guarantees. Most quality claims are first-party and the 25-second speed figure is for a six-second 720p Fast generation, not every resolution, duration or load condition. Scores should remain attached to the disclosed effort setting, agent harness, tool access, timeout, context-management policy and judge model. Moving a number into a procurement sheet without those conditions creates false comparability.
Architecture and access
Grok Imagine Video 1.5 and Video 1.5 Fast is described as image-conditioned video generation with jointly generated motion, physics, ambience, dialogue and sound effects with up to 10-second clips in the disclosed API example; no token context applies of stated context. Its access position at verification time is xAI Imagine API and consumer Grok surfaces; no open weights. That wording matters: open weights, source-available weights, an API, a product preview and a research demonstration give adopters very different rights and different levels of reproducibility.
Before deployment, record the exact model identifier or checkpoint, inference stack, quantisation, reasoning setting, region, price schedule and supplier terms. If the release uses a custom licence, read the licence itself rather than relying on the word “open” in launch copy. If it is API-only, preserve the dated documentation and change-notice route because the served snapshot can change without a downloadable artefact.
What an evaluation should test next
For Grok Imagine Video 1.5 and Video 1.5 Fast, a credible internal gate should include:
- a frozen set of representative tasks with pass, fail and abstain criteria;
- a matched baseline using the same tools, timeout, prompt budget and reviewer rubric;
- repeated runs to expose variance rather than reporting a single best attempt;
- latency, token use and total task cost alongside task success;
- adversarial, multilingual and long-context cases relevant to the real deployment; and
- rollback evidence showing the previous model can be restored safely.
The wider model change-control guide explains how to keep model, prompt, tool and corpus changes reconstructable. The AI dependency inventory guide covers the release and supplier records needed after deployment.
AIEngine verdict
Imagine Video 1.5 is a major media-model release. Creative teams should compare identity preservation, temporal stability, audio synchronisation and end-to-end queue time on fixed reference sets.
This is a launch assessment, not a certification. Benchmark leadership is useful evidence of where to test; it is not authorization to place the model in a high-impact workflow without domain evaluation, security review and an accountable owner.
Primary sources
- xAI Imagine Video 1.5 preview
- xAI Imagine Video 1.5 general release
- xAI Imagine Video 1.5 references update
Image provenance
Hero image: xAI official release artwork. The locally served WebP is a crop of the first-party release or model-card asset recorded in the repository provenance manifest.



