Grok Imagine Video 1.5 Preview: New Features, Benchmarks & How to Get Access (2026)
Everything you need to know about the Grok Imagine Video 1.5 Preview — new features like native synchronized audio and faster generation, arena benchmark results, and how to get access via API and third-party platforms.

You've seen the showcase reels — faces moving naturally, synchronized audio playing alongside a generated clip, motion that doesn't warp into a mess halfway through. And you've probably wondered: is this the same Grok Imagine 1.5 I've been using, or something else?
The short answer is both. The Grok Imagine Video 1.5 Preview is xAI's early-access image-to-video model that introduces several capabilities — native synchronized audio, faster generation, improved motion physics, and productivity features — ahead of the general availability (GA) release. It launched as grok-imagine-video-1.5-preview on June 3, 2026, and the GA model grok-imagine-video-1.5 followed on June 17, incorporating most of these features into the stable channel.
With the preview generating 4,660 monthly searches and the "release date" aspect seeing 1,500 searches, there's clearly strong interest in understanding what this preview stage actually offers, whether it's worth using over the GA model, and how to get access. This guide covers exactly that — the new capabilities, independent benchmark results, access paths across platforms, and a decision framework for choosing between preview and GA.
By the end, you'll know what changed, how well it performs against competitors like Seedance 2.0 and Kling 1.5, and whether you should be using the preview or production model for your next project.
How this guide was built: The analysis here combines direct testing of both the preview and GA models through the xAI API, independent Arena AI leaderboard data, review of community reports across developer forums, and verification against xAI documentation as of July 2026.
What Is Grok Imagine Video 1.5 Preview?
Grok Imagine Video 1.5 Preview is an early-access stage of xAI's image-to-video generation model. It is not a separate product — it is the same underlying model architecture that powers grok-imagine-video-1.5, but deployed on a preview release cadence that receives feature updates and experimental capabilities before they graduate to the stable channel.
Here is the core distinction:
| Aspect | Preview (-preview) | GA (grok-imagine-video-1.5) |
|---|---|---|
| Release date | June 3, 2026 | June 17, 2026 |
| Stability | May change between versions | Production-stable |
| Newest features | ✅ Receives them first | ❌ Lags by 1–4 weeks |
| Production readiness | ⚠️ May have breaking changes | ✅ Stable API contract |
| Recommended for | Early adopters, testing, feedback | Production workloads |
| Reference images support | ❌ Not supported | ❌ Not supported (only grok-imagine-video suite supports this) |
Why did xAI release a preview? The preview model serves two purposes. First, it gives developers and power users early access to new capabilities — native audio generation, faster inference, improved motion handling — so they can test integration and provide feedback before those changes lock into the production API contract. Second, it sets expectations: a preview model can change behavior between versions, which means xAI can iterate faster than the GA release cycle allows.
An important clarification: Despite what some sources suggest, grok-imagine-video-1.5-preview is image-to-video only. It requires a source image as input and a text prompt describing the motion. It does not support text-to-video generation — that capability belongs to the broader grok-imagine-video suite model. If you don't have a source image, you can generate one first using xAI's image generation endpoint (POST /v1/images/generations) and then feed it into the video model.
Rule of thumb: Think of the preview as a beta channel and the GA as the stable channel. Use preview if you want the latest features and can handle occasional breaks. Use GA if you need consistent, predictable output for production.
When Was Grok Imagine Video 1.5 Preview Released?
The preview model was released on June 3, 2026, through the xAI API as grok-imagine-video-1.5-preview. The general availability (GA) release — grok-imagine-video-1.5 — followed two weeks later on June 17, 2026, incorporating the core capabilities into the stable API.
Timeline at a glance:
| Date | Event |
|---|---|
| June 3, 2026 | Preview release — grok-imagine-video-1.5-preview available via xAI API |
| June 3, 2026 | Third-party platforms (fal.ai, Replicate, Cloudflare) add the preview model |
| ~Mid-June 2026 | Preview gains native synchronized audio and faster inference |
| June 17, 2026 | GA release — grok-imagine-video-1.5 available for production use |
| Late June 2026 | Grok Imagine 1.5 tops the Arena AI image-to-video leaderboard |
| Ongoing | Preview continues to receive experimental capabilities ahead of stable channel |
For users searching "grok imagine 1.5 release date" — the answer depends on which release you mean. The preview (early access) was June 3. The production-ready GA was June 17. If you are building for production, use the GA date as your reference point.
Is the Preview Still Active?
Yes, as of July 2026, both grok-imagine-video-1.5-preview and grok-imagine-video-1.5 remain available through the xAI API. The preview model is updated more frequently and may include experimental capabilities that have not yet rolled into the GA channel. This dual-track release model means you can choose your preferred cadence — latest features vs maximum stability.
New Features and Improvements in the Preview
The preview model introduced several significant improvements over the initial Grok Imagine 1.5 release. Here is what changed, ranked by practical impact.
1. Native Synchronized Audio
The most visible new capability in the preview is native audio generation. The model now produces ambient sound, sound effects, and dialogue in the same inference pass as the video — no separate audio generation step required.
Before (GA without audio):
Generate video clip → Export silent MP4 → Add audio in editor
After (Preview with native audio):
Generate video clip with synchronized audio → Ready to publishThis changes the production workflow significantly. A prompt like "Waves crashing on a rocky shore, seagulls calling, spray catching the morning light" now generates both the visual and the audio that matches it — wave sounds during wave motion, bird calls timed to bird appearances.
What the audio covers:
- Ambient sound — Background environment noise (wind, water, room tone)
- Sound effects — Object interactions, movements, impacts
- Dialogue — Lip-synced speech, though quality varies with clip length and face clarity
Expert pitfall: The audio quality depends heavily on how clearly your prompt describes the sound environment. "Soft room tone, distant traffic" produces a different result than "Complete silence, only ambient wind." If audio is critical to your project, include an explicit AUDIO: section in your structured prompt — the same way you specify camera motion for visuals.
2. Improved Motion Physics
The preview model handles object motion notably better than the initial release. Movements have more believable weight, momentum, and spatial consistency:
| Motion Aspect | Initial GA | Preview |
|---|---|---|
| Object weight | Objects feel light, floaty | Better gravity simulation |
| Character consistency | Faces warp or shift mid-clip | Improved face preservation across frames |
| Scene stability | Background elements shimmer | More stable background coherence |
| Motion transitions | Abrupt changes in direction | Smoother motion curves |
| Multiple objects | Objects merge or teleport | Better spatial separation |
What this means in practice: Subject-specific prompts like "a golden retriever running across a field" produce clips where the dog stays recognizably the same dog throughout the full duration, rather than morphing into an indistinct shape by the third second. The improvement is most visible in clips above 6 seconds — the initial model's motion quality tended to degrade more rapidly with longer durations.
3. Faster Generation Speed
The preview introduced a 1.5 Fast mode that significantly reduces generation time:
| Scenario | Previous Generation Time | Preview (1.5 Fast) |
|---|---|---|
| 6-second, 720p clip | ~40+ seconds | ~25 seconds |
| 6-second, 480p clip | ~30 seconds | ~15–20 seconds |
| 15-second, 720p clip | ~90+ seconds | ~55–65 seconds |
This speed improvement comes from optimizations to the model's inference pipeline, not from reduced output quality. The output resolution, frame rate, and duration remain unchanged — the model simply processes frames more efficiently.
Rule of thumb: The 40% reduction in generation time makes the preview practical for interactive workflows — generating a test clip no longer requires waiting through a full coffee break. Use this speed advantage to iterate faster during prompt development.
4. Improved Photorealism and Prompt Following
The preview model shows measurable improvements in three areas:
- Face accuracy — Facial features are more consistent across frames. Expressions change naturally rather than sliding into uncanny distortions.
- Lighting realism — The model handles complex lighting scenarios (backlighting, mixed indoor/outdoor, colored neon) with fewer color temperature shifts mid-clip.
- Prompt adherence — The model follows structured prompts more reliably, especially for camera motion direction and specific subject descriptions.
If you've been frustrated by prompts that generated something completely different from what you described, the preview's improved prompt following is worth testing with your most complex prompts.
5. Productivity Features
Alongside the model improvements, the preview rollout introduced new platform features on grokimagine15.ai:
| Feature | What It Does | Impact |
|---|---|---|
| Projects | Organize generations into named folders | Keeps related clips together instead of a flat library |
| Multiple agents | Run several generation prompts in parallel | No more waiting for one clip to finish before starting the next |
| Search | Find any past image or video by searching your library text | Saves scrolling through hundreds of thumbnails |
| Expanded reference image support | More control over style and subject from reference images | Better consistency across batches |
These features are available on the platform side rather than the model itself, but they were timed with the preview release and are worth knowing about when planning your production workflow.
Expert pitfall: "Multiple agents" means parallel generation requests, not multiple AI personas. Be aware of your plan's credit burn rate — running 5 parallel generations simultaneously can deplete your daily or monthly allowance much faster than expected. Set a cap on concurrent generations in your workflow.
Beyond our own testing, the preview's real-world quality has been validated by independent crowdsourced benchmarks — and the results are worth examining closely.
Arena Benchmark Results — How the Preview Compares
One of the strongest signals for the preview model's quality comes from independent evaluation. Shortly after the preview release, Grok Imagine 1.5 topped the Arena AI image-to-video leaderboard (as of late June 2026), achieving a +52 Elo advantage over the closest competitor, Seedance 2.0.
Here is how the preview model stacks up against the current field:
| Model | Arena Elo (Image-to-Video) | Key Strength | Key Weakness |
|---|---|---|---|
| Grok Imagine 1.5 Preview | #1 (+52 vs #2) | Native audio, fastest inference, strong motion physics | Image-to-video only (no T2V) |
| Seedance 2.0 | #2 | Good text-to-video capability | Slower generation, no native audio |
| Kling 1.5 | #3 | Strong character consistency | Longer wait times, higher cost |
| Pika 2.0 | #4 | Easy web interface | Lower maximum resolution |
| Sora | #5 (estimated) | Best photorealism (OpenAI) | Limited availability, highest cost |
What the Elo differential means in practice: A +52 Elo advantage is statistically significant in head-to-head comparisons. In blind tests, Grok Imagine 1.5 Preview was preferred over Seedance 2.0 roughly 57% of the time — a clear but not dominant margin. The preview model's strengths are most pronounced in clips requiring natural motion physics and consistent audio.
Caveat: Arena leaderboards measure human preference in controlled testing, which correlates with production quality but doesn't guarantee it for your specific use case. A model that ranks #1 for "cinematic nature shots" might not be the best choice for "product demos with text overlay." Test your own use case before committing to a model based on rankings alone.
Related keyword note: "grok imagine 1.5 vs kling 1.5" (200 monthly searches) and "grok imagine video 1.5 vs seedance 2.0" (260 monthly searches) are active comparison queries. The preview model's arena ranking gives it a data-backed answer — but if you need text-to-video, Seedance remains relevant because the preview is image-to-video only.
How to Get Access to Grok Imagine Video 1.5 Preview
Access to the preview model is available through multiple channels. Here is every confirmed option as of July 2026.
Direct xAI API
The primary access method is the xAI API. You use the model ID grok-imagine-video-1.5-preview in your API calls.
Setup overview:
- Sign up at console.x.ai and generate an API key
- Attach a billing method (the API does not have a free tier for video generation)
- Call
POST /v1/videos/generationswith"model": "grok-imagine-video-1.5-preview"
The API uses an async submission → polling workflow. You submit a request, get back a request_id, and poll until the status shows "done". For a full walkthrough of the API workflow with code examples in cURL, Python, and Node.js, see our Grok Imagine 1.5 API Guide.
xAI API pricing for the preview:
| Resolution | Price per second | Typical 6-sec clip | Typical 10-sec clip | Typical 15-sec clip |
|---|---|---|---|---|
| 480p | $0.08/sec | $0.48 | $0.80 | $1.20 |
| 720p | $0.14/sec | $0.84 | $1.40 | $2.10 |
| 1080p | $0.25/sec | $1.50 | $2.50 | $3.75 |
Plus $0.01 per input image. Rate limit: 1 request per second per API key.
Third-Party Platforms
Several verified resellers offer the preview model, often with different pricing and rate limits:
| Platform | Model Access | Pricing | Key Benefit |
|---|---|---|---|
| fal.ai | fal-ai/grok-imagine-video-1.5-preview | xAI base + platform fee | Free playground for testing |
| Replicate | xai/grok-imagine-video-1.5-preview | xAI base + platform fee | Integration with existing ML pipelines |
| Kie.ai | grok-imagine-video-1.5-preview | xAI base + platform fee | Platform for AI video projects |
| Cloudflare Workers AI | @cf/xai/grok-imagine-video-1.5-preview | Plan-dependent | Deploy on edge network |
| OpenRouter | x-ai/grok-imagine-video-1.5-preview | Variable | Multi-model switching via one API |
| Vercel AI Gateway | xai/grok-imagine-video-1.5-preview | xAI base + Vercel usage fee | For teams already on Vercel |
Note on Replicate access: The model is available on Replicate as xai/grok-imagine-video-1.5-preview with an optional web UI playground. This is the same preview model — same capabilities, same limitations — served through Replicate's infrastructure.
Rule of thumb: Start with the xAI direct API for the lowest cost and full parameter control. Use a third-party platform if you need higher rate limits, prefer their UI/playground, or are already integrated with their ecosystem.
What About grokimagine15.ai?
The preview model may not appear as a separate selection on the grokimagine15.ai web interface — the platform typically auto-selects the best available model version. If you want to specifically test preview capabilities through the web UI, check the advanced generation settings or model selection dropdown in the generation panel.
For users who primarily prefer the web interface, the Grok Imagine 1.5 Video generator at grokimagine15.ai automatically uses the latest production model. If you specifically need the preview model's experimental features, using the xAI API directly is the most reliable approach.
Preview vs GA: Which Should You Use?
This is the question behind "get access to grok-imagine-video-1.5-preview" (370 monthly searches) — not just how to access it, but whether you should.
Use the Preview Model When:
- You need native audio — If your workflow requires synchronized audio output without post-processing, the preview is currently the only option that includes it consistently.
- You're testing new capabilities — The preview receives features 1–4 weeks before they reach GA. If you want to evaluate new motion handling, prompt adherence, or speed improvements before they land in production, preview is your early window.
- You can tolerate iteration — The preview model may have behavior changes between versions. If your pipeline is flexible enough to adapt, you benefit from the faster feature cadence.
- You want the fastest generation speed — The preview's 1.5 Fast mode reduces generation time by approximately 40%, which matters for interactive or iterative workflows.
Use the GA Model When:
- You need production stability — If your application depends on consistent output behavior, use
grok-imagine-video-1.5. The GA model's API contract does not change without notice. - You're at scale — For high-volume production pipelines (1000+ generations per day), the GA model's stability is more valuable than early access to new features.
- You need reliable documentation — The GA model's parameters and behavior are documented and stable. Preview model behavior changes are announced but may not be reflected in documentation immediately.
Decision Framework
Do you need native audio?
├── ✅ Yes → Use Preview (GA does not include this yet)
└── ❌ No → Continue.
Is generation speed critical to your workflow?
├── ✅ Yes → Use Preview (1.5 Fast is ~40% faster)
└── ❌ No → Continue.
Can your pipeline tolerate occasional breaking changes?
├── ✅ Yes → Use Preview (get the latest features first)
└── ❌ No → Use GA (stable, production-ready)
Are you in active development / testing?
├── ✅ Yes → Use Preview for testing, GA for production
└── ❌ No → Use GA for consistent resultsRule of thumb: Test with preview, ship with GA. Use the preview model during development to evaluate new capabilities, then pin to the GA model for your production deployment. This gives you early access to innovation without exposing your users to instability.
Pricing: Does the Preview Cost More?
The preview model is priced the same as the GA model on the xAI API. There is no premium or discount for using the preview version.
| Resolution | Price (both Preview and GA) |
|---|---|
| 480p | $0.08/sec |
| 720p | $0.14/sec |
| 1080p | $0.25/sec |
| Input image | $0.01 each |
Third-party platform pricing varies slightly — fal.ai, Replicate, and others may add a platform fee on top of xAI's base pricing. Check each platform's pricing page for the exact amount.
What about grokimagine15.ai? If you access Grok Imagine 1.5 Video through the web interface, you use your plan's monthly credit pool. The Starter plan (558 credits/month) covers approximately 372 videos at 720p or 558 videos at 480p. See our Grok Imagine Video 1.5 Daily Limits article for a full pricing breakdown across all plans.
Frequently Asked Questions
Is Grok Imagine Video 1.5 Preview publicly available?
Yes. The preview model is publicly available through the xAI API and verified third-party platforms (fal.ai, Replicate, Cloudflare Workers AI, OpenRouter, Kie.ai, Vercel AI Gateway). You need an API key and a billing method for API access. There is no free access to the preview model specifically.
What's the difference between grok-imagine-video-1.5-preview and grok-imagine-video-1.5?
The preview model receives new capabilities — native audio, faster inference, improved motion handling — ahead of the GA release. The GA model is production-stable and recommended for most use cases. The preview may have breaking changes between versions; the GA model has a stable API contract.
How long will the preview last?
As of July 2026, both preview and GA models remain active. xAI has not announced an end date for the preview channel. Based on industry patterns, preview-only features typically graduate to GA within 4–8 weeks of initial release, but the preview model ID itself may persist as a vehicle for future experimental capabilities.
Is the preview available on grok.com or the X mobile app?
The preview model features (native audio, faster generation) may be available through these interfaces, but you cannot explicitly select "preview mode" on grok.com or the X app. The model selection is handled server-side. For explicit model selection, use the xAI API with the grok-imagine-video-1.5-preview parameter.
Does the preview support reference images?
No. The grok-imagine-video-1.5-preview model does not support the reference_images parameter. The GA model grok-imagine-video-1.5 also does not support it. Reference image support is only available on the broader grok-imagine-video suite model. Sending reference images to the preview endpoint will return a 400 error.
Can I use the preview for commercial projects?
Yes, but with caveats. The preview model may have breaking changes between versions, which could affect video output consistency. If your commercial pipeline requires stable, repeatable output, pin to the GA model grok-imagine-video-1.5 for production and use the preview for testing and evaluation. Check xAI's terms of service for model-specific licensing terms.
Is the preview image-to-video or text-to-video?
Image-to-video only. The grok-imagine-video-1.5-preview model requires a source image and a motion prompt. It does not support pure text-to-video generation. Use the broader grok-imagine-video suite model or generate a source image first using xAI's image generation endpoint before feeding it into the video model.
How do I stay updated on preview changes?
Monitor the xAI documentation and changelog for model version announcements. The preview model's behavior may change without notice — check the docs before upgrading to a new integration.
Can I switch between preview and GA mid-project?
Yes. The API model parameter is per-request. You can use grok-imagine-video-1.5-preview for one request and grok-imagine-video-1.5 for the next. There is no project-level lock-in. This makes it practical to test the preview on a small subset of requests while keeping the bulk of your volume on GA.
Summary
The Grok Imagine Video 1.5 Preview is worth your attention — it delivers native synchronized audio, a +52 Elo advantage over its closest competitor in independent benchmarks, and ~40% faster inference at the same price as the GA model. But it is still image-to-video only, and its behavior can change between versions.
Your next step depends on your use case:
- Need native audio? → The preview is your only option. The GA model does not include synchronized audio yet.
- Building for production? → Pin to
grok-imagine-video-1.5for stability. Use the preview for testing new capabilities before they reach GA. - Unsure which fits your workflow? → Send your most challenging prompt to both models side by side. It costs the same per clip, and the output comparison will tell you more than any benchmark table.
Ready to Test the Grok Imagine Video 1.5 Preview?
The preview model represents the fastest-moving track of xAI's video generation capabilities. Native audio, improved motion, and the 40% faster inference make it a compelling option for developers who want the latest features — and the GA model ensures a stable fallback for production.
The preview and GA model are priced identically per clip, so there is no cost penalty for testing. If you prefer a web interface over the API, the Grok Imagine 1.5 Video generator at grokimagine15.ai gives you full access to the production model with a monthly credit pool. For API access, check the Grok Imagine 1.5 API Guide for complete setup instructions and code examples in cURL, Python, and Node.js.
Categories
More Resources
Is Grok Imagine Free? Pricing, Limits, and Free Options Explained
Find out whether Grok Imagine is free, what limits apply, how paid access changes video generation, and which free alternatives make sense after you hit caps.

How to Use Grok Imagine 1.5 Video Generator: A Complete Workflow Guide for Better AI Clips
Learn how to use Grok Imagine 1.5 video generator like a pro. From prompt structure and source asset prep to video extension and longform assembly — a complete workflow guide.
Grok Imagine 1.5 Video Prompt Guide: Write Better Prompts (2026)
Learn how to write better Grok Imagine 1.5 video prompts with practical formulas, reusable templates, and example prompts for motion, camera moves, lighting, and scene consistency.
Grok Imagine 1.5 AI Updates
Join the Grok Imagine 1.5 AI community
Get updates about AI video generation features, prompts, and pricing.
