Standard Grok Imagine generation and natural-language image editing for routine iterations.
Brands
xAI
Grok Imagine models for expressive image and video generation.
Models
Image
Grok Imagine Image
View details →Grok Imagine generates images from prompts and can edit or combine source images using natural-language directions. The Quality variant adds 1K and 2K output choices and is the better fit when resolution matters, while the standard entry is suitable for ordinary iteration. For edits, keep each reference's role explicit so the model can distinguish subject, style, and composition sources.
Higher-resolution Grok Imagine generation and editing with 1K and 2K output choices.
Models
Video
Grok Imagine Video
View details →The main Grok Imagine Video model can generate a clip from text, animate an image, incorporate reference images, edit a source video, or continue a video from its end. These modes use different inputs and constraints, while Grok Imagine Video 1.5 in Omnipix is the direct image-to-video path and should not be presented as a replacement for every workflow of the main model.
General Grok video model for text, image, visual-reference, edit, and extension workflows.
Focused Grok Imagine Video 1.5 variant for animating a supplied image.