Grok Imagine
Generative image and video creation from xAI.
xAI's Grok-based generative-media product and API for creating and editing images and generating video.

Recent stories
A game developer documents a Grok Bot workflow that creates consistent game art asset sheets with Grok Imagine, evaluates results, and exports transparent PNGs. Separate bots summarize Slack activity and send bug reports to a coding agent for reviewable fixes.
Creators posted Grok Imagine Image 2.0 demos for object-segment edits, reference blending, text rendering, resizing, and background removal. The samples focus on posters, covers, product-style images, and diagrams.
Grok Imagine Video 1.5 is now available inside Leonardo, and side-by-side tests across five prompts put Seedance ahead on quality while Grok stayed faster and cheaper. Try the shared access point and prompt set if you want to compare output, speed, and cost yourself.
Grok Imagine Video 1.5 is now live with general API access, and Video 1.5 Fast is rolling out to consumers. xAI says the update cuts 720p render time to about 25 seconds from 40-plus and improves realism and physics.
xAI opened Grok Imagine 1.5 Preview in its Imagine API, moving the model from benchmark chatter into direct creator access. The same-day Cloudflare AI Gateway support gives teams another route to run Grok models in production.
fal added Grok Imagine Video 1.5, and creator posts immediately tested it against Seedance 2.0 and Gemini Omni on fight scenes, lip-sync, and reference-driven clips. The early comparisons put it into the serious creator model mix, but not clearly ahead of Seedance in real-world use.
Posts reported Grok Imagine Agent Mode going live on the web as an open-canvas creative agent, with demos showing brand ideation inside one workspace. The change matters because Grok is moving from single-prompt turns toward iterative visual brainstorming on a persistent canvas; watch the beta for workflow limits.
Creator posts say Grok Imagine's video update can make one-shot clips with spoken audio, stronger lip sync and support for multiple speakers, pets and varied face angles. The demos also show selfie-to-scene transforms and timeline prompting, but the rollout is documented mainly through independent testing.