Text-to-Image Generation
Converts plain-language text prompts into high-resolution images. Users describe what they want and receive a generated image in return.
Grok Imagine (grok-imagine-image) is a text-to-image generation model developed by xAI, the AI company founded by Elon Musk. It is part of the Grok Imagine family and allows users and developers to generate high-resolution images from plain-language text descriptions. The model was unveiled in early August 2025 and expanded from a subscriber-only feature to a broadly available tool accessible via xAI's developer API. It sits alongside the more premium grok-imagine-image-pro variant, serving as the standard, faster option in the family. Grok Imagine supports up to 300 requests per minute, making it suited for applications that require image generation at volume. It accepts a 131,072-token context window and is accessible through xAI's API for integration into apps, tools, and workflows. The model is best suited for developers and creators who need reliable, high-throughput image generation for use cases such as prototyping, content creation, or building image-powered products. No artistic expertise is required to use it — a text description is sufficient to produce a detailed image.
High-signal model metadata in a structured two-column overview table.
The entity that provides this model.
The number of tokens supported by the input context window.
The number of tokens that can be generated by the model in a single request.
Whether the model's code is available for public use.
When the model was first released.
When the model's knowledge was last updated.
The providers that offer this model. This is not an exhaustive list.
Types of data this model can process.
A fuller summary of positioning, capabilities, and source-specific details for Grok Imagine.
Grok Imagine (grok-imagine-image) is a text-to-image generation model developed by xAI, the AI company founded by Elon Musk. It is part of the Grok Imagine family and allows users and developers to generate high-resolution images from plain-language text descriptions. The model was unveiled in early August 2025 and expanded from a subscriber-only feature to a broadly available tool accessible via xAI's developer API. It sits alongside the more premium grok-imagine-image-pro variant, serving as the standard, faster option in the family.
Grok Imagine supports up to 300 requests per minute, making it suited for applications that require image generation at volume. It accepts a 131,072-token context window and is accessible through xAI's API for integration into apps, tools, and workflows. The model is best suited for developers and creators who need reliable, high-throughput image generation for use cases such as prototyping, content creation, or building image-powered products. No artistic expertise is required to use it — a text description is sufficient to produce a detailed image.
Converts plain-language text prompts into high-resolution images. Users describe what they want and receive a generated image in return.
Supports up to 300 requests per minute, enabling image generation at scale for production applications.
Available through xAI's developer API, allowing seamless embedding into apps, tools, and automated workflows.
Accepts image URLs as input, enabling workflows that reference or build upon existing images alongside text prompts.
Supports select-type inputs for configuring generation parameters such as image size or style at request time.
Primary API pricing shown in the same “quick compare” spirit as the reference page.
Additional usage-cost dimensions synced into the project for this model.
Places where this model is available, based on the synced detail-page metadata.
The configurable options currently documented for this model.
Parameters currently listed by OpenRouter or the local catalog for this model.
Official model cards, release notes, docs, and other references synced from the source page.
Grok Imagine discussions are most active in r/grok, r/AIJailbreak, r/ChatGPT. Top Reddit threads cluster around benchmark and model-comparison threads, safety and censorship questions, coding workflow discussions.
The strongest match in this snapshot has 2508 upvotes and 419 comments.
You can start with the two models, [Eros](https://huggingface.co/TenStrip/LTX2.3-10Eros), which is better for I2V, and [Sulphur](https://huggingface.co/SulphurAI/Sulphur-2-base/tree/main), which works for both I2V and T2V. If you don't know what any of that means, you've got a long road ahead of you, but I promise it'll be worth it in the end.
This is not an ad and this is not a paid service. You can run this on your PC for free, right now. Just letting ya'll know that you no longer have to bother with Grok. The video I attached below was first attempt that I generated on my PC in <5 minutes.
NSFW warning:
* [Generation result](https://files.catbox.moe/a2nhbx.mp4)
* [Attempt 2](https://files.catbox.moe/q3n8kx.mp4)
EDIT: I've seen a lot of people saying you need a 4090 or 5090 to run LTX, and that's just not true. You can run it on much weaker hardware, the real question is how much you're willing to compromise on speed, resolution, and workflow setup.
For normal use, 12GB of VRAM is a solid baseline. A 3060 12GB or anything better is enough to get started, and people have even managed to run LTX on 8GB cards or lower with quantization and other tricks, but that's more of a technical workaround than something I'd recommend if you want a smooth experience.
RAM matters a lot too, and people keep ignoring that part. I'd treat 32GB as the bare minimum, while 48GB or 64GB is a much better place to be, especially if you don't want your system constantly leaning on pagefile and slowing everything down. If you're using a slow drive, it's even worse.
ComfyUI has also improved a lot here. It can offload parts of the workflow between VRAM and system memory, which is why cards that look too weak on paper can still run models they technically shouldn't fit, just much slower.
So no, you do not need some insane flagship GPU to use LTX. What stronger hardware really buys you is speed and less pain. For reference, I'm on a 5070 Ti and a 10-second 720p video still takes me around 5 minutes to generate.
As a lot of you have already seen and posted, the limits have increased, the moderation has gotten much worse and the prices are exactly the same. Anything beyond a G-rated Disney prompt gets flagged at this point. Like 99% of my prompts are failing, i can't even ask someone to pick their freaking nose w/out it being turned down. It's truly unusable at this point. They they need to go back about 2-3 upgrades or else they're about to lose a whole lot of subs...
Hey everyone,
I've been having a lot of fun with Grok Imagine since it launched, but the moderation has gotten so aggressive lately that it's basically killing the experience. It feels like every other prompt gets blocked or heavily censored for no good reason, even really tame stuff. Super frustrating.
I'm looking for decent **online** AI image to video generators as alternatives (not interested in local setups). I’ve tried Seedance and it’s been pretty solid so far, but I want to see what else is out there.
What are you guys using these days? Bonus points if it has good prompt adherence, decent speed, and isn’t insanely censored.
Drop your recommendations below 👇
You just need to go via Ask section in Grok (not directly via Imagine), select 'edit picture', it takes you to Imagine section and then you can edit your t2v images the same way as it was possible until yesterday.
Hey guys,
I’ve seen a few posts here saying that Grok Imagine on Venice.ai is pretty much uncensored, but I wanted to ask directly to confirm.
Has anyone been testing it lately? Especially with NSFW prompts?
Grok Imagine (grok-imagine-image) has a context window of 131,072 tokens.
The model supports up to 300 requests per minute, making it suitable for high-volume production use cases.
Grok Imagine was unveiled in early August 2025 and subsequently expanded from a subscriber-only feature to a broadly available tool via xAI's public API.
The grok-imagine-image model is the standard variant in the Grok Imagine family, designed for speed and accessibility. The grok-imagine-image-pro variant is the more premium option. Pricing and capability differences are detailed in xAI's official models documentation.
Grok Imagine is available through xAI's developer API. You can find integration details, endpoint references, and pricing in xAI's Image Generation API Reference and Official Pricing & Models Documentation.
Continue browsing adjacent models from the same provider.