
grok-imagine-imageGrok Imagine Image is xAI Grok's image model, focused on synchronous text-to-image and image editing (19s median in production) with strong in-image text rendering and an unusually wide aspect-ratio set (1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 2:1, 1:2, 19.5:9, 9:19.5, 20:9, 9:20, auto) — square, standard portrait/landscape and ultra-wide all covered. It is the speed-and-cost standard channel; for higher precision and finer detail with multi-round optimization, step up to Grok Imagine Image Pro. A common pattern is to draft at volume here and refine key images on Pro.
Generated image will appear here
Enter a prompt and click Generate
Generate from text descriptions — 19s median, 22s p90 in production
Pass an input image to edit; aspect_ratio is ignored in edit mode
Legible typography inside images for posters and graphics
1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 2:1, 1:2, 19.5:9, 9:19.5, 20:9, 9:20, auto
Grok Imagine Image is a Image Generation API provided by xAI. Grok Imagine Image is xAI Grok's image model, focused on synchronous text-to-image and image editing (19s median in production) with strong in-image text rendering and an unusually wide aspect-ratio set (1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 2:1, 1:2, 19.5:9, 9:19.5, 20:9, 9:20, auto) — square, standard portrait/landscape and ultra-wide all covered. It is the speed-and-cost standard channel; for higher precision and finer detail with multi-round optimization, step up to Grok Imagine Image Pro. A common pattern is to draft at volume here and refine key images on Pro. Through APIMODELS platform, you can access this model via a unified API with transparent pay-as-you-go pricing. Current pricing: per image: $0.03.
Generate high-quality product visuals for online stores, ads, and marketing materials.
Create eye-catching visual content for social platforms to boost engagement and brand visibility.
Produce concept art for characters, scenes, and props to accelerate game development.
Design posters, banners, and promotional graphics at a fraction of traditional design costs.
Grok Imagine Image is available through APIMODELS at: per image: $0.03. Billing is pay-as-you-go — you only pay for what you generate.
Sign up at APIMODELS, get your API key, and call our unified API endpoint. We provide detailed API documentation with code examples in cURL, Python, and Node.js.
APIMODELS offers the same Grok Imagine Image model through our aggregation platform. We provide a unified API interface so you do not need separate accounts for each provider - one API key to access all models.
Grok Imagine Image is xAI Grok's image model, focused on fast (~4s) text-to-image and image editing with strong in-image text rendering and an unusually wide aspect-ratio set (1:1, 16:9, 9:16, 4:3, 3:4, 3:2, 2:3, 2:1, 1:2, 19.5:9, 9:19.5, 20:9, 9:20, auto) — square, standard portrait/landscape and ultra-wide all covered.
It does both text-to-image and image editing, with a dozen-plus aspect ratios from square through standard portrait/landscape to ultra-wide 20:9 — handy for social, banners, wallpapers and posters. Generation is fast (~4s) with solid text rendering.
The standard channel (grok-imagine-image) is built for speed and low cost — good for everyday batch generation. The Pro channel (grok-imagine-image-pro) has stronger understanding and finer detail for higher-precision work and multi-round optimization. A common pattern is to draft at volume on standard, then refine key images on Pro.
On APIMODELS, Grok Imagine Image runs alongside 60+ models on one API key and one balance, so choosing is about fit, not lock-in. It supports Text to Image, Image Editing, Wide Aspect Ratio Set, Synchronous, No Polling, Text Rendering, and you can weigh it on price and capability against other Image Generation models, then switch by changing a single model-name string — no new account or integration. Browse every Image Generation option with live pricing at apimodels.app/models.
Grok Imagine Image supports: Text to Image, Image Editing, Wide Aspect Ratio Set, Synchronous, No Polling, Text Rendering. See the APIMODELS docs for full parameters and call examples.
Yes. APIMODELS exposes Grok Imagine Image through a single unified API and one key — no separate provider accounts, and no need to handle each provider's regional network access yourself.
We support Stripe (Visa, Mastercard, and other international cards) and Alipay. Credits are available instantly after payment.
Prompts shared by their authors — copy and adapt them. Each one credits its author and links back to the original post.

Weathered sailor on a fishing boat
Create a photorealistic candid photograph of an elderly sailor standing on a small fishing boat. He has weathered skin with visible wrinkles, pores, and sun texture, and a few faded traditional sailor tattoos on his arms. He is calmly adjusting a net while his dog sits nearby on the deck. Shot like a 35mm film photograph, medium close-up at eye level, using a 50mm lens. Soft coastal daylight, shallow depth of field, subtle film grain, natural color balance. The image should feel honest and unposed, with real skin texture, worn materials, and everyday detail. No glamorization, no heavy retouching.
by OpenAI

Automatic coffee machine workflow infographic
Create a detailed Infographic of the functioning and flow of an automatic coffee machine like a Jura. From bean basket, to grinding, to scale, water tank, boiler, etc. I'd like to understand technically and visually the flow.
by OpenAI

Thread streetwear ad with exact typography
Give me a cool in culture ad / fashion shot for a brand called Thread. It's a hip young street brand. The ad shows a group of friends hanging out together with the tagline "Yours to Create." Make it feel like a polished campaign image for a youth streetwear audience: stylish, contemporary, energetic, and tasteful. Use clean composition, strong color direction, natural poses, and premium fashion photography cues. Render the tagline exactly once, clearly and legibly, integrated into the ad layout. No extra text, no watermarks, no unrelated logos.
by OpenAI
We curate copy-ready prompt libraries — every entry shows its full text and a sample result, ready to adapt.
How to get access, regional availability, and how this model compares with its alternatives.