Generate Image is a full AI image studio with 20+ models in a single tool. Create from text prompts, edit existing images with natural language, compose new shots from multiple references, and upscale to 4K — all without switching between different apps.
Models range from fast drafts to photorealistic renders, with specialists for typography, editorial photography, product shots, and artistic styles. The default model handles most tasks well; call list_models when you need something specific. Models approximate fonts, so use a text-focused model for drafts, then typeset pixel-exact copy in a design tool. Checked 2026-08-26 against this manifest.
What you can do
- text_to_image — generate images from text prompts with full control over size, style, aspect ratio, and model
- edit_image — modify an existing image using natural language: change backgrounds, add objects, remove elements, apply style transfers
- image_to_image — compose a new image from up to 4 reference photos, combining scenes, personas, products, and outfits
- check_image — poll for a pending result when a model is still processing
- upscale_image — increase resolution up to 10x from inside the same workflow
- list_models — browse all available models with capabilities and supported parameters
Who it's for
Marketers producing campaign assets, product teams visualizing concepts, content creators building visual content, and developers adding image generation to AI workflows.
How to use it
- Use text_to_image with a descriptive prompt — the default model works well without setting a model parameter
- Use edit_image to modify an existing photo with a natural language instruction
- Use image_to_image to combine reference images from your library (scenes, personas, products, outfits) into one new composition
- Use list_models to pick a specialist model for typography, photorealism, or artistic styles
- Use check_image before delivery and upscale_image only after approving the composition
Getting started
Try text_to_image with a wireless headphone on marble under soft studio light. Representative manifest-schema shape: images, model, seed [redacted]. Visual accuracy is not guaranteed; timing varies.