Skip to content
Tools / Generate Image
Generate Image icon

Generate Image

AI image generation, 20+ models

4.4(7 reviews)

2 free skills. Paid skills start at $0.005. The final price is shown before running. Failed paid calls do not charge.

What is verified

Catalog facts and aggregate ToolRouter calls. Success uses all recorded calls, including caller errors in the total, and is not a controlled benchmark.

6Maintained skills
7Integrated providers
$0.005Paid calls from
1,130Recorded calls
94.6%Successful calls
2026-09-11Usage updated
2026-09-09Tool updated

Generate Image is a full AI image studio with 20+ models in a single tool. Create from text prompts, edit existing images with natural language, compose new shots from multiple references, and upscale to 4K — all without switching between different apps.

Models range from fast drafts to photorealistic renders, with specialists for typography, editorial photography, product shots, and artistic styles. The default model handles most tasks well; call list_models when you need something specific. Models approximate fonts, so use a text-focused model for drafts, then typeset pixel-exact copy in a design tool. Checked 2026-08-26 against this manifest.

What you can do

  • text_to_image — generate images from text prompts with full control over size, style, aspect ratio, and model
  • edit_image — modify an existing image using natural language: change backgrounds, add objects, remove elements, apply style transfers
  • image_to_image — compose a new image from up to 4 reference photos, combining scenes, personas, products, and outfits
  • check_image — poll for a pending result when a model is still processing
  • upscale_image — increase resolution up to 10x from inside the same workflow
  • list_models — browse all available models with capabilities and supported parameters

Who it's for

Marketers producing campaign assets, product teams visualizing concepts, content creators building visual content, and developers adding image generation to AI workflows.

How to use it

  1. Use text_to_image with a descriptive prompt — the default model works well without setting a model parameter
  2. Use edit_image to modify an existing photo with a natural language instruction
  3. Use image_to_image to combine reference images from your library (scenes, personas, products, outfits) into one new composition
  4. Use list_models to pick a specialist model for typography, photorealism, or artistic styles
  5. Use check_image before delivery and upscale_image only after approving the composition

Getting started

Try text_to_image with a wireless headphone on marble under soft studio light. Representative manifest-schema shape: images, model, seed [redacted]. Visual accuracy is not guaranteed; timing varies.

Permissions and setup

  • fal.ai API Key (secret): Optional: use your own fal.ai key instead of the platform default Official setup
  • Prodia API Token (secret): Optional: use your own Prodia token instead of the platform default Official setup
  • Higgsfield API Key (secret): Optional: use your own Higgsfield key instead of the platform default Official setup
  • Photalabs API Key (secret): Optional: use your own Photalabs key instead of the platform default Official setup
  • Google AI API Key (secret): Optional: use your own Google AI key instead of the platform default Official setup
  • OpenRouter API Key (secret): Optional: use your own OpenRouter key instead of the platform default Official setup
  • ElevenLabs API Key (secret): Optional: use your own ElevenLabs key instead of the platform default Official setup
Text to ImagePricing: paid

Generate an image from a text prompt. 20+ models with different quality, speed, style, and cost tradeoffs. Returns image URL(s) and downloadable assets.

Returns: Image URL(s), downloadable asset(s) via asset system, model used, seed, dimensions, and request metadata
Edit ImagePricing: paid

Edit, transform, or inpaint an existing image using AI. Provide an image URL and a text prompt describing the edit. Supports style transfer, object removal/addition, background changes, and targeted edits with masks.

Returns: Edited image URL(s), downloadable asset(s), source image URL, model used, seed, and request metadata
Image to ImagePricing: paid

Compose a new image from up to 4 reference images plus a prompt. Combine scenes, personas, outfits, products into one shot. Pass scene_file_id, persona_file_id, product_file_ids[], outfit_file_ids[], and/or raw image_urls[]. Default model nano-banana-2.

Returns: Composed image URL(s), downloadable asset(s), reference URLs used, model, seed, and request metadata
Check Image StatusPricing: free

Check on a pending fal-backed image request and retrieve it if ready. Use after text_to_image or edit_image returns status: pending with a fal_request_id.

Returns: Completed image URL(s) if ready, or the current queued/running/failed status with retry guidance
Upscale ImagePricing: paid

AI-upscale an image up to 10x resolution. Enhances detail and sharpness while maintaining quality. Choose a target resolution (720p-4K) or a scale factor (1-10x).

Returns: Upscaled image URL, downloadable asset, upscale factor or target resolution, and request metadata
List ModelsPricing: free

List all available image generation models with capabilities, pricing, supported parameters, and descriptions. Filter by capability type. Models not in the registry can still be used by passing their model endpoint ID directly.

Returns: Array of model objects with key, capabilities, pricing, supported parameters, and descriptions
Loading reviews...

Loading activity...

v0.192026-09-09
  • Added OpenAI GPT Image 2.5 — "gpt-image-2.5-flare" (fast default, up to 50% faster than GPT Image 2) and "gpt-image-2.5-sunburst" (precision tier for campaign and product work). Both have edit variants ("-edit") that take up to 16 reference images plus an optional mask.
  • Quality now accepts xhigh and max (plus auto) for GPT Image models — available on edit_image and image_to_image as well as text_to_image.
  • Fixed edit_image and image_to_image with GPT Image edit models — reference images were not reaching the model.
  • Default look when no style reference is given is now a candid iPhone camera-roll photo instead of an editorial lifestyle shoot. Pass style_reference_prompt or a saved style to override.
v0.182026-08-27
  • Added request-time same-model provider switching and retryable failover through backend.order, backend.only, backend.ignore, and backend.allow_fallbacks.
v0.172026-08-20
  • Added Seedream 5.0 Pro and Lite for higher-quality image generation and editing. Nano Banana 2 remains the default.
v0.162026-04-28
  • Added typography guidance to model descriptions — ideogram-v3 and gpt-image-2 are now clearly recommended for text-heavy designs; gallery previews now show the generation prompt instead of a generic label.
v0.152026-04-22
  • Quality-tier billing now exact across every model — low/medium/high calls on gpt-image-2 bill at $0.01 / $0.04 / $0.15 instead of a flat rate. Applies universally: every provider now returns the true per-request cost.
v0.142026-04-21
  • Added OpenAI GPT Image 2 — excellent prompt adherence and typography. Model key "gpt-image-2" works for text-to-image, editing, and mask-based inpainting; add image_urls to trigger edit mode. Also available via OpenRouter as "openrouter-gpt-image-2".
v0.132026-04-14
  • Accept personas, scenes, products, and outfits from your file library
v0.122026-04-13
  • Added image_to_image skill — compose new images from up to 4 reference URLs or saved file IDs (scenes, personas, products, outfits) plus a prompt. Pass scene_file_id, persona_file_id, product_file_ids, outfit_file_ids, or raw image_urls and the gateway resolves and merges them automatically.
v0.112026-04-06
  • Style references now supported — pass a style reference when generating or editing images for consistent visual direction
v0.102026-04-04
  • Ideogram style values are now case-insensitive — pass "realistic" or "REALISTIC", both work (valid: REALISTIC, GENERAL, DESIGN, AUTO)
v0.092026-03-31
  • Added Google as a direct provider — 6 new models: Nano Banana, Nano Banana 2, Nano Banana Pro, Imagen 4 Fast, Imagen 4, Imagen 4 Ultra. Direct Google API access with token-based billing for Nano Banana and flat-rate for Imagen 4.
v0.082026-03-27
  • Added Photalabs as a direct provider — phota, phota-edit, and phota-enhance models now route via the Photalabs API instead of fal.ai
v0.072026-03-24
  • Default model changed to nano-banana-2
  • Removed hardcoded model names from descriptions — agents discover models dynamically via list_models
v0.062026-03-24
  • Added Higgsfield Soul model — realistic human-centric UGC photos (model key: "soul")
v0.052026-03-23
  • Added Prodia as an alternative provider — 10 new models (prodia-flux-2-pro, prodia-sdxl, prodia-recraft-v4, prodia-gemini-3-pro, etc.)
v0.042026-03-23
  • upscale_image now delegates to image-upscale tool via composition
v0.032026-03-22
  • Fuzzy model resolution — natural names like "nano banana 2" now resolve correctly
  • Broadened live model discovery to catch more fal.ai categories and tags
  • Added subtitle, expanded description, and agent instructions
v0.022026-03-20
  • Initial release
v0.012026-04-03
  • Expanded Nano Banana aspect-ratio support to include auto, ultrawide, and extreme ratios
  • fal-backed image generation now returns recoverable pending results with check_image instead of silently degrading on long queue waits
  • fal live-pricing lookup now falls back to static registry pricing to avoid pricing-endpoint outages breaking image generation

Copy these instructions to use Generate Image in Claude, ChatGPT, Copilot, and more.

Related Tools

Related Categories

Frequently Asked Questions

Can I create an image from a text prompt?

Yes. `text_to_image` turns a prompt into a new image and returns downloadable image assets.

Can I edit an image I already have?

Yes. `edit_image` can change backgrounds, add or remove objects, and make targeted visual edits.

Can I upscale an image for higher resolution?

Yes. `upscale_image` can enhance an image up to 10x resolution, including 4K-level output.

How do I know which model to use?

Use `list_models` to compare capabilities, speed, and style strengths before you generate.