image
generate, edit and upscale images. 94 apps on inference shell.
run via API, SDK, or belt CLI.

Gemini Nano Banana 2.1
google/gemini-nano-banana-2-1
gemini nano banana 2.1 via vertex ai — google's newest image model. better design, mask editing and subject consistency than nano banana 2, and faster. up to 4k, web search grounding.

grok-imagine-image-2-0
xai/grok-imagine-image-2-0
generate and edit images with xai's grok imagine image 2.0. text-to-image or editing with up to 5 source images, low or medium quality, 1k or 2k output.

GPT Image 2.5 Sunburst
openai/gpt-image-2-5-sunburst
gpt image 2.5 sunburst — openai's most capable image model, built for premium creative and editing workflows with tighter control across edits. text-to-image, reference-image editing, mask inpainting, transparent backgrounds, quality up to max.

GPT Image 2.5 Flare
openai/gpt-image-2-5-flare
gpt image 2.5 flare — openai's fast, high-quality everyday image model. higher quality than gpt image 2 at 50% lower latency. text-to-image, reference-image editing, mask inpainting, transparent backgrounds, quality up to max.

Krea 2 Medium
krea/krea-2-medium
krea 2 medium — expressive illustrations, ~10s per generation, 1.5k native resolution

Krea 2 Large
krea/krea-2-large
krea 2 large — photorealistic generation, ~25s per generation, 2k native resolution

Krea 2 Medium Turbo
krea/krea-2-medium-turbo
krea 2 medium turbo — fastest k2 model, ~3s per generation, 1.5k native resolution

Runway Gen-4 Image Turbo
runway/gen-4-image-turbo
runway gen-4 image turbo — fast text-to-image generation with reference image support. 2 credits per image at any resolution.

Runway Gen-4 Image
runway/gen-4-image
runway gen-4 image — high-quality text-to-image generation with optional reference images and @tag syntax. 5 credits per 720p, 8 credits per 1080p.

P-Image Ideogram
pruna/p-image-ideogram
high-quality text-to-image generation with strong typography and prompt understanding, built with ideogram

Seedream 5.0 Pro
bytedance/seedream-5-pro
bytedance's flagship seedream 5.0 pro image model via byteplus ark api. precision creation and editing with pixel-level regional edits, intelligent layer understanding, complex infographic generation, multi-reference blending (up to 10 images), and native text rendering in 14 languages.

Gemini 3.1 Flash Lite Image
google/gemini-3-1-flash-lite-image
gemini 3.1 flash lite image (nanobanana 2 lite) via vertex ai — ultra-low latency image generation

MAI Image 2.5
microsoft/mai-image-2-5
mai image 2.5 — microsoft's photorealistic image generation and editing model with fine-grained pixel-level control.

Reve Remix
reve/remix
reve remix — create images from text and 1-6 reference images combined.

Reve Create
reve/create
reve create — generate images from text with best-in-class prompt adherence and text rendering.

Gemini 3 Pro Image
google/gemini-3-pro-image
gemini 3 pro image (nanobanana pro) via vertex ai - advanced image generation model powered by google cloud

Gemini 3.1 Flash Image
google/gemini-3-1-flash-image
gemini 3.1 flash image (nanobanana 2) via vertex ai - advanced image generation model powered by google cloud

Kling Image V2
klingai/image-v2
kling image v2 (kolors v2.0) - text-to-image with 2k resolution, multi-image reference, and restyle. restyle output matches input resolution.

Kling Image O1
klingai/image-o1
kling image o1 (kolors image-o1) - omni image generation with element control. text-to-image and image-to-image at 1k/2k. $0.028/image.

Kling Image 3O
klingai/image-3o
kling image 3o (kolors image-3o) - most capable image model with native 4k, series-image generation, and element control. $0.028/image (4k $0.056).

Kling Image V1
klingai/image-v1
kling image v1 (kolors v1.0) - basic text-to-image and image-to-image generation. cheapest option at $0.0035/image.

Kling Image V1.5
klingai/image-v1-5
kling image v1.5 (kolors v1.5) - text-to-image with subject and face reference for character consistency. generate images preserving a person's appearance.

Kling Image V2.1
klingai/image-v2-1
kling image v2.1 (kolors v2.1) - text-to-image and multi-image reference generation. combine multiple images for complex compositions.

Kling Image V3
klingai/image-v3
kling image v3 (kolors v3.0) - latest image generation model with 1k/2k resolution support. highest quality text-to-image.

Grok Imagine Quality
xai/grok-imagine-image-quality
generate and edit high-quality images using xai's grok imagine quality model. supports 1k and 2k output resolutions with text-to-image and image editing.

Bria Generate
bria/generate
generate images from text prompts using bria fibo

Bria Generate Lite
bria/generate-lite
fast image generation from text prompts using bria fibo lite

GPT Image 2
openai/gpt-image-2
generate and edit images using openai's gpt image 2 model. supports text-to-image, image editing with reference images, mask-based inpainting, and transparent backgrounds.

PATINA Text to Material
patina/text-to-material
generates seamlessly tiling pbr materials up to 8k from a text prompt (optional image-to-image and inpainting) via fal.ai patina.

Wan 2.7 Image Pro
alibaba/wan-2-7-image-pro
wan 2.7 image pro is alibaba's professional image generation model supporting text-to-image, image editing, and multi-reference generation with up to 4k high-definition output

Wan 2.7 Image
alibaba/wan-2-7-image
wan 2.7 image is alibaba's fast image generation model supporting text-to-image, image editing, and multi-reference image generation with up to 2k resolution

Phota Generate
phota/generate
generate images from text prompts with identity-preserved subjects via [[profile_id]] syntax

Pruna FLUX.2 Klein 4B
pruna/flux-2-klein-4b
lightweight 4b parameter model with excellent speed-to-quality ratio

Pruna Z-Image Turbo LoRA
pruna/z-image-turbo-lora
fast generation with lora support for unique styles and personalized outputs

Pruna Z-Image Turbo
pruna/z-image-turbo
ultra-fast turbo image generation with minimal latency

Pruna Qwen-Image Fast
pruna/qwen-image-fast
fast qwen-based image generation with creativity control

Pruna Qwen-Image
pruna/qwen-image
advanced text-to-image generation with optional lora weights and prompt enhancement

Pruna Wan Image Small
pruna/wan-image-small
fast, efficient text-to-image optimized for rapid prototyping and batch generation
![P-Image-LoRA FLUX.1 [dev]](https://cloud.inference.sh/app/files/t/65hmp52a/4nuafhbd.png)
P-Image-LoRA FLUX.1 [dev]
pruna/flux-dev-lora
text-to-image and image-to-image generation with custom lora weights from huggingface
![P-Image FLUX.1 [dev]](https://cloud.inference.sh/app/files/t/65hmp52a/m26ne4vk.png)
P-Image FLUX.1 [dev]
pruna/flux-dev
advanced text-to-image generation with multiple aspect ratios, speed optimizations, and high-quality outputs

P-Image-LoRA
pruna/p-image-lora
pruna's flagship fast text-to-image with custom lora style support

P-Image
pruna/p-image
pruna's flagship fast text-to-image with multiple aspect ratios and prompt enhancement

Seedream 5 Lite
bytedance/seedream-5-lite
generate high-quality 2k-3k images from text prompts with single or multi-image input. supports text-to-image, image-to-image, and multi-reference image blending using bytedance's seedream 5 lite model via byteplus ark api.

Qwen-Image-2.0 Pro
alibaba/qwen-image-2-pro
qwen-image-2.0 pro offers enhanced text rendering, fine-grained realism, photorealistic scenes, and stronger semantic adherence for professional image generation and editing

Qwen-Image-2.0
alibaba/qwen-image-2
qwen-image-2.0 is alibaba's multimodal image generation model that integrates image generation and editing with enhanced text-rendering, realistic textures, and photorealistic scenes

Grok Imagine Pro
xai/grok-imagine-image-pro
generate and edit images using xai's grok imagine pro model. supports text-to-image and image editing with multiple aspect ratios.

Grok Imagine
xai/grok-imagine-image
generate and edit images using xai's grok imagine model. supports text-to-image and image editing with multiple aspect ratios.
![FLUX.1 [dev] LoRA](https://cloud.inference.sh/app/files/t/65hmp52a/r975rghg.png)
FLUX.1 [dev] LoRA
falai/flux-dev-lora
text-to-image and image-to-image generation with flux.1 [dev] lora support. custom style adaptation and fine-tuned model variations from black forest labs.
![FLUX.2 [klein] LoRA](https://cloud.inference.sh/app/files/t/65hmp52a/wbkhy3ot.png)
FLUX.2 [klein] LoRA
falai/flux-2-klein-lora
text-to-image and image-to-image generation with flux.2 [klein] lora support. available in 4b and 9b parameter sizes. custom style adaptation and fine-tuned model variations from black forest labs.

Seedream 3.0 T2I
bytedance/seedream-3-0-t2i
generate cinematic quality images from text prompts with accurate text rendering using bytedance's seedream 3.0 t2i model via byteplus ark api.

Seedream 4.0
bytedance/seedream-4-0
generate high-quality 2k-4k images from text prompts with optional image-to-image generation using bytedance's seedream 4.0 model via byteplus ark api.

Seedream 4.5
bytedance/seedream-4-5
generate high-quality 2k-4k images from text prompts with optional image-to-image generation using bytedance's seedream 4.5 model via byteplus ark api.

Imagine Art 1.5 Pro Preview
falai/imagine-art-1-5-pro-preview
advanced text-to-image model creating ultra-high-fidelity 4k visuals with lifelike realism and refined aesthetics.

Gemini 2.5 Flash Image
google/gemini-2-5-flash-image
gemini 2.5 flash image (nanobanana) via vertex ai - advanced image generation model powered by google cloud

CogView4 6B
infsh/cogview4-6b
generates high-quality images from text, capable of producing detailed visuals up to 2048x2048 resolution.

ByteDance USO
infsh/bytedance-uso
a unified image editor that allows users to generate images by combining any subject with any style efficiently, preserving identity and consistency.

OmniZero
infsh/omni-zero
creates stylized portraits instantly without needing specific training data.

HiDream-I1 Fast
infsh/hidream-i1-fast
a rapid image generator that produces high-quality images in various styles very quickly.

Chroma 1 HD
infsh/chroma-1-hd
creates high-definition images from text descriptions, primarily designed as a robust base for developers to customize and build upon.
![FLUX.1 [dev]](https://cloud.inference.sh/app/files/t/65hmp52a/ra5s25du.png)
FLUX.1 [dev]
infsh/flux-1-dev
generates images from text descriptions for non-commercial purposes.

Reve Edit
reve/edit
reve edit — edit images with natural language instructions. top 3 on lmarena leaderboard.

Bria Expand
bria/expand
expand image canvas with ai-generated content matching the original scene

Bria Eraser
bria/erase
remove objects from images using mask-based inpainting while preserving quality

Bria Gen Fill
bria/gen-fill
generative fill — replace masked regions with ai-generated content guided by a text prompt

Bria Product Packshot
bria/product-packshot
generate professional 2000x2000 product packshot images

Bria Product Shadow
bria/product-shadow
add realistic shadows to product cutout images

Bria Image Edit
bria/edit
edit an image using natural language text instructions

Phota Edit
phota/edit
edit images with text prompts while preserving identity of known subjects

Phota Enhance
phota/enhance
automatically enhance photo quality — lighting, composition, color, and sharpness

FLUX.1 Fill
infsh/flux-1-fill
fills designated areas in existing images based on a descriptive text input.

Bria Replace Background
bria/replace-background
replace image background with ai-generated content from a text prompt or reference image

PATINA Image to Material
patina/image-to-material
predicts seamless high-resolution pbr material maps (basecolor, normal, roughness, metalness, height) from a single input image via fal.ai patina.

Pruna Qwen-Image Edit Plus
pruna/qwen-image-edit-plus
edit images using text instructions with multi-image support and pose transfer

P-Image-Edit LoRA
pruna/p-image-edit-lora
fast image editing with custom lora styles for unique transformations

P-Image-Edit
pruna/p-image-edit
fast image editing with text instructions and multi-image support

Image Resize & Crop
eval/resize
crop and resize images — center crop to square, face-aware crop, or custom resize

Image Degradation
eval/blur
degrade images with blur, noise, and compression for benchmark testing

Image Resize
infsh/image-resize
resize images by width, height, scale factor, or megapixel target

Inference Shell Stitch Images
infsh/stitch-images
combine multiple photos horizontally or vertically into a single image or collage.

Krea 2 Medium Turbo LoRA Training
krea/krea-2-medium-turbo-train
train custom lora styles for krea 2 medium turbo generation

Krea 2 Medium LoRA Training
krea/krea-2-medium-train
train custom lora styles for krea 2 medium generation

Krea 2 Large LoRA Training
krea/krea-2-large-train
train custom lora styles for krea 2 large generation

Bria Background Removal
bria/rmbg
remove the background from an image, producing a transparent cutout. the general-purpose background removal — for product-specific cutouts, use product-cutout instead. output can be passed to replace-background, blur-background, or any editing app.

Bria Product Cutout
bria/product-cutout
cut out product from image with transparent background

BiRefNet Background Removal
infsh/birefnet
removes backgrounds from images quickly and accurately.

Bria Increase Resolution
bria/increase-resolution
upscale images 2x or 4x (max 8192x8192) while preserving original content

P-Image-Upscale
pruna/p-image-upscale
ai-powered image upscaling up to 128 megapixels with detail and realism enhancement

Thera Image Upscaler
infsh/thera
upscales images to any size without blurriness or jagged edges, maintaining high detail through its unique neural heat field technology.
explore more on inference shell
image is one of the categories on the grid. discover hundreds of apps across image, video, audio, and more.
we use cookies
we use cookies to ensure you get the best experience on our website. for more information on how we use cookies, please see our cookie policy.
by clicking "accept", you agree to our use of cookies.
learn more.





